Ling-2.6-flash

inclusionAI released 2026-04-21

$0.015 per million tokens, balanced

Nothing is both better and cheaper.

Under a balanced workload, no other model in the catalogue scores higher and costs less. This model is on the value frontier.

Every price dimension

Input$0.010/M
Output$0.030/M
Cached input$0.0020/M

Independent scores

Intelligence14.2
Coding25.3
Agentic2.3

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window262K
Max output33K
Input modestext
Tool useyes
Reasoningno

Retirement scheduled for 2026-08-24. This model is no longer available.

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

Markdown for LLMs