Ling-3.0-flash

inclusionAI released 2026-07-23

$0.032 per million tokens, balanced

Nothing is both better and cheaper.

Under a balanced workload, no other model in the catalogue scores higher and costs less. This model is on the value frontier.

Every price dimension

Input$0.021/M
Output$0.063/M
Cached input$0.0042/M

Independent scores

Intelligence37.8
Coding50.6
Agentic29.3

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window262K
Max output33K
Input modestext
Tool useyes
Reasoningoptional

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Markdown for LLMs