Qwen3.5-122B-A10B

Qwen released 2026-02-25

$0.715 per million tokens, balanced

MiniMax M3 is both better and cheaper.

It scores +12.6 higher and costs 27% less ($0.525/M against $0.715/M) on this workload — and it does everything this model does.

DeepSeek V4 Flash 0731 is cheaper still (85% less) but drops no image, video input.

Every price dimension

Input$0.260/M
Output$2.08/M

Independent scores

Intelligence32.8
Coding45.7
Agentic21.3
LMArena Elo1417.9default

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window262K
Max output66K
Input modestext, image, video
Tool useyes
Reasoningoptional

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Markdown for LLMs