Qwen3.5 397B A17B

Qwen released 2026-02-16

$0.877 per million tokens, balanced

Gemini 3.7 Flash is both better and cheaper.

It scores +21.7 higher and costs 15% less ($0.750/M against $0.877/M) on this workload — and it does everything this model does.

DeepSeek V4 Flash 0731 is cheaper still (88% less) but drops no image, video input.

Every price dimension

Input$0.390/M
Output$2.34/M

Independent scores

Intelligence34.3
Coding48.2
Agentic19.8
LMArena Elo1438.3default

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window262K
Max output66K
Input modestext, image, video
Tool useyes
Reasoningoptional

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

Markdown for LLMs