Qwen3 32B

Qwen released 2025-04-28

$0.130 per million tokens, balanced

DeepSeek V4 Flash 0731 is both better and cheaper.

It scores +40.4 higher and costs 19% less ($0.105/M against $0.130/M) on this workload — and it does everything this model does.

Every price dimension

Input$0.080/M
Output$0.280/M

Independent scores

Intelligence11.4
Coding15.3
Agentic1.8
LMArena Elo1340.1default

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window131K
Max output16K
Input modestext
Tool useyes
Reasoningoptional
Knowledge cutoff2025-03-31

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

Markdown for LLMs