Qwen3 235B A22B Thinking 2507

Qwen released 2025-07-25

$0.748 per million tokens, balanced

DeepSeek V4 Flash 0731 is both better and cheaper.

It scores +31.9 higher and costs 86% less ($0.105/M against $0.748/M) on this workload — and it does everything this model does.

MiniMax M2.7 is cheaper still (44% less) but drops 262K → 205K context.

Every price dimension

Input$0.230/M
Output$2.30/M

Independent scores

Intelligence19.9
Coding22.1
Agentic3.8
LMArena Elo1413.8default

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window262K
Input modestext
Tool useyes
Reasoningalways on
Knowledge cutoff2025-06-30

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Markdown for LLMs