MODEL PROOF
Qwen3 235B A22B Thinking 2507
BETTER VALUE OPTION
A stored alternative has an equal-or-higher measured score and an equal-or-lower price.
Serving precision differs between offers.
MiMo-V2.6-Pro · Δ score 75.8 points · 28% lower measured price · $0.540 / 1M tokens
Gemma 4 31B is a further stored option with these losses: 118K → 16K max output
Lifecycle: retirement date · 09 Oct 2026
- Independent LMArena score
- 1,415.2 ± 6.55 · 8,774 votes
- Context window
- 131.07K
- Maximum output
- 117.96K
- Input modalities
- text
- Output modalities
- text
- Published input price
- $0.230 / 1M tokens
- Published output price
- $2.30 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from Alibaba: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Every price dimension · USD per million tokens
| Input | $0.230/M |
|---|---|
| Output | $2.30/M |
| Cached input | Unknown/M |
| Cache write | Unknown/M |
| Cache write 1h | Unknown/M |
| Reasoning | Unknown/M |
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.300/M from Novita.
3 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| Alibabaalibaba Endpoint terms
| $0.230 | $2.30 | undeclared | 99.74% |
| Novitanovita/fp8 Endpoint terms
| $0.300 | $3.00 | fp8 | 99.72% |
| Venicevenice/fp8 Endpoint terms
| $0.450 | $3.50 | fp8 | 99.94% |
Independent scores
| LMArena Elo | 1415.2default |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
- 3D #102 1005
- Websites #107 1059
- Code categories #107 1037
- Data visualisation #113 957
- UI components #113 952
- Game development #116 970
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
Capability
| Context window | 131K |
|---|---|
| Max output | 117K |
| Input modes | text |
| Tool use | yes |
| Reasoning | always on |
| Knowledge cutoff | 2025-06-30 |
| Open weights | Unknown |
Retirement scheduled for 2026-10-09. Do not start new work on it.
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Questions this page answers
What does Qwen3 235B A22B Thinking 2507 cost?
$0.230/M in, $2.30/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Qwen3 235B A22B Thinking 2507 have an independent quality score?
Qwen3 235B A22B Thinking 2507 has an LMArena score in this catalogue.
What beats Qwen3 235B A22B Thinking 2507?
MiMo-V2.6-Pro has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for Qwen3 235B A22B Thinking 2507?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest Qwen3 235B A22B Thinking 2507 endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.300/M from Novita.
MODEL MONUMENT