MODEL PROOF

Qwen3 235B A22B Instruct 2507

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$0.153 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

Gemma 4 26B A4B · Δ score 14.7 points · 7% lower measured price · $0.142 / 1M tokens

Gemma 4 31B is a further stored option with these losses: 236K → 16K max output

Independent LMArena score
1,419.8 ± 2.57 · 95,670 votes
Context window
262.14K
Maximum output
235.93K
Input modalities
text
Output modalities
text
Published input price
$0.087 / 1M tokens
Published output price
$0.350 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Every price dimension · USD per million tokens
Input$0.087/M
Output$0.350/M
Cached input$0.018/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.087/M from GMICloud.

10 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
GMICloudgmicloud/fp8
Endpoint terms
Cached input /M
$0.018
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
75% · already included in these rates
Uptime · last 30 minutes
90.50%
$0.087 $0.350 fp8 98.69%
DeepInfradeepinfra/fp8
Endpoint terms
Cached input /M
Unknown
Context limit
262,144 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
44.14%
$0.090 $0.550 fp8 95.88%
Novitanovita/fp8
Endpoint terms
Cached input /M
Unknown
Context limit
131,072 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
77.47%
$0.090 $0.580 fp8 96.86%
Parasailparasail/fp8
Endpoint terms
Cached input /M
$0.050
Context limit
131,072 tokens
Output limit
117,964 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.95%
$0.140 $0.800 fp8 99.90%
Alibabaalibaba
Endpoint terms
Cached input /M
Unknown
Context limit
131,072 tokens
Output limit
32,768 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.99%
$0.149 $0.598 undeclared 99.92%
Venicevenice/fp8
Endpoint terms
Cached input /M
Unknown
Context limit
128,000 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
43.99%
$0.150 $0.750 fp8 96.11%
Nebiusnebius/fp8
Endpoint terms
Cached input /M
Unknown
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
81.01%
$0.200 $0.600 fp8 92.00%
StreamLakestreamlake
Endpoint terms
Cached input /M
Unknown
Context limit
128,000 tokens
Output limit
32,000 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
40% · already included in these rates
Uptime · last 30 minutes
90.03%
$0.210 $0.840 undeclared 98.25%
Googlegoogle-vertex/us-south1
Endpoint terms
Cached input /M
Unknown
Context limit
262,144 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.32%
$0.220 $0.880 undeclared 99.47%
Googlegoogle-vertex/us-south1
Endpoint terms
Cached input /M
Unknown
Context limit
262,144 tokens
Output limit
16,384 tokens
Tools
Not listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.40%
$0.250 $1.00 undeclared 99.64%

Independent scores

LMArena Elo1419.8default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Capability

Context window262K
Max output236K
Input modestext
Tool useyes
Reasoningno
Knowledge cutoff2025-06-30
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-09-23
Quality dataverified

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

Questions this page answers

What does Qwen3 235B A22B Instruct 2507 cost?

$0.087/M in, $0.350/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Qwen3 235B A22B Instruct 2507 have an independent quality score?

Qwen3 235B A22B Instruct 2507 has an LMArena score in this catalogue.

What beats Qwen3 235B A22B Instruct 2507?

Gemma 4 26B A4B has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Qwen3 235B A22B Instruct 2507?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest Qwen3 235B A22B Instruct 2507 endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.087/M from GMICloud.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Qwen3 235B A22B Instruct 2507 — Undominated.ai dominance verdict

Markdown
[![Qwen3 235B A22B Instruct 2507 — Undominated.ai dominance verdict](https://undominated.ai/badge/qwen__qwen3-235b-a22b-2507.svg)](https://undominated.ai/models/qwen__qwen3-235b-a22b-2507/)
HTML
<a href="https://undominated.ai/models/qwen__qwen3-235b-a22b-2507/"><img src="https://undominated.ai/badge/qwen__qwen3-235b-a22b-2507.svg" alt="Qwen3 235B A22B Instruct 2507 — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/qwen__qwen3-235b-a22b-2507.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask