MODEL PROOF

Qwen3.5-35B-A3B

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$0.355 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

GLM 5.3 Flash · Δ score 75.2 points · 33% lower measured price · $0.238 / 1M tokens

DeepSeek V4.1 Flash is a further stored option with these losses: no video input

Independent LMArena score
1,394.4 ± 4.21 · 30,408 votes
Context window
262.14K
Maximum output
32.77K
Input modalities
text, image, video
Output modalities
text
Published input price
$0.140 / 1M tokens
Published output price
$1.00 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.

Cheaper right now: $0.248/M at Darkbloom — 4-bit It does not pass the like-for-like test, so it never sets a rank.

Every price dimension · USD per million tokens
Input$0.140/M
Output$1.00/M
Cached input$0.050/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.140/M from DeepInfra.

7 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Darkbloomdarkbloom/fp4
Endpoint terms
Cached input /M
$0.040
Context limit
262,144 tokens
Output limit
32,768 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.90%
$0.080 $0.750 fp4 99.81%
DeepInfradeepinfra/fp8
Endpoint terms
Cached input /M
$0.050
Context limit
262,144 tokens
Output limit
81,920 tokens
Tools
Not listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.140 $1.00 fp8 99.91%
Parasailparasail/fp8
Endpoint terms
Cached input /M
$0.050
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.150 $1.00 fp8 99.95%
Alibabaalibaba
Endpoint terms
Cached input /M
Unknown
Context limit
262,144 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
$0.163 $1.30 undeclared 89.37%
AtlasCloudatlas-cloud/fp8
Endpoint terms
Cached input /M
$0.225
Context limit
262,144 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.225 $1.80 fp8 99.64%
SiliconFlowsiliconflow/fp8
Endpoint terms
Cached input /M
$0.150
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Not listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.240 $1.80 fp8 96.71%
Venicevenice
Endpoint terms
Cached input /M
$0.156
Context limit
256,000 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.313 $1.25 undeclared 99.89%

Independent scores

LMArena Elo1394.4default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Capability

Context window262K
Max output32K
Input modestext, image, video
Tool useyes
Reasoningoptional
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataverified

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

Questions this page answers

What does Qwen3.5-35B-A3B cost?

$0.140/M in, $1.00/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Qwen3.5-35B-A3B have an independent quality score?

Qwen3.5-35B-A3B has an LMArena score in this catalogue.

What beats Qwen3.5-35B-A3B?

GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Qwen3.5-35B-A3B?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest Qwen3.5-35B-A3B endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.140/M from DeepInfra.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Qwen3.5-35B-A3B — Undominated.ai dominance verdict

Markdown
[![Qwen3.5-35B-A3B — Undominated.ai dominance verdict](https://undominated.ai/badge/qwen__qwen3.5-35b-a3b.svg)](https://undominated.ai/models/qwen__qwen3.5-35b-a3b/)
HTML
<a href="https://undominated.ai/models/qwen__qwen3.5-35b-a3b/"><img src="https://undominated.ai/badge/qwen__qwen3.5-35b-a3b.svg" alt="Qwen3.5-35B-A3B — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/qwen__qwen3.5-35b-a3b.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask