MODEL PROOF

Qwen2.5 72B Instruct

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$0.370 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

GLM 5.3 Flash · Δ score 200.6 points · 36% lower measured price · $0.238 / 1M tokens

Gemma 3 4B is a further stored option with these losses: no tool use

Independent LMArena score
1,269 ± 4.06 · 39,406 votes
Context window
32.77K
Maximum output
16.38K
Input modalities
text
Output modalities
text
Published input price
$0.360 / 1M tokens
Published output price
$0.400 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.

Every price dimension · USD per million tokens
Input$0.360/M
Output$0.400/M
Cached inputUnknown/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.380/M from Novita.

2 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
DeepInfradeepinfra/fp8
Endpoint terms
Cached input /M
Unknown
Context limit
32,768 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.63%
$0.360 $0.400 fp8 98.80%
Novitanovita/bf16
Endpoint terms
Cached input /M
Unknown
Context limit
32,000 tokens
Output limit
8,192 tokens
Tools
Listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
92.76%
$0.380 $0.400 bf16 97.40%

Independent scores

LMArena Elo1269default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Capability

Context window32K
Max output16K
Input modestext
Tool useyes
Reasoningno
Knowledge cutoff2024-06-30
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataverified

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

Questions this page answers

What does Qwen2.5 72B Instruct cost?

$0.360/M in, $0.400/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Qwen2.5 72B Instruct have an independent quality score?

Qwen2.5 72B Instruct has an LMArena score in this catalogue.

What beats Qwen2.5 72B Instruct?

GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Qwen2.5 72B Instruct?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest Qwen2.5 72B Instruct endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.380/M from Novita.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Qwen2.5 72B Instruct — Undominated.ai dominance verdict

Markdown
[![Qwen2.5 72B Instruct — Undominated.ai dominance verdict](https://undominated.ai/badge/qwen__qwen-2.5-72b-instruct.svg)](https://undominated.ai/models/qwen__qwen-2.5-72b-instruct/)
HTML
<a href="https://undominated.ai/models/qwen__qwen-2.5-72b-instruct/"><img src="https://undominated.ai/badge/qwen__qwen-2.5-72b-instruct.svg" alt="Qwen2.5 72B Instruct — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/qwen__qwen-2.5-72b-instruct.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask