MODEL PROOF
Qwen2.5 72B Instruct
BETTER VALUE OPTION
A stored alternative has an equal-or-higher measured score and an equal-or-lower price.
Serving precision differs between offers.
GLM 5.3 Flash · Δ score 200.6 points · 36% lower measured price · $0.238 / 1M tokens
Gemma 3 4B is a further stored option with these losses: no tool use
- Independent LMArena score
- 1,269 ± 4.06 · 39,406 votes
- Context window
- 32.77K
- Maximum output
- 16.38K
- Input modalities
- text
- Output modalities
- text
- Published input price
- $0.360 / 1M tokens
- Published output price
- $0.400 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.
Every price dimension · USD per million tokens
| Input | $0.360/M |
|---|---|
| Output | $0.400/M |
| Cached input | Unknown/M |
| Cache write | Unknown/M |
| Cache write 1h | Unknown/M |
| Reasoning | Unknown/M |
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.380/M from Novita.
2 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| DeepInfradeepinfra/fp8 Endpoint terms
| $0.360 | $0.400 | fp8 | 98.80% |
| Novitanovita/bf16 Endpoint terms
| $0.380 | $0.400 | bf16 | 97.40% |
Independent scores
| LMArena Elo | 1269default |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Capability
| Context window | 32K |
|---|---|
| Max output | 16K |
| Input modes | text |
| Tool use | yes |
| Reasoning | no |
| Knowledge cutoff | 2024-06-30 |
| Open weights | Unknown |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
Questions this page answers
What does Qwen2.5 72B Instruct cost?
$0.360/M in, $0.400/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Qwen2.5 72B Instruct have an independent quality score?
Qwen2.5 72B Instruct has an LMArena score in this catalogue.
What beats Qwen2.5 72B Instruct?
GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for Qwen2.5 72B Instruct?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest Qwen2.5 72B Instruct endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.380/M from Novita.
MODEL MONUMENT