---
title: "Z.ai: GLM 5 — price, capability and what beats it · Undominated.ai"
canonical: https://undominated.ai/models/z-ai__glm-5/
description: "Z.ai: GLM 5: $0.700/M in, $2.24/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements."
---

# Z.ai: GLM 5 — price, capability and what beats it · Undominated.ai

> Z.ai: GLM 5: $0.700/M in, $2.24/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements.

[Check base rates for this workload](/check/z-ai__glm-5/) [Cheaper alternatives](/alternatives/z-ai__glm-5/) [Family](/families/glm/) [Compare models](https://undominated.ai/compare/?models=z-ai%2Fglm-5) Comparison context

Choose up to three standard models in Compare.

Requirements and prompt length apply in Compare. This page keeps its stated price basis and workload controls.

MODEL PROOF

# GLM 5

Z.ai · 11 Feb 2026

 Balanced Summarise Chat Code gen Agentic

BETTER VALUE OPTION

## A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

 **$1.09** / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

[MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/) · Δ score 44.7 points · 50% lower measured price · $0.540 / 1M tokens

Evidence as of 06 Oct 2026

 [openrouter/models](https://openrouter.ai/api/v1/models) [LMArena · cc-by-4.0](https://huggingface.co/datasets/lmarena-ai/leaderboard-dataset) Independent LMArena score 1,446.3 ± 4.23 · 29,174 votes
 Context window 204.8K
 Maximum output 128K
 Input modalities text
 Output modalities text
 Published input price $0.700 / 1M tokens
 Published output price $2.24 / 1M tokens
 Pricing kind fixed

Inspect [complete billing conditions](#billing-details) and endpoint terms below.

Price from Baidu: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.

Cheaper right now: $0.930/M at GMICloud — promotion 40% It does not pass the like-for-like test, so it never sets a rank.

 Every price dimension · USD per million tokens

| Input | $0.700 /M |
| --- | --- |
| Output | $2.24 /M |
| Cached input | $0.140 /M |
| Cache write | Unknown /M |
| Cache write 1h | Unknown /M |
| Reasoning | Unknown /M |

### Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.700/M from Baidu.

 8 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

| Seller | Input | Output | Precision | Uptime (1d) |
| --- | --- | --- | --- | --- |
| Seller StreamLake streamlake/fp8 Endpoint terms Cached input /M $0.120 Context limit 198,000 tokens Output limit 128,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 40% · already included in these rates Uptime · last 30 minutes 99.96% | Input $0.600 | Output $1.92 | Precision fp8 | Uptime (1d) 99.89% |
| Seller GMICloud gmicloud/fp8 Endpoint terms Cached input /M $0.120 Context limit 202,752 tokens Output limit 182,476 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount 40% · already included in these rates Uptime · last 30 minutes 99.47% | Input $0.600 | Output $1.92 | Precision fp8 | Uptime (1d) 99.41% |
| Seller Baidu baidu/fp8 Endpoint terms Cached input /M $0.140 Context limit 202,752 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.700 | Output $2.24 | Precision fp8 | Uptime (1d) 99.46% |
| Seller SiliconFlow siliconflow/fp8 Endpoint terms Cached input /M $0.200 Context limit 204,800 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 98.68% | Input $0.950 | Output $2.55 | Precision fp8 | Uptime (1d) 99.66% |
| Seller Venice venice/fp8 Endpoint terms Cached input /M $0.200 Context limit 198,000 tokens Output limit 32,000 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.19% | Input $1.00 | Output $3.20 | Precision fp8 | Uptime (1d) 98.79% |
| Seller Novita novita/fp8 Endpoint terms Cached input /M $0.200 Context limit 202,800 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $1.00 | Output $3.20 | Precision fp8 | Uptime (1d) 99.99% |
| Seller Z.AI z-ai/fp8 Endpoint terms Cached input /M $0.200 Context limit 202,752 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.92% | Input $1.00 | Output $3.20 | Precision fp8 | Uptime (1d) 99.96% |
| Seller Amazon Bedrock amazon-bedrock Endpoint terms Cached input /M Unknown Context limit 202,752 tokens Output limit 131,072 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $1.00 | Output $3.20 | Precision undeclared | Uptime (1d) 99.22% |

### Independent scores

| LMArena Elo | 1446.3 default |
| --- | --- |

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

#### Best rankings by task

 - Htmlslides #18 1147
- Godot games #21 1144
- Native Android apps #24 1161
- Full-stack apps #27 1121
- Mobile apps #32 1133
- Code categories #33 1252
- 3D #33 1237
- Game development #34 1247

Per-task Elo and rank from [Design Arena](https://designarena.ai), via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

### Capability

| Context window | 204K |
| --- | --- |
| Max output | 128K |
| Input modes | text |
| Tool use | yes |
| Reasoning | optional |
| Open weights | Unknown |

### Provenance

| Price source | [openrouter.ai](https://openrouter.ai/api/v1/models) |
| --- | --- |
| Fetched | 2026-10-06 |
| Quality data | verified |

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

## Questions this page answers

 What does GLM 5 cost?

$0.700/M in, $2.24/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

 Does GLM 5 have an independent quality score?

GLM 5 has an LMArena score in this catalogue.

 What beats GLM 5?

MiMo-V2.6-Pro has an equal-or-higher measured score and an equal-or-lower price.

 Does this page use Artificial Analysis scores for GLM 5?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

 Is the cheapest GLM 5 endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.700/M from Baidu.

MODEL MONUMENT

## Save or share this proof

 Download proof card · 1200 × 630 Preview and more sharing options

Z.ai: GLM 5. Balanced workload. Better value option available. Evidence as of 06 Oct 2026.

 Copy model link Download square card · 1080 × 1080 Copy landscape image Share landscape card

## Related

 - Both better and cheaper [MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/)
- Nearby on the capability axis [GPT-5.6 Terra](/models/openai__gpt-5.6-terra/)
- Same vendor [GLM 4.6](/models/z-ai__glm-4.6/)
- Z.ai models [Z.ai](/providers/z-ai/)
- Where this model appears [spreads](/spreads/)
- Nearby on the capability axis [GPT-6.1 Sol](/models/openai__gpt-6.1-sol/)
- Same vendor [GLM 5V Turbo](/models/z-ai__glm-5v-turbo/)
- Where this model appears [traps](/traps/)
- Nearby on the capability axis [Qwen3.6 Max Preview](/models/qwen__qwen3.6-max-preview/)
- Same vendor [GLM 4.7](/models/z-ai__glm-4.7/)
- Where this model appears [significance](/significance/)
- Where this model appears [cheapest-at](/cheapest-at/)

## Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

 Markdown Copy
 [![GLM 5 — Undominated.ai dominance verdict](https://undominated.ai/badge/z-ai__glm-5.svg)](https://undominated.ai/models/z-ai__glm-5/)
 HTML Copy
 <a href="https://undominated.ai/models/z-ai__glm-5/"><img src="https://undominated.ai/badge/z-ai__glm-5.svg" alt="GLM 5 — Undominated.ai dominance verdict" height="36"></a>

Direct file: [https://undominated.ai/badge/z-ai__glm-5.svg](https://undominated.ai/badge/z-ai__glm-5.svg) — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

## Continue your investigation

 - [Explore model families](/families/)
- [Compare a shortlist](/compare/)
- [Check a model](/check/)
- [Understand the evidence](/methodology/)
