MODEL PROOF
GLM 5.3 Flash
FRONTIER
Nothing is both better and cheaper under this workload.
Serving precision differs between offers.
- Independent LMArena score
- 1,469.6 ± 5.32 · 22,971 votes
- Context window
- 1.05M
- Maximum output
- 943.72K
- Input modalities
- text, image, video
- Output modalities
- text
- Published input price
- $0.150 / 1M tokens
- Published output price
- $0.500 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from Z.AI: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.119/M at DeepInfra — promotion 50% · 4-bit It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
| Input | $0.150/M |
|---|---|
| Output | $0.500/M |
| Cached input | $0.030/M |
| Cache write | Unknown/M |
| Cache write 1h | Unknown/M |
| Reasoning | Unknown/M |
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.150/M from BaseTen.
33 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| Relacerelace Endpoint terms
| $0.028 | $0.500 | undeclared | 99.86% |
| OpenInferenceopen-inference/fp4 Endpoint terms
| $0.031 | $0.688 | fp4 | 99.25% |
| Sail Researchsail-research/us Endpoint terms
| $0.045 | $0.600 | fp4 | 98.64% |
| Sail Researchsail-research/fp4 Endpoint terms
| $0.045 | $0.600 | fp4 | 98.47% |
| DeepInfradeepinfra/fp4 Endpoint terms
| $0.075 | $0.250 | fp4 | 99.20% |
| InferenceNetinference-net/fp4 Endpoint terms
| $0.080 | $0.500 | fp4 | 99.85% |
| Novitanovita/fp8 Endpoint terms
| $0.084 | $0.280 | fp8 | 92.49% |
| StreamLakestreamlake/fp8 Endpoint terms
| $0.087 | $0.290 | fp8 | 99.10% |
| Decartdecart/fp4 Endpoint terms
| $0.087 | $0.290 | fp4 | 98.88% |
| GMICloudgmicloud/fp8 Endpoint terms
| $0.090 | $0.300 | fp8 | 91.84% |
| Waferwafer Endpoint terms
| $0.100 | $0.500 | undeclared | 99.90% |
| DekaLLMdekallm Endpoint terms
| $0.100 | $1.00 | undeclared | 99.66% |
| Near AInear-ai/fp8 Endpoint terms
| $0.105 | $0.350 | fp8 | 99.66% |
| Phalaphala/fp8 Endpoint terms
| $0.113 | $0.375 | fp8 | 98.31% |
| BaseTenbaseten/fp8 Endpoint terms
| $0.150 | $0.500 | fp8 | 99.91% |
| AtlasCloudatlas-cloud/fp8 Endpoint terms
| $0.150 | $0.500 | fp8 | 80.48% |
| SiliconFlowsiliconflow/fp8 Endpoint terms
| $0.150 | $0.500 | fp8 | 98.94% |
| BaseTenbaseten/fp8 Endpoint terms
| $0.150 | $0.500 | fp8 | 99.96% |
| Z.AIz-ai/fp8 Endpoint terms
| $0.150 | $0.500 | fp8 | 98.90% |
| Crusoecrusoe/fp4 Endpoint terms
| $0.150 | $0.500 | fp4 | 99.67% |
| Parasailparasail/fp4 Endpoint terms
| $0.150 | $0.500 | fp4 | 99.68% |
| Modalmodal/nvfp4 Endpoint terms
| $0.150 | $0.500 | nvfp4 | 99.37% |
| CoreWeavecoreweave/nvfp4 Endpoint terms
| $0.150 | $0.500 | nvfp4 | 99.94% |
| Fireworksfireworks Endpoint terms
| $0.150 | $0.500 | undeclared | 96.10% |
| Friendlifriendli Endpoint terms
| $0.150 | $0.500 | undeclared | 99.15% |
| DigitalOceandigitalocean Endpoint terms
| $0.150 | $0.500 | undeclared | 99.67% |
| Togethertogether Endpoint terms
| $0.150 | $0.500 | undeclared | 99.68% |
| Venicevenice Endpoint terms
| $0.150 | $0.500 | undeclared | 98.59% |
| Morphmorph/fp8 Endpoint terms
| $0.185 | $0.646 | fp8 | 96.93% |
| Inceptroninceptron/fp8 Endpoint terms
| $0.225 | $0.600 | fp8 | 97.94% |
| Fireworksfireworks/us Endpoint terms
| $0.225 | $0.750 | undeclared | 99.13% |
| Cloudflarecloudflare Endpoint terms
| $0.300 | $1.00 | undeclared | 99.19% |
| Rekareka Endpoint terms
| $0.350 | $1.75 | undeclared | 99.68% |
Independent scores
| LMArena Elo | 1469.6default |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
- SVG #10 1289
- UI components #11 1324
- 3D #11 1327
- ASCII art #12 1276
- Code categories #17 1288
- Game development #17 1294
- Websites #21 1280
- Data visualisation #24 1272
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
Capability
| Context window | 1M |
|---|---|
| Max output | 943K |
| Input modes | text, image, video |
| Tool use | yes |
| Reasoning | always on |
| Open weights | Unknown |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Questions this page answers
What does GLM 5.3 Flash cost?
$0.150/M in, $0.500/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does GLM 5.3 Flash have an independent quality score?
GLM 5.3 Flash has an LMArena score in this catalogue.
What beats GLM 5.3 Flash?
Nothing is both better and cheaper than GLM 5.3 Flash.
Does this page use Artificial Analysis scores for GLM 5.3 Flash?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest GLM 5.3 Flash endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.150/M from BaseTen.
MODEL MONUMENT