MODEL PROOF

GLM 5.3 Flash

FRONTIER

Nothing is both better and cheaper under this workload.

$0.238 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

Independent LMArena score
1,469.6 ± 5.32 · 22,971 votes
Context window
1.05M
Maximum output
943.72K
Input modalities
text, image, video
Output modalities
text
Published input price
$0.150 / 1M tokens
Published output price
$0.500 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from Z.AI: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.

Cheaper right now: $0.119/M at DeepInfra — promotion 50% · 4-bit It does not pass the like-for-like test, so it never sets a rank.

Every price dimension · USD per million tokens
Input$0.150/M
Output$0.500/M
Cached input$0.030/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.150/M from BaseTen.

33 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Relacerelace
Endpoint terms
Cached input /M
$0.028
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.81%
$0.028 $0.500 undeclared 99.86%
OpenInferenceopen-inference/fp4
Endpoint terms
Cached input /M
$0.031
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.93%
$0.031 $0.688 fp4 99.25%
Sail Researchsail-research/us
Endpoint terms
Cached input /M
$0.029
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.44%
$0.045 $0.600 fp4 98.64%
Sail Researchsail-research/fp4
Endpoint terms
Cached input /M
$0.029
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.14%
$0.045 $0.600 fp4 98.47%
DeepInfradeepinfra/fp4
Endpoint terms
Cached input /M
$0.015
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
97.56%
$0.075 $0.250 fp4 99.20%
InferenceNetinference-net/fp4
Endpoint terms
Cached input /M
$0.025
Context limit
1,048,576 tokens
Output limit
262,144 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.21%
$0.080 $0.500 fp4 99.85%
Novitanovita/fp8
Endpoint terms
Cached input /M
$0.017
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
44% · already included in these rates
Uptime · last 30 minutes
98.81%
$0.084 $0.280 fp8 92.49%
StreamLakestreamlake/fp8
Endpoint terms
Cached input /M
$0.017
Context limit
1,024,000 tokens
Output limit
128,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
42% · already included in these rates
Uptime · last 30 minutes
99.20%
$0.087 $0.290 fp8 99.10%
Decartdecart/fp4
Endpoint terms
Cached input /M
$0.017
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
42% · already included in these rates
Uptime · last 30 minutes
99.94%
$0.087 $0.290 fp4 98.88%
GMICloudgmicloud/fp8
Endpoint terms
Cached input /M
$0.018
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
40% · already included in these rates
Uptime · last 30 minutes
98.56%
$0.090 $0.300 fp8 91.84%
Waferwafer
Endpoint terms
Cached input /M
$0.095
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.99%
$0.100 $0.500 undeclared 99.90%
DekaLLMdekallm
Endpoint terms
Cached input /M
$0.040
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.80%
$0.100 $1.00 undeclared 99.66%
Near AInear-ai/fp8
Endpoint terms
Cached input /M
$0.025
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
30% · already included in these rates
Uptime · last 30 minutes
99.56%
$0.105 $0.350 fp8 99.66%
Phalaphala/fp8
Endpoint terms
Cached input /M
$0.023
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
25% · already included in these rates
Uptime · last 30 minutes
99.04%
$0.113 $0.375 fp8 98.31%
BaseTenbaseten/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.150 $0.500 fp8 99.91%
AtlasCloudatlas-cloud/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.57%
$0.150 $0.500 fp8 80.48%
SiliconFlowsiliconflow/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
262,144 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.36%
$0.150 $0.500 fp8 98.94%
BaseTenbaseten/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.99%
$0.150 $0.500 fp8 99.96%
Z.AIz-ai/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.37%
$0.150 $0.500 fp8 98.90%
Crusoecrusoe/fp4
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.92%
$0.150 $0.500 fp4 99.67%
Parasailparasail/fp4
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.20%
$0.150 $0.500 fp4 99.68%
Modalmodal/nvfp4
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.07%
$0.150 $0.500 nvfp4 99.37%
CoreWeavecoreweave/nvfp4
Endpoint terms
Cached input /M
$0.050
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.89%
$0.150 $0.500 nvfp4 99.94%
Fireworksfireworks
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.91%
$0.150 $0.500 undeclared 96.10%
Friendlifriendli
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.90%
$0.150 $0.500 undeclared 99.15%
DigitalOceandigitalocean
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
97.97%
$0.150 $0.500 undeclared 99.67%
Togethertogether
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,575 tokens
Output limit
943,717 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.66%
$0.150 $0.500 undeclared 99.68%
Venicevenice
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
97.41%
$0.150 $0.500 undeclared 98.59%
Morphmorph/fp8
Endpoint terms
Cached input /M
$0.040
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.185 $0.646 fp8 96.93%
Inceptroninceptron/fp8
Endpoint terms
Cached input /M
$0.090
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.65%
$0.225 $0.600 fp8 97.94%
Fireworksfireworks/us
Endpoint terms
Cached input /M
$0.045
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.87%
$0.225 $0.750 undeclared 99.13%
Cloudflarecloudflare
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.71%
$0.300 $1.00 undeclared 99.19%
Rekareka
Endpoint terms
Cached input /M
$0.070
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
30% · already included in these rates
Uptime · last 30 minutes
98.17%
$0.350 $1.75 undeclared 99.68%

Independent scores

LMArena Elo1469.6default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Best rankings by task

  • SVG #10 1289
  • UI components #11 1324
  • 3D #11 1327
  • ASCII art #12 1276
  • Code categories #17 1288
  • Game development #17 1294
  • Websites #21 1280
  • Data visualisation #24 1272

Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

Capability

Context window1M
Max output943K
Input modestext, image, video
Tool useyes
Reasoningalways on
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataverified

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Questions this page answers

What does GLM 5.3 Flash cost?

$0.150/M in, $0.500/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does GLM 5.3 Flash have an independent quality score?

GLM 5.3 Flash has an LMArena score in this catalogue.

What beats GLM 5.3 Flash?

Nothing is both better and cheaper than GLM 5.3 Flash.

Does this page use Artificial Analysis scores for GLM 5.3 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest GLM 5.3 Flash endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.150/M from BaseTen.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

GLM 5.3 Flash — Undominated.ai dominance verdict

Markdown
[![GLM 5.3 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/z-ai__glm-5.3-flash.svg)](https://undominated.ai/models/z-ai__glm-5.3-flash/)
HTML
<a href="https://undominated.ai/models/z-ai__glm-5.3-flash/"><img src="https://undominated.ai/badge/z-ai__glm-5.3-flash.svg" alt="GLM 5.3 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/z-ai__glm-5.3-flash.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask