MODEL PROOF

GLM 5.3 Flash

FRONTIER

Nothing is both better and cheaper under this workload.

$0.237 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

Independent LMArena score
1,471.9 ± 6.53 · 10,038 votes
Context window
1.31M
Maximum output
943.72K
Input modalities
text, image, video
Output modalities
text
Published input price
$0.150 / 1M tokens
Published output price
$0.500 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Every price dimension · USD per million tokens
Input$0.150/M
Output$0.500/M
Cached input$0.050/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.090/M from GMICloud.

31 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
DeepInfradeepinfra/fp4
Endpoint terms
Cached input /M
$0.015
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
99.07%
$0.075 $0.250 fp4 99.26%
Morphmorph
Endpoint terms
Cached input /M
$0.018
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
12% · already included in these rates
Uptime · last 30 minutes
77.48%
$0.088 $0.308 undeclared 83.81%
Waferwafer
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.93%
$0.089 $0.350 undeclared 99.90%
GMICloudgmicloud/fp8
Endpoint terms
Cached input /M
$0.018
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
40% · already included in these rates
Uptime · last 30 minutes
99.26%
$0.090 $0.300 fp8 99.24%
InferenceNetinference-net/fp4
Endpoint terms
Cached input /M
$0.020
Context limit
1,048,576 tokens
Output limit
128,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.31%
$0.090 $0.280 fp4 98.67%
OpenInferenceopen-inference/fp4
Endpoint terms
Cached input /M
$0.025
Context limit
1,048,576 tokens
Output limit
102,400 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.98%
$0.100 $0.500 fp4 99.92%
Relacerelace
Endpoint terms
Cached input /M
$0.020
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.95%
$0.100 $0.360 undeclared 99.91%
Phalaphala/fp8
Endpoint terms
Cached input /M
$0.025
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
15% · already included in these rates
Uptime · last 30 minutes
98.99%
$0.128 $0.425 fp8 99.49%
Novitanovita/fp8
Endpoint terms
Cached input /M
$0.026
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
12% · already included in these rates
Uptime · last 30 minutes
99.80%
$0.132 $0.440 fp8 99.50%
StreamLakestreamlake/fp8
Endpoint terms
Cached input /M
$0.028
Context limit
1,024,000 tokens
Output limit
128,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
6% · already included in these rates
Uptime · last 30 minutes
99.85%
$0.141 $0.470 fp8 99.31%
Sail Researchsail-research/us
Endpoint terms
Cached input /M
$0.029
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.33%
$0.142 $0.475 fp8 99.76%
Io Netio-net/fp8
Endpoint terms
Cached input /M
$0.029
Context limit
262,144 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
5% · already included in these rates
Uptime · last 30 minutes
99.82%
$0.142 $0.475 fp8 99.53%
BaseTenbaseten/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.60%
$0.150 $0.500 fp8 96.66%
Inceptroninceptron/fp8
Endpoint terms
Cached input /M
$0.070
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.32%
$0.150 $0.500 fp8 98.56%
Near AInear-ai/fp8
Endpoint terms
Cached input /M
$0.035
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.84%
$0.150 $0.500 fp8 99.31%
AtlasCloudatlas-cloud/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.75%
$0.150 $0.500 fp8 99.46%
SiliconFlowsiliconflow/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
262,144 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.84%
$0.150 $0.500 fp8 99.63%
Rekareka/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.94%
$0.150 $0.500 fp8 99.06%
Parasailparasail/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.75%
$0.150 $0.500 fp8 98.87%
BaseTenbaseten/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.80%
$0.150 $0.500 fp8 98.34%
Z.AIz-ai/fp8
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.84%
$0.150 $0.500 fp8 99.78%
Crusoecrusoe/fp4
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
93.53%
$0.150 $0.500 fp4 92.21%
CoreWeavecoreweave/nvfp4
Endpoint terms
Cached input /M
$0.050
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.69%
$0.150 $0.500 nvfp4 99.65%
Fireworksfireworks
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.60%
$0.150 $0.500 undeclared 99.55%
Friendlifriendli
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.91%
$0.150 $0.500 undeclared 99.14%
DigitalOceandigitalocean
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.95%
$0.150 $0.500 undeclared 95.45%
Togethertogether
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,575 tokens
Output limit
943,717 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.83%
$0.150 $0.500 undeclared 99.74%
Venicevenice
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.59%
$0.150 $0.500 undeclared 98.97%
NextBitnextbit/fp8
Endpoint terms
Cached input /M
$0.033
Context limit
1,048,576 tokens
Output limit
128,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.22%
$0.165 $0.550 fp8 98.70%
Cloudflarecloudflare
Endpoint terms
Cached input /M
$0.030
Context limit
1,310,720 tokens
Output limit
1,179,648 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
89.40%
$0.300 $1.00 undeclared 94.48%
Modalmodal/fp8
Endpoint terms
Cached input /M
$0.090
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.72%
$0.450 $1.50 fp8 99.74%

Independent scores

LMArena Elo1471.9default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Best rankings by task

  • 3d #7of 100 1350
  • svg #8of 76 1308
  • uicomponent #9of 104 1333
  • asciiart #10of 61 1283
  • gamedev #15of 107 1306
  • codecategories #16of 108 1292
  • dataviz #21of 106 1277
  • website #22of 114 1279

Capability

Context window1.3M
Max output944K
Input modestext, image, video
Tool useyes
Reasoningalways on
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-09-23
Quality dataverified

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Questions this page answers

What does GLM 5.3 Flash cost?

$0.150/M in, $0.500/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does GLM 5.3 Flash have an independent quality score?

GLM 5.3 Flash has an LMArena score in this catalogue.

What beats GLM 5.3 Flash?

Nothing is both better and cheaper than GLM 5.3 Flash.

Does this page use Artificial Analysis scores for GLM 5.3 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest GLM 5.3 Flash endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.090/M from GMICloud.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

GLM 5.3 Flash — Undominated.ai dominance verdict

Markdown
[![GLM 5.3 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/z-ai__glm-5.3-flash.svg)](https://undominated.ai/models/z-ai__glm-5.3-flash/)
HTML
<a href="https://undominated.ai/models/z-ai__glm-5.3-flash/"><img src="https://undominated.ai/badge/z-ai__glm-5.3-flash.svg" alt="GLM 5.3 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/z-ai__glm-5.3-flash.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask