Price from Parasail: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.
Cheaper right now: $0.261/M at StreamLake — promotion 88% It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
Input
$0.450/M
Output
$3.48/M
Cached input
$0.100/M
Cache write
Unknown/M
Cache write 1h
Unknown/M
Reasoning
Unknown/M
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.450/M from Parasail.
15 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
Seller
Input
Output
Precision
Uptime (1d)
Seller
Relacerelace/fp4 Endpoint terms
Cached input /M
$0.210
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
97.62%
Input$0.207
Output$4.20
Precisionfp4
Uptime (1d)99.85%
Seller
StreamLakestreamlake/fp8 Endpoint terms
Cached input /M
$0.017
Context limit
1,024,000 tokens
Output limit
384,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
88% · already included in these rates
Uptime · last 30 minutes
99.60%
Input$0.209
Output$0.418
Precisionfp8
Uptime (1d)98.92%
Seller
Parasailparasail/fp8 Endpoint terms
Cached input /M
$0.100
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.73%
Input$0.450
Output$3.48
Precisionfp8
Uptime (1d)98.64%
Seller
GMICloudgmicloud/fp8 Endpoint terms
Cached input /M
$0.080
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
45% · already included in these rates
Uptime · last 30 minutes
95.15%
Input$0.957
Output$1.91
Precisionfp8
Uptime (1d)97.03%
Seller
DigitalOceandigitalocean Endpoint terms
Cached input /M
$0.209
Context limit
1,048,576 tokens
Output limit
384,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.44%
Input$1.04
Output$2.09
Precisionundeclared
Uptime (1d)99.36%
Seller
Cloudflarecloudflare Endpoint terms
Cached input /M
$0.200
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
92.01%
Input$1.15
Output$2.55
Precisionundeclared
Uptime (1d)98.19%
Seller
Rekareka Endpoint terms
Cached input /M
$0.240
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.85%
Input$1.20
Output$12.00
Precisionundeclared
Uptime (1d)97.97%
Seller
DeepInfradeepinfra/fp8 Endpoint terms
Cached input /M
$0.100
Context limit
1,048,576 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.90%
Input$1.30
Output$2.60
Precisionfp8
Uptime (1d)99.82%
Seller
Alibabaalibaba/fp8 Endpoint terms
Cached input /M
$0.118
Context limit
1,000,000 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
95.99%
Input$1.42
Output$2.83
Precisionfp8
Uptime (1d)97.85%
Seller
SiliconFlowsiliconflow/fp8 Endpoint terms
Cached input /M
$0.135
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.96%
Input$1.50
Output$3.14
Precisionfp8
Uptime (1d)99.56%
Seller
Novitanovita/fp8 Endpoint terms
Cached input /M
$0.135
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.91%
Input$1.60
Output$3.20
Precisionfp8
Uptime (1d)99.97%
Seller
Venicevenice Endpoint terms
Cached input /M
$0.330
Context limit
1,000,000 tokens
Output limit
32,768 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$1.65
Output$3.30
Precisionundeclared
Uptime (1d)99.29%
Seller
AtlasCloudatlas-cloud/fp4 Endpoint terms
Cached input /M
$0.130
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.26%
Input$1.68
Output$3.38
Precisionfp4
Uptime (1d)98.99%
Seller
Baidubaidu/fp8 Endpoint terms
Cached input /M
$0.140
Context limit
1,048,576 tokens
Output limit
393,216 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.85%
Input$1.69
Output$3.38
Precisionfp8
Uptime (1d)99.91%
Seller
Azureazure/us Endpoint terms
Cached input /M
$0.160
Context limit
1,048,576 tokens
Output limit
384,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.83%
Input$1.91
Output$3.83
Precisionundeclared
Uptime (1d)99.41%
Independent scores
LMArena Elo
1445.2high
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
3D#241262
Godot games#321059
Game development#361242
Code categories#371246
ASCII art#401160
Web apps#411000
Websites#421241
Full-stack apps#43948
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
For scheduled or cache-heavy work, confirm the selected seller’s billing rules before estimating savings. A hosted listing does not establish another endpoint’s time-of-day tariff. The accepted record lists open weights, but deployment still needs a licence and hardware review. Use the current price table and proof for supported measurements, and confirm the tariff with the endpoint you intend to use.
Strengths
Open weights are recorded
Text input, tool use and cached-input pricing are listed
Weaknesses
Time-of-day terms must be established for the actual endpoint
A missing licence detail needs investigation before self-hosting
Reach for it when
Cache-heavy text workload evaluations
Self-hosting assessments with a verified licence
Avoid it if
Your savings calculation relies on an unverified schedule discount
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Questions this page answers
What does DeepSeek V4 Pro 0423 cost?
$0.450/M in, $3.48/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does DeepSeek V4 Pro 0423 have an independent quality score?
DeepSeek V4 Pro 0423 has an LMArena score in this catalogue.
What beats DeepSeek V4 Pro 0423?
GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for DeepSeek V4 Pro 0423?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest DeepSeek V4 Pro 0423 endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.450/M from Parasail.
MODEL MONUMENT
Save or share this proof
Preview and more sharing options
DeepSeek: DeepSeek V4 Pro 0423. Balanced workload. Better value option available. Evidence as of 06 Oct 2026.
The verdict as one SVG, for a README or a docs page. It states this model's dominance
status at the balanced workload on the LMArena lens, with the date it was computed, and
it is rebuilt with the catalogue — so it changes when the verdict changes, including to
one you would rather it did not.
Markdown
[](https://undominated.ai/models/deepseek__deepseek-v4-pro/)