Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.413/M at CoreWeave — 4-bit It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
Input
$0.280/M
Output
$1.10/M
Cached input
$0.056/M
Cache write
Unknown/M
Cache write 1h
Unknown/M
Reasoning
Unknown/M
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.280/M from DeepInfra.
13 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
Seller
Input
Output
Precision
Uptime (1d)
Seller
CoreWeavecoreweave/fp4 Endpoint terms
Cached input /M
$0.050
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.23%
Input$0.230
Output$0.960
Precisionfp4
Uptime (1d)98.47%
Seller
GMICloudgmicloud/fp8 Endpoint terms
Cached input /M
$0.048
Context limit
1,048,576 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
20% · already included in these rates
Uptime · last 30 minutes
98.79%
Input$0.240
Output$0.960
Precisionfp8
Uptime (1d)99.22%
Seller
DeepInfradeepinfra/fp8 Endpoint terms
Cached input /M
$0.056
Context limit
524,288 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.77%
Input$0.280
Output$1.10
Precisionfp8
Uptime (1d)99.86%
Seller
StreamLakestreamlake/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,000,000 tokens
Output limit
512,000 tokens
Tools
Not listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
99.43%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)98.69%
Seller
Venicevenice/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)97.21%
Seller
Parasailparasail/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,048,576 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.95%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.86%
Seller
AtlasCloudatlas-cloud/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,300 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.62%
Seller
Novitanovita/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,000,000 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.76%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.44%
Seller
Minimaxminimax/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
512,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.86%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.67%
Seller
Togethertogether Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
471,859 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.82%
Input$0.300
Output$1.20
Precisionundeclared
Uptime (1d)99.26%
Seller
Maramara Endpoint terms
Cached input /M
Unknown
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.00%
Input$0.600
Output$2.40
Precisionundeclared
Uptime (1d)93.42%
Seller
SambaNovasambanova Endpoint terms
Cached input /M
$0.060
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.18%
Input$0.600
Output$2.40
Precisionundeclared
Uptime (1d)96.17%
Seller
ModelRunmodelrun/fp4 Endpoint terms
Cached input /M
$0.150
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
97.42%
Input$0.750
Output$3.00
Precisionfp4
Uptime (1d)97.84%
Independent scores
LMArena Elo
1432.1default
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
Htmlslides#121180
Python pptxslides#141204
Full-stack apps#181196
Agenticgamedev#191158
Web apps#211192
Mobile apps#231167
Native Android apps#231171
Websites#311264
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
People search MiniMax. This page is MiniMax M3, not M2.7 and not a MiniMax subscription. This catalogue does not publish a parameter count. A missing count is not a guess. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Questions this page answers
What does MiniMax M3 cost?
$0.280/M in, $1.10/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does MiniMax M3 have an independent quality score?
MiniMax M3 has an LMArena score in this catalogue.
What beats MiniMax M3?
GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for MiniMax M3?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest MiniMax M3 endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.280/M from DeepInfra.
How many parameters does MiniMax M3 have?
This catalogue does not publish a parameter count for MiniMax M3. A missing count is not a guess.
MODEL MONUMENT
Save or share this proof
Preview and more sharing options
MiniMax: MiniMax M3. Balanced workload. Better value option available. Evidence as of 06 Oct 2026.
The verdict as one SVG, for a README or a docs page. It states this model's dominance
status at the balanced workload on the LMArena lens, with the date it was computed, and
it is rebuilt with the catalogue — so it changes when the verdict changes, including to
one you would rather it did not.