The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.240/M from GMICloud.
13 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
Seller
Input
Output
Precision
Uptime (1d)
Seller
CoreWeavecoreweave/fp4 Endpoint terms
Cached input /M
$0.050
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.81%
Input$0.230
Output$0.960
Precisionfp4
Uptime (1d)99.18%
Seller
GMICloudgmicloud/fp8 Endpoint terms
Cached input /M
$0.048
Context limit
1,048,576 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
20% · already included in these rates
Uptime · last 30 minutes
99.88%
Input$0.240
Output$0.960
Precisionfp8
Uptime (1d)99.78%
Seller
DeepInfradeepinfra/fp8 Endpoint terms
Cached input /M
$0.056
Context limit
524,288 tokens
Output limit
512,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.36%
Input$0.280
Output$1.10
Precisionfp8
Uptime (1d)99.34%
Seller
StreamLakestreamlake/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,000,000 tokens
Output limit
512,000 tokens
Tools
Not listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
99.91%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.46%
Seller
Venicevenice/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
95.35%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)89.97%
Seller
Parasailparasail/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,048,576 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)96.24%
Seller
AtlasCloudatlas-cloud/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,300 tokens
Output limit
524,288 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)97.86%
Seller
Novitanovita/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
1,000,000 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.94%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.62%
Seller
Minimaxminimax/fp8 Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
512,000 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.85%
Input$0.300
Output$1.20
Precisionfp8
Uptime (1d)99.03%
Seller
Togethertogether Endpoint terms
Cached input /M
$0.060
Context limit
524,288 tokens
Output limit
471,859 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.99%
Input$0.300
Output$1.20
Precisionundeclared
Uptime (1d)99.89%
Seller
Maramara Endpoint terms
Cached input /M
Unknown
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
Input$0.600
Output$2.40
Precisionundeclared
Uptime (1d)90.61%
Seller
SambaNovasambanova Endpoint terms
Cached input /M
Unknown
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
64.69%
Input$0.600
Output$2.40
Precisionundeclared
Uptime (1d)95.78%
Seller
ModelRunmodelrun/fp4 Endpoint terms
Cached input /M
$0.150
Context limit
1,048,576 tokens
Output limit
943,718 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.750
Output$3.00
Precisionfp4
Uptime (1d)99.79%
Independent scores
LMArena Elo
1433.5default
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
People search MiniMax. This page is MiniMax M3, not M2.7 and not a MiniMax subscription. This catalogue does not publish a parameter count. A missing count is not a guess. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Questions this page answers
What does MiniMax M3 cost?
$0.300/M in, $1.20/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does MiniMax M3 have an independent quality score?
MiniMax M3 has an LMArena score in this catalogue.
What beats MiniMax M3?
GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for MiniMax M3?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest MiniMax M3 endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.240/M from GMICloud.
How many parameters does MiniMax M3 have?
This catalogue does not publish a parameter count for MiniMax M3. A missing count is not a guess.
MODEL MONUMENT
Save or share this proof
Preview and more sharing options
MiniMax: MiniMax M3. Balanced workload. Better value option available. Evidence as of 23 Sept 2026.
The verdict as one SVG, for a README or a docs page. It states this model's dominance
status at the balanced workload on the LMArena lens, with the date it was computed, and
it is rebuilt with the catalogue — so it changes when the verdict changes, including to
one you would rather it did not.