MODEL PROOF

Kimi K2.6

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$1.71 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

GLM 5.3 Flash · Δ score 17 points · 86% lower measured price · $0.237 / 1M tokens

Gemini 3.8 Flash is a further stored option with these losses: 236K → 66K max output

Independent LMArena score
1,454.9 ± 4.52 · 37,502 votes
Context window
262.14K
Maximum output
235.93K
Input modalities
text, image
Output modalities
text
Published input price
$0.950 / 1M tokens
Published output price
$4.00 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Every price dimension · USD per million tokens
Input$0.950/M
Output$4.00/M
Cached input$0.160/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.700/M from Crusoe.

22 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Baidubaidu/fp4
Endpoint terms
Cached input /M
$0.073
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
54.6% · already included in these rates
Uptime · last 30 minutes
100.00%
$0.431 $1.82 fp4 99.91%
Inceptroninceptron/int4
Endpoint terms
Cached input /M
$0.117
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.77%
$0.432 $2.38 int4 99.74%
Chuteschutes/int4
Endpoint terms
Cached input /M
$0.050
Context limit
262,144 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
98.43%
$0.500 $2.85 int4 98.27%
DigitalOceandigitalocean
Endpoint terms
Cached input /M
$0.114
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
93.53%
$0.570 $2.40 undeclared 99.51%
Decartdecart/fp4
Endpoint terms
Cached input /M
$0.099
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.587 $2.47 fp4 99.22%
StreamLakestreamlake/fp8
Endpoint terms
Cached input /M
$0.101
Context limit
256,000 tokens
Output limit
230,400 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
37% · already included in these rates
Uptime · last 30 minutes
99.15%
$0.599 $2.52 fp8 99.09%
CoreWeavecoreweave/fp4
Endpoint terms
Cached input /M
$0.150
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.99%
$0.650 $3.41 fp4 99.90%
Crusoecrusoe/bf16
Endpoint terms
Cached input /M
$0.350
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.40%
$0.700 $3.50 bf16 98.30%
DeepInfradeepinfra/fp4
Endpoint terms
Cached input /M
$0.150
Context limit
262,144 tokens
Output limit
16,384 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.89%
$0.750 $3.50 fp4 98.75%
Venicevenice/int4
Endpoint terms
Cached input /M
$0.160
Context limit
256,000 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.80%
$0.750 $3.50 int4 98.26%
Parasailparasail/int4
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.74%
$0.750 $3.50 int4 99.65%
SiliconFlowsiliconflow/fp8
Endpoint terms
Cached input /M
$0.140
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.770 $3.40 fp8 99.76%
Novitanovita
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.800 $3.40 undeclared 99.75%
GMICloudgmicloud/fp8
Endpoint terms
Cached input /M
$0.144
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
10% · already included in these rates
Uptime · last 30 minutes
Unknown
$0.855 $3.60 fp8 92.20%
AtlasCloudatlas-cloud/int4
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.950 $4.00 int4 98.11%
Moonshot AImoonshotai/int4
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.950 $4.00 int4 99.92%
BaseTenbaseten/fp4
Endpoint terms
Cached input /M
$0.160
Context limit
262,000 tokens
Output limit
235,800 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
96.65%
$0.950 $4.00 fp4 95.90%
BaseTenbaseten/fp4
Endpoint terms
Cached input /M
$0.160
Context limit
262,000 tokens
Output limit
235,800 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
$0.950 $4.00 fp4 —
Cloudflarecloudflare
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
$0.950 $4.00 undeclared 92.18%
Fireworksfireworks
Endpoint terms
Cached input /M
$0.160
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
$0.950 $4.00 undeclared 0.00%
Sail Researchsail-research/int4
Endpoint terms
Cached input /M
$0.200
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$1.00 $4.00 int4 99.99%
Phalaphala
Endpoint terms
Cached input /M
$0.370
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$1.09 $4.60 undeclared 99.35%

Independent scores

LMArena Elo1454.9default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Best rankings by task

  • agentichtmlslides #2of 9 1248
  • agenticslides(html) #2of 9 1252
  • agenticslides #4of 9 1187
  • agenticslides(python-pptx) #4of 9 1186
  • pptxslides #4of 8 1181
  • webapps #7of 40 1268
  • htmlslides #10of 23 1187
  • 3d #18of 100 1297

Capability

Context window262K
Max output236K
Input modestext, image
Tool useyes
Reasoningoptional
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-09-23
Quality dataverified

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

Questions this page answers

What does Kimi K2.6 cost?

$0.950/M in, $4.00/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Kimi K2.6 have an independent quality score?

Kimi K2.6 has an LMArena score in this catalogue.

What beats Kimi K2.6?

GLM 5.3 Flash has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Kimi K2.6?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest Kimi K2.6 endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.700/M from Crusoe.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Kimi K2.6 — Undominated.ai dominance verdict

Markdown
[![Kimi K2.6 — Undominated.ai dominance verdict](https://undominated.ai/badge/moonshotai__kimi-k2.6.svg)](https://undominated.ai/models/moonshotai__kimi-k2.6/)
HTML
<a href="https://undominated.ai/models/moonshotai__kimi-k2.6/"><img src="https://undominated.ai/badge/moonshotai__kimi-k2.6.svg" alt="Kimi K2.6 — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/moonshotai__kimi-k2.6.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask