MODEL PROOF

Gemini 2.5 Flash

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$0.850 / 1M tokens Balanced · 3 tokens in per 1 out

Reasoning tokens are billed separately.

Gemini 3.5 Flash Lite · Δ score 17.7 points · same measured price · $0.850 / 1M tokens

MiMo-V2.6-Pro is a further stored option with these losses: no file input

Lifecycle: retirement date · 20 Oct 2026

Independent LMArena score
1,417 ± 2.4 · 125,012 votes
Context window
1.05M
Maximum output
65.54K
Input modalities
file, image, text, audio, video
Output modalities
text
Published input price
$0.300 / 1M tokens
Published output price
$2.50 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.

Cheaper right now: $0.425/M at Google AI Studio — delivery tier (flex) It does not pass the like-for-like test, so it never sets a rank.

Every price dimension · USD per million tokens
Input$0.300/M
Output$2.50/M
Cached input$0.030/M
Cache write$0.083/M
Cache write 1hUnknown/M
Reasoning$2.50/M
Web search$0.014/call
Batch discount50%

Reasoning tokens are billed separately at $2.50/M, on top of output. Its share of your bill depends on the reasoning tokens used.

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

7 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Google AI Studiogoogle-ai-studio/flex
Endpoint terms
Cached input /M
$0.015
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.150 $1.25 undeclared 99.99%
Googlegoogle-vertex/eu
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
91.05%
$0.300 $2.50 undeclared 96.07%
Googlegoogle-vertex/global
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
97.06%
$0.300 $2.50 undeclared 98.82%
Google AI Studiogoogle-ai-studio
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.89%
$0.300 $2.50 undeclared 99.87%
Googlegoogle-vertex
Endpoint terms
Cached input /M
$0.030
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
92.51%
$0.300 $2.50 undeclared 91.72%
Googlegoogle-vertex/global/priority
Endpoint terms
Cached input /M
$0.054
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.67%
$0.540 $4.50 undeclared 99.86%
Google AI Studiogoogle-ai-studio/priority
Endpoint terms
Cached input /M
$0.054
Context limit
1,048,576 tokens
Output limit
65,535 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
Unknown
$0.540 $4.50 undeclared —

Independent scores

LMArena Elo1417default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Best rankings by task

  • Data visualisation #80 1142
  • SVG #80 1019
  • UI components #90 1098
  • 3D #91 1079
  • Code categories #92 1108
  • Websites #93 1122
  • Game development #96 1078

Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

Capability

Context window1M
Max output65K
Input modesfile, image, text, audio, video
Tool useyes
Reasoningoptional
Knowledge cutoff2025-01-31
Open weightsno

Retirement scheduled for 2026-10-20. Do not start new work on it.

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataverified
Cross-checkedvendor page

Our take

editorial — not a measurement

For an existing Gemini deployment, measure whether migration improves completed-task quality and cost. The accepted record distinguishes audio input from text-token pricing, so an audio-heavy estimate must use the relevant billing dimension. Check current benchmark evidence rather than assuming a newer model is better on every task.

Strengths

  • Audio, video, image, file and text input are listed
  • The accepted pricing record distinguishes audio input

Weaknesses

  • A text-only token estimate can misprice audio workloads
  • Cache retention and delivery mode need explicit assumptions

Reach for it when

  • Migration evaluations for existing Gemini applications
  • Multimodal workloads with measured input composition

Avoid it if

  • Your estimate treats every input modality as text
  • Another model has demonstrated a better result on your tasks

Sources: ai.google.dev

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Questions this page answers

What does Gemini 2.5 Flash cost?

$0.300/M in, $2.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Gemini 2.5 Flash have an independent quality score?

Gemini 2.5 Flash has an LMArena score in this catalogue.

What beats Gemini 2.5 Flash?

Gemini 3.5 Flash Lite has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Gemini 2.5 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Gemini 2.5 Flash — Undominated.ai dominance verdict

Markdown
[![Gemini 2.5 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/google__gemini-2.5-flash.svg)](https://undominated.ai/models/google__gemini-2.5-flash/)
HTML
<a href="https://undominated.ai/models/google__gemini-2.5-flash/"><img src="https://undominated.ai/badge/google__gemini-2.5-flash.svg" alt="Gemini 2.5 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/google__gemini-2.5-flash.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask