MODEL PROOF
Gemini 2.5 Flash
BETTER VALUE OPTION
A stored alternative has an equal-or-higher measured score and an equal-or-lower price.
Reasoning tokens are billed separately.
Gemini 3.5 Flash Lite · Δ score 17.7 points · same measured price · $0.850 / 1M tokens
MiMo-V2.6-Pro is a further stored option with these losses: no file input
Lifecycle: retirement date · 20 Oct 2026
- Independent LMArena score
- 1,417 ± 2.4 · 125,012 votes
- Context window
- 1.05M
- Maximum output
- 65.54K
- Input modalities
- file, image, text, audio, video
- Output modalities
- text
- Published input price
- $0.300 / 1M tokens
- Published output price
- $2.50 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.425/M at Google AI Studio — delivery tier (flex) It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
| Input | $0.300/M |
|---|---|
| Output | $2.50/M |
| Cached input | $0.030/M |
| Cache write | $0.083/M |
| Cache write 1h | Unknown/M |
| Reasoning | $2.50/M |
| Web search | $0.014/call |
| Batch discount | 50% |
Reasoning tokens are billed separately at $2.50/M, on top of output. Its share of your bill depends on the reasoning tokens used.
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
7 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| Google AI Studiogoogle-ai-studio/flex Endpoint terms
| $0.150 | $1.25 | undeclared | 99.99% |
| Googlegoogle-vertex/eu Endpoint terms
| $0.300 | $2.50 | undeclared | 96.07% |
| Googlegoogle-vertex/global Endpoint terms
| $0.300 | $2.50 | undeclared | 98.82% |
| Google AI Studiogoogle-ai-studio Endpoint terms
| $0.300 | $2.50 | undeclared | 99.87% |
| Googlegoogle-vertex Endpoint terms
| $0.300 | $2.50 | undeclared | 91.72% |
| Googlegoogle-vertex/global/priority Endpoint terms
| $0.540 | $4.50 | undeclared | 99.86% |
| Google AI Studiogoogle-ai-studio/priority Endpoint terms
| $0.540 | $4.50 | undeclared | — |
Independent scores
| LMArena Elo | 1417default |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
- Data visualisation #80 1142
- SVG #80 1019
- UI components #90 1098
- 3D #91 1079
- Code categories #92 1108
- Websites #93 1122
- Game development #96 1078
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
Capability
| Context window | 1M |
|---|---|
| Max output | 65K |
| Input modes | file, image, text, audio, video |
| Tool use | yes |
| Reasoning | optional |
| Knowledge cutoff | 2025-01-31 |
| Open weights | no |
Retirement scheduled for 2026-10-20. Do not start new work on it.
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
| Cross-checked | vendor page |
Our take
editorial — not a measurementFor an existing Gemini deployment, measure whether migration improves completed-task quality and cost. The accepted record distinguishes audio input from text-token pricing, so an audio-heavy estimate must use the relevant billing dimension. Check current benchmark evidence rather than assuming a newer model is better on every task.
Strengths
- Audio, video, image, file and text input are listed
- The accepted pricing record distinguishes audio input
Weaknesses
- A text-only token estimate can misprice audio workloads
- Cache retention and delivery mode need explicit assumptions
Reach for it when
- Migration evaluations for existing Gemini applications
- Multimodal workloads with measured input composition
Avoid it if
- Your estimate treats every input modality as text
- Another model has demonstrated a better result on your tasks
Sources: ai.google.dev
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Questions this page answers
What does Gemini 2.5 Flash cost?
$0.300/M in, $2.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Gemini 2.5 Flash have an independent quality score?
Gemini 2.5 Flash has an LMArena score in this catalogue.
What beats Gemini 2.5 Flash?
Gemini 3.5 Flash Lite has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for Gemini 2.5 Flash?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
MODEL MONUMENT