---
title: "Google: Gemini 2.5 Flash — price, capability and what beats it · Undominated.ai"
canonical: https://undominated.ai/models/google__gemini-2.5-flash/
description: "Google: Gemini 2.5 Flash: $0.300/M in, $2.50/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements."
---

# Google: Gemini 2.5 Flash — price, capability and what beats it · Undominated.ai

> Google: Gemini 2.5 Flash: $0.300/M in, $2.50/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements.

[Check base rates for this workload](/check/google__gemini-2.5-flash/) [Cheaper alternatives](/alternatives/google__gemini-2.5-flash/) [Family](/families/gemini-2/) [Compare models](https://undominated.ai/compare/?models=google%2Fgemini-2.5-flash) Comparison context

Choose up to three standard models in Compare.

Requirements and prompt length apply in Compare. This page keeps its stated price basis and workload controls.

MODEL PROOF

# Gemini 2.5 Flash

Google · 17 Jun 2025 · proprietary

 Balanced Summarise Chat Code gen Agentic

BETTER VALUE OPTION

## A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

 **$0.850** / 1M tokens Balanced · 3 tokens in per 1 out

Reasoning tokens are billed separately.

[Gemini 3.5 Flash Lite](/models/google__gemini-3.5-flash-lite/) · Δ score 17.7 points · same measured price · $0.850 / 1M tokens

[MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/) is a further stored option with these losses: no file input

Lifecycle: retirement date · 20 Oct 2026

Evidence as of 06 Oct 2026

 [openrouter/models](https://openrouter.ai/api/v1/models) [Vendor cross-check](https://ai.google.dev/gemini-api/docs/pricing) [LMArena · cc-by-4.0](https://huggingface.co/datasets/lmarena-ai/leaderboard-dataset) Independent LMArena score 1,417 ± 2.4 · 125,012 votes
 Context window 1.05M
 Maximum output 65.54K
 Input modalities file, image, text, audio, video
 Output modalities text
 Published input price $0.300 / 1M tokens
 Published output price $2.50 / 1M tokens
 Pricing kind fixed

Inspect [complete billing conditions](#billing-details) and endpoint terms below.

Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.

Cheaper right now: $0.425/M at Google AI Studio — delivery tier (flex) It does not pass the like-for-like test, so it never sets a rank.

 Every price dimension · USD per million tokens

| Input | $0.300 /M |
| --- | --- |
| Output | $2.50 /M |
| Cached input | $0.030 /M |
| Cache write | $0.083 /M |
| Cache write 1h | Unknown /M |
| Reasoning | $2.50 /M |
| Web search | $0.014 /call |
| Batch discount | 50% |

**Reasoning tokens are billed separately** at $2.50/M, on top of output. Its share of your bill depends on the reasoning tokens used.

### Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

 7 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

| Seller | Input | Output | Precision | Uptime (1d) |
| --- | --- | --- | --- | --- |
| Seller Google AI Studio google-ai-studio/flex Endpoint terms Cached input /M $0.015 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $0.150 | Output $1.25 | Precision undeclared | Uptime (1d) 99.99% |
| Seller Google google-vertex/eu Endpoint terms Cached input /M $0.030 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 91.05% | Input $0.300 | Output $2.50 | Precision undeclared | Uptime (1d) 96.07% |
| Seller Google google-vertex/global Endpoint terms Cached input /M $0.030 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 97.06% | Input $0.300 | Output $2.50 | Precision undeclared | Uptime (1d) 98.82% |
| Seller Google AI Studio google-ai-studio Endpoint terms Cached input /M $0.030 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.89% | Input $0.300 | Output $2.50 | Precision undeclared | Uptime (1d) 99.87% |
| Seller Google google-vertex Endpoint terms Cached input /M $0.030 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 92.51% | Input $0.300 | Output $2.50 | Precision undeclared | Uptime (1d) 91.72% |
| Seller Google google-vertex/global/priority Endpoint terms Cached input /M $0.054 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.67% | Input $0.540 | Output $4.50 | Precision undeclared | Uptime (1d) 99.86% |
| Seller Google AI Studio google-ai-studio/priority Endpoint terms Cached input /M $0.054 Context limit 1,048,576 tokens Output limit 65,535 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes Unknown | Input $0.540 | Output $4.50 | Precision undeclared | Uptime (1d) — |

### Independent scores

| LMArena Elo | 1417 default |
| --- | --- |

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

#### Best rankings by task

 - Data visualisation #80 1142
- SVG #80 1019
- UI components #90 1098
- 3D #91 1079
- Code categories #92 1108
- Websites #93 1122
- Game development #96 1078

Per-task Elo and rank from [Design Arena](https://designarena.ai), via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

### Capability

| Context window | 1M |
| --- | --- |
| Max output | 65K |
| Input modes | file, image, text, audio, video |
| Tool use | yes |
| Reasoning | optional |
| Knowledge cutoff | 2025-01-31 |
| Open weights | no |

**Retirement scheduled for 2026-10-20.** Do not start new work on it.

### Provenance

| Price source | [openrouter.ai](https://openrouter.ai/api/v1/models) |
| --- | --- |
| Fetched | 2026-10-06 |
| Quality data | verified |
| Cross-checked | [vendor page](https://ai.google.dev/gemini-api/docs/pricing) |

## Our take

 editorial — not a measurement

For an existing Gemini deployment, measure whether migration improves completed-task quality and cost. The accepted record distinguishes audio input from text-token pricing, so an audio-heavy estimate must use the relevant billing dimension. Check current benchmark evidence rather than assuming a newer model is better on every task.

### Strengths

 - Audio, video, image, file and text input are listed
- The accepted pricing record distinguishes audio input

### Weaknesses

 - A text-only token estimate can misprice audio workloads
- Cache retention and delivery mode need explicit assumptions

### Reach for it when

 - Migration evaluations for existing Gemini applications
- Multimodal workloads with measured input composition

### Avoid it if

 - Your estimate treats every input modality as text
- Another model has demonstrated a better result on your tasks

Sources: [ai.google.dev](https://ai.google.dev/gemini-api/docs/pricing)

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

## Questions this page answers

 What does Gemini 2.5 Flash cost?

$0.300/M in, $2.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

 Does Gemini 2.5 Flash have an independent quality score?

Gemini 2.5 Flash has an LMArena score in this catalogue.

 What beats Gemini 2.5 Flash?

Gemini 3.5 Flash Lite has an equal-or-higher measured score and an equal-or-lower price.

 Does this page use Artificial Analysis scores for Gemini 2.5 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

MODEL MONUMENT

## Save or share this proof

 Download proof card · 1200 × 630 Preview and more sharing options

Google: Gemini 2.5 Flash. Balanced workload. Better value option available. Evidence as of 06 Oct 2026.

 Copy model link Download square card · 1080 × 1080 Copy landscape image Share landscape card

## Related

 - Both better and cheaper [Gemini 3.5 Flash Lite](/models/google__gemini-3.5-flash-lite/)
- Higher score and cheaper, with a named trade [MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/)
- Nearby on the capability axis [DeepSeek V3.1 Terminus](/models/deepseek__deepseek-v3.1-terminus/)
- Same vendor [Gemma 4 26B A4B](/models/google__gemma-4-26b-a4b-it/)
- Google models [Google](/providers/google/)
- Where this model appears [deprecations](/deprecations/)
- Nearby on the capability axis [Gemini 3.1 Flash Lite Preview](/models/google__gemini-3.1-flash-lite-preview/)
- Same vendor [Gemma 4 31B](/models/google__gemma-4-31b-it/)
- Where this model appears [traps](/traps/)
- Nearby on the capability axis [Qwen3.5-122B-A10B](/models/qwen__qwen3.5-122b-a10b/)
- Same vendor [Gemini 2.5 Pro](/models/google__gemini-2.5-pro/)
- Where this model appears [significance](/significance/)

## Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

 Markdown Copy
 [![Gemini 2.5 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/google__gemini-2.5-flash.svg)](https://undominated.ai/models/google__gemini-2.5-flash/)
 HTML Copy
 <a href="https://undominated.ai/models/google__gemini-2.5-flash/"><img src="https://undominated.ai/badge/google__gemini-2.5-flash.svg" alt="Gemini 2.5 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: [https://undominated.ai/badge/google__gemini-2.5-flash.svg](https://undominated.ai/badge/google__gemini-2.5-flash.svg) — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

## Continue your investigation

 - [Explore model families](/families/)
- [Compare a shortlist](/compare/)
- [Check a model](/check/)
- [Understand the evidence](/methodology/)
