---
title: "Google: Gemini 3.5 Flash — price, capability and what beats it · Undominated.ai"
canonical: https://undominated.ai/models/google__gemini-3.5-flash/
description: "Google: Gemini 3.5 Flash: $1.50/M in, $9.00/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements."
---

# Google: Gemini 3.5 Flash — price, capability and what beats it · Undominated.ai

> Google: Gemini 3.5 Flash: $1.50/M in, $9.00/M out. Independent capability scores, and alternatives compared under recorded workload and capability requirements.

[Check base rates for this workload](/check/google__gemini-3.5-flash/) [Cheaper alternatives](/alternatives/google__gemini-3.5-flash/) [Family](/families/gemini-3/) [Compare models](https://undominated.ai/compare/?models=google%2Fgemini-3.5-flash) Comparison context

Choose up to three standard models in Compare.

Requirements and prompt length apply in Compare. This page keeps its stated price basis and workload controls.

MODEL PROOF

# Gemini 3.5 Flash

Google · 19 May 2026 · proprietary

 Balanced Summarise Chat Code gen Agentic

BETTER VALUE OPTION

## A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

 **$3.38** / 1M tokens Balanced · 3 tokens in per 1 out

Reasoning tokens are billed separately.

[Gemini 3.8 Flash](/models/google__gemini-3.8-flash/) · Δ score 19.3 points · 11% lower measured price · $3.00 / 1M tokens

[MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/) is a further stored option with these losses: no file input

Evidence as of 06 Oct 2026

 [openrouter/models](https://openrouter.ai/api/v1/models) [Vendor cross-check](https://ai.google.dev/gemini-api/docs/pricing) [LMArena · cc-by-4.0](https://huggingface.co/datasets/lmarena-ai/leaderboard-dataset) Independent LMArena score 1,477.7 ± 4.13 · 47,187 votes
 Context window 1.05M
 Maximum output 65.54K
 Input modalities text, image, video, file, audio
 Output modalities text
 Published input price $1.50 / 1M tokens
 Published output price $9.00 / 1M tokens
 Pricing kind fixed

Inspect [complete billing conditions](#billing-details) and endpoint terms below.

Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.

Cheaper right now: $1.69/M at Google — delivery tier (flex) It does not pass the like-for-like test, so it never sets a rank.

 Every price dimension · USD per million tokens

| Input | $1.50 /M |
| --- | --- |
| Output | $9.00 /M |
| Cached input | $0.150 /M |
| Cache write | $0.083 /M |
| Cache write 1h | Unknown /M |
| Reasoning | $9.00 /M |
| Web search | $0.014 /call |
| Batch discount | 50% |

**Reasoning tokens are billed separately** at $9.00/M, on top of output. Its share of your bill depends on the reasoning tokens used.

### Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

 7 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

| Seller | Input | Output | Precision | Uptime (1d) |
| --- | --- | --- | --- | --- |
| Seller Google google-vertex/global/flex Endpoint terms Cached input /M $0.075 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes Unknown | Input $0.750 | Output $4.50 | Precision undeclared | Uptime (1d) 98.68% |
| Seller Google AI Studio google-ai-studio/flex Endpoint terms Cached input /M $0.075 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes Unknown | Input $0.750 | Output $4.50 | Precision undeclared | Uptime (1d) 99.96% |
| Seller Google google-vertex/global Endpoint terms Cached input /M $0.150 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.65% | Input $1.50 | Output $9.00 | Precision undeclared | Uptime (1d) 98.86% |
| Seller Google AI Studio google-ai-studio Endpoint terms Cached input /M $0.150 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 99.91% | Input $1.50 | Output $9.00 | Precision undeclared | Uptime (1d) 99.88% |
| Seller Google google-vertex/us Endpoint terms Cached input /M $0.165 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes Unknown | Input $1.65 | Output $9.90 | Precision undeclared | Uptime (1d) — |
| Seller Google google-vertex/global/priority Endpoint terms Cached input /M $0.270 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $2.70 | Output $16.20 | Precision undeclared | Uptime (1d) 99.86% |
| Seller Google AI Studio google-ai-studio/priority Endpoint terms Cached input /M $0.270 Context limit 1,048,576 tokens Output limit 65,536 tokens Tools Listed by endpoint Reasoning Listed by endpoint Promotional discount None reported Uptime · last 30 minutes 100.00% | Input $2.70 | Output $16.20 | Precision undeclared | Uptime (1d) 99.93% |

### Independent scores

| LMArena Elo | 1477.7 medium |
| --- | --- |

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

#### Best rankings by task

 - Agenticslides(python pptx) #2 1242
- Pptxslides #2 1244
- Agenticslides #3 1244
- Agentichtmlslides #6 1162
- Agenticslides(html) #6 1162
- Python pptxslides #9 1247
- Agenticgamedev #12 1180
- SVG #13 1254

Per-task Elo and rank from [Design Arena](https://designarena.ai), via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.

### Capability

| Context window | 1M |
| --- | --- |
| Max output | 65K |
| Input modes | text, image, video, file, audio |
| Tool use | yes |
| Reasoning | always on |
| Knowledge cutoff | 2025-01-01 |
| Open weights | no |

### Provenance

| Price source | [openrouter.ai](https://openrouter.ai/api/v1/models) |
| --- | --- |
| Fetched | 2026-10-06 |
| Quality data | verified |
| Cross-checked | [vendor page](https://ai.google.dev/gemini-api/docs/pricing) |

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

## Questions this page answers

 What does Gemini 3.5 Flash cost?

$1.50/M in, $9.00/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

 Does Gemini 3.5 Flash have an independent quality score?

Gemini 3.5 Flash has an LMArena score in this catalogue.

 What beats Gemini 3.5 Flash?

Gemini 3.8 Flash has an equal-or-higher measured score and an equal-or-lower price.

 Does this page use Artificial Analysis scores for Gemini 3.5 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

MODEL MONUMENT

## Save or share this proof

 Download proof card · 1200 × 630 Preview and more sharing options

Google: Gemini 3.5 Flash. Balanced workload. Better value option available. Evidence as of 06 Oct 2026.

 Copy model link Download square card · 1080 × 1080 Copy landscape image Share landscape card

## Related

 - Both better and cheaper [Gemini 3.8 Flash](/models/google__gemini-3.8-flash/)
- Higher score and cheaper, with a named trade [MiMo-V2.6-Pro](/models/xiaomi__mimo-v2.6-pro/)
- Nearby on the capability axis [Muse Spark 1.1](/models/meta__muse-spark-1.1/)
- Same vendor [Gemini 3.1 Pro Preview](/models/google__gemini-3.1-pro-preview/)
- Google models [Google](/providers/google/)
- Where this model appears [disagreement](/disagreement/)
- Nearby on the capability axis [Kimi K3](/models/moonshotai__kimi-k3/)
- Same vendor [Gemini 3.7 Flash](/models/google__gemini-3.7-flash/)
- Where this model appears [effort](/effort/)
- Nearby on the capability axis [Gemini 3.6 Flash](/models/google__gemini-3.6-flash/)
- Same vendor [Gemini 2.5 Pro](/models/google__gemini-2.5-pro/)
- Where this model appears [traps](/traps/)

## Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

 Markdown Copy
 [![Gemini 3.5 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/google__gemini-3.5-flash.svg)](https://undominated.ai/models/google__gemini-3.5-flash/)
 HTML Copy
 <a href="https://undominated.ai/models/google__gemini-3.5-flash/"><img src="https://undominated.ai/badge/google__gemini-3.5-flash.svg" alt="Gemini 3.5 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: [https://undominated.ai/badge/google__gemini-3.5-flash.svg](https://undominated.ai/badge/google__gemini-3.5-flash.svg) — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

## Continue your investigation

 - [Explore model families](/families/)
- [Compare a shortlist](/compare/)
- [Check a model](/check/)
- [Understand the evidence](/methodology/)
