MODEL PROOF
Gemini 3.7 Flash
BETTER VALUE OPTION
A stored alternative has an equal-or-higher measured score and an equal-or-lower price.
Reasoning tokens are billed separately.
Gemini 3.8 Flash · Δ score 10 points · same measured price · $3.00 / 1M tokens
MiMo-V2.6-Pro is a further stored option with these losses: no file input
- Independent LMArena score
- 1,487 ± 5.15 · 21,539 votes
- Context window
- 1.05M
- Maximum output
- 65.54K
- Input modalities
- text, image, video, file, audio
- Output modalities
- text
- Published input price
- $1.50 / 1M tokens
- Published output price
- $7.50 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.750 in · $3.75 out per million tokens at Google, a promotion of 50%. The price on this page is Google’s standard rate: the ranking uses the standard rate, never a promotion.
Every price dimension · USD per million tokens
| Input | $1.50/M |
|---|---|
| Output | $7.50/M |
| Cached input | $0.150/M |
| Cache write | $0.083/M |
| Cache write 1h | Unknown/M |
| Reasoning | $7.50/M |
| Web search | $0.014/call |
Reasoning tokens are billed separately at $7.50/M, on top of output. Its share of your bill depends on the reasoning tokens used.
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
6 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| Googlegoogle-vertex/global/flex Endpoint terms
| $0.375 | $1.88 | undeclared | 99.06% |
| Google AI Studiogoogle-ai-studio/flex Endpoint terms
| $0.375 | $1.88 | undeclared | 99.94% |
| Google AI Studiogoogle-ai-studio Endpoint terms
| $0.750 | $3.75 | undeclared | 99.86% |
| Googlegoogle-vertex/global Endpoint terms
| $0.750 | $3.75 | undeclared | 99.11% |
| Googlegoogle-vertex/global/priority Endpoint terms
| $1.35 | $6.75 | undeclared | 99.62% |
| Google AI Studiogoogle-ai-studio/priority Endpoint terms
| $1.35 | $6.75 | undeclared | 99.99% |
Independent scores
| LMArena Elo | 1487high |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
- Agenticgamedev #5 1235
- Native Android apps #6 1255
- Godot games #6 1264
- Game development #7 1334
- ASCII art #7 1289
- Websites #8 1311
- Mobile apps #8 1227
- Data visualisation #9 1327
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
Capability
| Context window | 1M |
|---|---|
| Max output | 65K |
| Input modes | text, image, video, file, audio |
| Tool use | yes |
| Reasoning | always on |
| Open weights | no |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
| Cross-checked | vendor page |
Our take
editorial — not a measurementEvaluate Flash for recurring multimodal work, using the current proof to judge measured capability. The retained pricing source describes promotional terms, so confirm the dated vendor schedule before committing to a long-term budget. Standard and batch delivery are distinct offers; do not substitute a batch rate into a synchronous workload.
Strengths
- Audio, video, image, file and text input are listed
- Cache and batch pricing can be compared separately
Weaknesses
- A promotional rate is not evidence of a permanent future price
- Cache storage terms need checking alongside token rates
Reach for it when
- Multimodal production evaluations
- Workflows that can compare standard and delayed delivery
Avoid it if
- Your forecast assumes unchanged promotional terms
- You have not established the required delivery mode
Sources: ai.google.dev
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Questions this page answers
What does Gemini 3.7 Flash cost?
$1.50/M in, $7.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Gemini 3.7 Flash have an independent quality score?
Gemini 3.7 Flash has an LMArena score in this catalogue.
What beats Gemini 3.7 Flash?
Gemini 3.8 Flash has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for Gemini 3.7 Flash?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
MODEL MONUMENT