MODEL PROOF

Gemini 3.7 Flash

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$1.50 / 1M tokens Balanced · 3 tokens in per 1 out

Reasoning tokens are billed separately.

Gemini 3.8 Flash · Δ score 4.2 points · same measured price · $1.50 / 1M tokens

Independent LMArena score
1,490.5 ± 8.18 · 5,640 votes
Context window
1.05M
Maximum output
65.54K
Input modalities
text, image, video, file, audio
Output modalities
text
Published input price
$0.750 / 1M tokens
Published output price
$3.75 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Every price dimension · USD per million tokens
Input$0.750/M
Output$3.75/M
Cached input$0.075/M
Cache write$0.042/M
Cache write 1hUnknown/M
Reasoning$3.75/M
Web search$0.014/call
Batch discount50%

Reasoning tokens are billed separately at $3.75/M, on top of output. Its share of your bill depends on the reasoning tokens used.

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

6 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Googlegoogle-vertex/global/flex
Endpoint terms
Cached input /M
$0.037
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
100.00%
$0.375 $1.88 undeclared 92.98%
Google AI Studiogoogle-ai-studio/flex
Endpoint terms
Cached input /M
$0.037
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
100.00%
$0.375 $1.88 undeclared 99.96%
Google AI Studiogoogle-ai-studio
Endpoint terms
Cached input /M
$0.075
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
99.94%
$0.750 $3.75 undeclared 99.86%
Googlegoogle-vertex/global
Endpoint terms
Cached input /M
$0.075
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
99.59%
$0.750 $3.75 undeclared 99.26%
Googlegoogle-vertex/global/priority
Endpoint terms
Cached input /M
$0.135
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
100.00%
$1.35 $6.75 undeclared 99.76%
Google AI Studiogoogle-ai-studio/priority
Endpoint terms
Cached input /M
$0.135
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
50% · already included in these rates
Uptime · last 30 minutes
100.00%
$1.35 $6.75 undeclared 99.97%

Independent scores

LMArena Elo1490.5high

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Best rankings by task

  • mobileapps #5of 40 1261
  • agenticgamedev #5of 24 1235
  • androidnative #6of 35 1253
  • website #7of 114 1311
  • gamedev #7of 107 1344
  • dataviz #8of 106 1325
  • codecategories #9of 108 1317
  • 3d #11of 100 1335

Capability

Context window1M
Max output66K
Input modestext, image, video, file, audio
Tool useyes
Reasoningalways on
Open weightsno

Provenance

Price sourceopenrouter.ai
Fetched2026-09-23
Quality dataverified
Cross-checkedvendor page

Our take

editorial — not a measurement

Evaluate Flash for recurring multimodal work, using the current proof to judge measured capability. The retained pricing source describes promotional terms, so confirm the dated vendor schedule before committing to a long-term budget. Standard and batch delivery are distinct offers; do not substitute a batch rate into a synchronous workload.

Strengths

  • Audio, video, image, file and text input are listed
  • Cache and batch pricing can be compared separately

Weaknesses

  • A promotional rate is not evidence of a permanent future price
  • Cache storage terms need checking alongside token rates

Reach for it when

  • Multimodal production evaluations
  • Workflows that can compare standard and delayed delivery

Avoid it if

  • Your forecast assumes unchanged promotional terms
  • You have not established the required delivery mode

Sources: ai.google.dev

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Questions this page answers

What does Gemini 3.7 Flash cost?

$0.750/M in, $3.75/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Gemini 3.7 Flash have an independent quality score?

Gemini 3.7 Flash has an LMArena score in this catalogue.

What beats Gemini 3.7 Flash?

Gemini 3.8 Flash has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Gemini 3.7 Flash?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Gemini 3.7 Flash — Undominated.ai dominance verdict

Markdown
[![Gemini 3.7 Flash — Undominated.ai dominance verdict](https://undominated.ai/badge/google__gemini-3.7-flash.svg)](https://undominated.ai/models/google__gemini-3.7-flash/)
HTML
<a href="https://undominated.ai/models/google__gemini-3.7-flash/"><img src="https://undominated.ai/badge/google__gemini-3.7-flash.svg" alt="Gemini 3.7 Flash — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/google__gemini-3.7-flash.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask