Gemini 3.6 Flash

Google released 2026-07-21 proprietary

$1.50 per million tokens, balanced

Gemini 3.7 Flash is both better and cheaper.

It scores +4.4 higher and costs 50% less ($0.750/M against $1.50/M) on this workload — and it does everything this model does.

DeepSeek V4 Flash 0731 is cheaper still (93% less) but drops no image, video, file, audio input.

The same model via batch is 50% cheaper ($0.750/M) — same weights, different latency.

Every price dimension

Input$0.750/M
Output$3.75/M
Cached input$0.075/M
Cache write$0.042/M
Reasoning$3.75/M
Web search$0.014/call
Batch discount50%

Reasoning tokens are billed separately at $3.75/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.

Independent scores

Intelligence51.6
Coding69.2
Agentic40.5
LMArena Elo1476.5high

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Best rankings by task

  • website #5of 151 1323
  • agenticgamedev #7of 34 1199
  • dataviz #8of 142 1322
  • uicomponent #8of 140 1330
  • codecategories #9of 145 1319
  • asciiart #9of 83 1283
  • mobileapps #9of 56 1249
  • 3d #10of 136 1324

Capability

Context window1.0M
Max output66K
Input modestext, image, video, file, audio
Tool useyes
Reasoningalways on
Open weightsno

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified
Cross-checkedvendor page

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Markdown for LLMs