MODEL PROOF
Gemini 3.8 Flash
FRONTIER
Nothing is both better and cheaper under this workload.
Reasoning tokens are billed separately.
- Independent LMArena score
- 1,497 ± 5.04 · 26,298 votes
- Context window
- 1.05M
- Maximum output
- 65.54K
- Input modalities
- text, image, video, file, audio
- Output modalities
- text
- Published input price
- $1.50 / 1M tokens
- Published output price
- $7.50 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.750 in · $3.75 out per million tokens at Google, a promotion of 50%. The price on this page is Google’s standard rate: the ranking uses the standard rate, never a promotion.
Every price dimension · USD per million tokens
| Input | $1.50/M |
|---|---|
| Output | $7.50/M |
| Cached input | $0.150/M |
| Cache write | $0.083/M |
| Cache write 1h | Unknown/M |
| Reasoning | $7.50/M |
| Web search | $0.014/call |
Reasoning tokens are billed separately at $7.50/M, on top of output. Its share of your bill depends on the reasoning tokens used.
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
6 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| Google AI Studiogoogle-ai-studio/flex Endpoint terms
| $0.375 | $1.88 | undeclared | 99.92% |
| Googlegoogle-vertex/global/flex Endpoint terms
| $0.375 | $1.88 | undeclared | 99.30% |
| Google AI Studiogoogle-ai-studio Endpoint terms
| $0.750 | $3.75 | undeclared | 99.78% |
| Googlegoogle-vertex/global Endpoint terms
| $0.750 | $3.75 | undeclared | 97.52% |
| Google AI Studiogoogle-ai-studio/priority Endpoint terms
| $1.35 | $6.75 | undeclared | 99.69% |
| Googlegoogle-vertex/global/priority Endpoint terms
| $1.35 | $6.75 | undeclared | 99.65% |
Independent scores
| LMArena Elo | 1497high |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Best rankings by task
- Godot games #5 1273
- Mobile apps #6 1233
- SVG #8 1292
- Websites #9 1308
- ASCII art #10 1283
- Code categories #11 1309
- Game development #11 1319
- Full-stack apps #11 1242
Per-task Elo and rank from Design Arena, via OpenRouter’s model feed. The rank is the feed’s, and it counts models OpenRouter does not list.
Capability
| Context window | 1M |
|---|---|
| Max output | 65K |
| Input modes | text, image, video, file, audio |
| Tool use | yes |
| Reasoning | always on |
| Open weights | no |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Questions this page answers
What does Gemini 3.8 Flash cost?
$1.50/M in, $7.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Gemini 3.8 Flash have an independent quality score?
Gemini 3.8 Flash has an LMArena score in this catalogue.
What beats Gemini 3.8 Flash?
Nothing is both better and cheaper than Gemini 3.8 Flash.
Does this page use Artificial Analysis scores for Gemini 3.8 Flash?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
MODEL MONUMENT