Gemini 3.1 Flash Lite Preview
Google released 2026-03-03
DeepSeek V4 Flash 0731 scores higher and costs less — but you would give something up.
+26.2 on the capability index and 81% cheaper ($0.105/M against $0.563/M). What you lose:
- no image, video, file, audio input
Capability scores do not measure context length, output ceiling or which inputs a model accepts, so a higher score does not mean a drop-in replacement.
Every price dimension
| Input | $0.250/M |
|---|---|
| Output | $1.50/M |
| Cached input | $0.025/M |
| Cache write | $0.083/M |
| Reasoning | $1.50/M |
| Web search | $0.014/call |
Reasoning tokens are billed separately at $1.50/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.
Independent scores
| Intelligence | 25.6 |
|---|---|
| Coding | 34.7 |
| Agentic | 6.5 |
Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.
Best rankings by task
- asciiart #23of 83 1194
Capability
| Context window | 1.0M |
|---|---|
| Max output | 66K |
| Input modes | text, image, video, file, audio |
| Tool use | yes |
| Reasoning | optional |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | verified |
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...