Gemini 2.5 Flash Lite
Google released 2025-07-22
No independent quality score.
No benchmark we track has measured this model. That is not the same as measuring it and finding it wanting — we simply cannot rank it, so we do not.
Every price dimension
| Input | $0.100/M |
|---|---|
| Output | $0.400/M |
| Cached input | $0.010/M |
| Cache write | $0.083/M |
| Reasoning | $0.400/M |
| Web search | $0.014/call |
Reasoning tokens are billed separately at $0.400/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.
Independent scores
No independent benchmark has measured this model.
Capability
| Context window | 1.0M |
|---|---|
| Max output | 66K |
| Input modes | text, image, file, audio, video |
| Tool use | yes |
| Reasoning | optional |
| Knowledge cutoff | 2025-01-31 |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | unrated |
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...