Gemini 2.5 Flash Lite

Google released 2025-07-22

$0.175 per million tokens, balanced

No independent quality score.

No benchmark we track has measured this model. That is not the same as measuring it and finding it wanting — we simply cannot rank it, so we do not.

Every price dimension

Input$0.100/M
Output$0.400/M
Cached input$0.010/M
Cache write$0.083/M
Reasoning$0.400/M
Web search$0.014/call

Reasoning tokens are billed separately at $0.400/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.

Independent scores

No independent benchmark has measured this model.

Capability

Context window1.0M
Max output66K
Input modestext, image, file, audio, video
Tool useyes
Reasoningoptional
Knowledge cutoff2025-01-31

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataunrated

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Markdown for LLMs