Gemini 3.1 Flash Lite Preview

Google released 2026-03-03

$0.563 per million tokens, balanced

DeepSeek V4 Flash 0731 scores higher and costs less — but you would give something up.

+26.2 on the capability index and 81% cheaper ($0.105/M against $0.563/M). What you lose:

  • no image, video, file, audio input

Capability scores do not measure context length, output ceiling or which inputs a model accepts, so a higher score does not mean a drop-in replacement.

Every price dimension

Input$0.250/M
Output$1.50/M
Cached input$0.025/M
Cache write$0.083/M
Reasoning$1.50/M
Web search$0.014/call

Reasoning tokens are billed separately at $1.50/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.

Independent scores

Intelligence25.6
Coding34.7
Agentic6.5

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Best rankings by task

  • asciiart #23of 83 1194

Capability

Context window1.0M
Max output66K
Input modestext, image, video, file, audio
Tool useyes
Reasoningoptional

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Markdown for LLMs