---
title: "Google: Gemini 3.1 Flash Lite — price, capability and what beats it · Undominated.ai"
canonical: https://undominated.ai/models/google__gemini-3.1-flash-lite/
description: "Google: Gemini 3.1 Flash Lite: $0.250/M in, $1.50/M out. Independent capability scores, context-tier pricing, and the models that are both better and cheaper."
---

# Google: Gemini 3.1 Flash Lite — price, capability and what beats it · Undominated.ai

> Google: Gemini 3.1 Flash Lite: $0.250/M in, $1.50/M out. Independent capability scores, context-tier pricing, and the models that are both better and cheaper.

[Leaderboard](/) / Google

# Gemini 3.1 Flash Lite

Google · released 2026-05-07 · proprietary

 $0.563 per million tokens, balanced Balanced Summarise Chat Code gen Agentic

## No independent quality score.

No benchmark we track has measured this model. That is not the same as measuring it and finding it wanting — we simply cannot rank it, so we do not.

### Every price dimension

| Input | $0.250 /M |
| --- | --- |
| Output | $1.50 /M |
| Cached input | $0.025 /M |
| Cache write | $0.083 /M |
| Reasoning | $1.50 /M |
| Web search | $0.014 /call |
| Batch discount | 50% |

**Reasoning tokens are billed separately** at $1.50/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.

### Independent scores

| LMArena Elo | 1414.8 preview |
| --- | --- |

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

### Capability

| Context window | 1.0M |
| --- | --- |
| Max output | 66K |
| Input modes | text, image, video, file, audio |
| Tool use | yes |
| Reasoning | optional |
| Open weights | no |

### Provenance

| Price source | openrouter.ai |
| --- | --- |
| Fetched | 2026-08-24 |
| Quality data | unrated |
| Cross-checked | vendor page |

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
