Gemini 2.5 Pro
Google released 2025-06-17 proprietary
Gemini 3.7 Flash is both better and cheaper.
It scores +30.1 higher and costs 78% less ($0.750/M against $3.44/M) on this workload — and it does everything this model does.
GLM 5.3 is cheaper still (37% less) but drops no image, file, audio, video input.
The same model via batch is 50% cheaper ($1.72/M) — same weights, different latency.
Every price dimension
| Input | $1.25/M |
|---|---|
| Output | $10.00/M |
| Cached input | $0.125/M |
| Cache write | $0.375/M |
| Reasoning | $10.00/M |
| Web search | $0.014/call |
| Batch discount | 50% |
Past 200,000 tokens the price changes. Input goes to $2.50/M (2×) and output to $15.00/M. The headline rate does not apply to a long-context workload.
Reasoning tokens are billed separately at $10.00/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.
Independent scores
| Intelligence | 25.9 |
|---|---|
| Coding | 33.3 |
| Agentic | 7.2 |
| LMArena Elo | 1457.3default |
Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.
Capability
| Context window | 1.0M |
|---|---|
| Max output | 66K |
| Input modes | text, image, file, audio, video |
| Tool use | yes |
| Reasoning | always on |
| Knowledge cutoff | 2025-01-31 |
| Open weights | no |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | verified |
| Cross-checked | vendor page |
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...