Gemini 2.5 Pro

Google released 2025-06-17 proprietary

$3.44 per million tokens, balanced

Gemini 3.7 Flash is both better and cheaper.

It scores +30.1 higher and costs 78% less ($0.750/M against $3.44/M) on this workload — and it does everything this model does.

GLM 5.3 is cheaper still (37% less) but drops no image, file, audio, video input.

The same model via batch is 50% cheaper ($1.72/M) — same weights, different latency.

Every price dimension

Input$1.25/M
Output$10.00/M
Cached input$0.125/M
Cache write$0.375/M
Reasoning$10.00/M
Web search$0.014/call
Batch discount50%

Past 200,000 tokens the price changes. Input goes to $2.50/M (2×) and output to $15.00/M. The headline rate does not apply to a long-context workload.

Reasoning tokens are billed separately at $10.00/M, on top of output. On a reasoning-heavy workload this can be 40% of the bill and it does not appear in the advertised price.

Independent scores

Intelligence25.9
Coding33.3
Agentic7.2
LMArena Elo1457.3default

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window1.0M
Max output66K
Input modestext, image, file, audio, video
Tool useyes
Reasoningalways on
Knowledge cutoff2025-01-31
Open weightsno

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified
Cross-checkedvendor page

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Markdown for LLMs