Claude Sonnet 4

Anthropic released 2025-05-22

$6.00 per million tokens, balanced

Gemini 3.7 Flash is both better and cheaper.

It scores +26.2 higher and costs 88% less ($0.750/M against $6.00/M) on this workload — and it does everything this model does.

GLM 5.3 is cheaper still (64% less) but drops no image, file input.

Every price dimension

Input$3.00/M
Output$15.00/M
Cached input$0.300/M
Cache write$3.75/M
Cache write 1h$6.00/M
Web search$0.01/call

Past 200,000 tokens the price changes. Input goes to $6.00/M (2×) and output to $22.50/M. The headline rate does not apply to a long-context workload.

Independent scores

Intelligence29.8
Coding37.6
Agentic17.6

Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.

Capability

Context window1M
Max output64K
Input modesimage, text, file
Tool useyes
Reasoningoptional
Knowledge cutoff2025-01-31

Provenance

Price sourceopenrouter.ai
Fetched2026-08-24
Quality dataverified

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

Markdown for LLMs