Claude Sonnet 4
Anthropic released 2025-05-22
Gemini 3.7 Flash is both better and cheaper.
It scores +26.2 higher and costs 88% less ($0.750/M against $6.00/M) on this workload — and it does everything this model does.
GLM 5.3 is cheaper still (64% less) but drops no image, file input.
Every price dimension
| Input | $3.00/M |
|---|---|
| Output | $15.00/M |
| Cached input | $0.300/M |
| Cache write | $3.75/M |
| Cache write 1h | $6.00/M |
| Web search | $0.01/call |
Past 200,000 tokens the price changes. Input goes to $6.00/M (2×) and output to $22.50/M. The headline rate does not apply to a long-context workload.
Independent scores
| Intelligence | 29.8 |
|---|---|
| Coding | 37.6 |
| Agentic | 17.6 |
Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.
Capability
| Context window | 1M |
|---|---|
| Max output | 64K |
| Input modes | image, text, file |
| Tool use | yes |
| Reasoning | optional |
| Knowledge cutoff | 2025-01-31 |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | verified |
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...