GPT-5.4 Mini
OpenAI released 2026-03-17 proprietary
GPT-5.6 Luna is both better and cheaper.
It scores +11.4 higher and costs 73% less ($0.450/M against $1.69/M) on this workload — and it does everything this model does.
Gemini 3.7 Flash is cheaper still (56% less) but drops 128K → 66K max output.
The same model via batch is 50% cheaper ($0.844/M) — same weights, different latency.
Every price dimension
| Input | $0.750/M |
|---|---|
| Output | $4.50/M |
| Cached input | $0.075/M |
| Web search | $0.01/call |
| Batch discount | 50% |
Independent scores
| Intelligence | 40.9 |
|---|---|
| Coding | 56.1 |
| Agentic | 31.5 |
| LMArena Elo | 1412.1high |
Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.
Capability
| Context window | 400K |
|---|---|
| Max output | 128K |
| Input modes | file, image, text |
| Tool use | yes |
| Reasoning | optional |
| Knowledge cutoff | 2025-08-31 |
| Open weights | no |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | verified |
| Cross-checked | vendor page |
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...