Is GPT-4 Turbo a good deal?

Whether anything in this catalogue beats GPT-4 Turbo on both quality and price, and what you give up if it does. A computation on the current catalogue, not an opinion.

12 undominated of 145 · Oct 6, 2026

As of Oct 6, 2026, GPT-4 Turbo is dominated for Balanced on LMArena. Claude Opus 5.5 scores 240.0 higher and costs 47% less, with a covering envelope. 12 of 145 rated, priced standard models are undominated.

Inspect model evidence Compare differences & requirements

Current model

GPT-4 Turbo

Balanced

3 tokens in per 1 out

Claude Opus 5.5 is both better and cheaper than GPT-4 Turbo: 240.0 points higher on LMArena and 47% less per million tokens, $7.00 cheaper at this mix.

GPT-4 Turbo

LMArena Elo · higher is better

Scale starts at 1040 Elo

1271.7

Effective $/M · Balanced · lower is better

$15.00/M

Claude Opus 5.5

LMArena Elo · higher is better

Scale starts at 1040 Elo

1511.7

Effective $/M · Balanced · lower is better

$8.00/M

GPT-4 Turbo takes text, image, returns up to 4,096 tokens from a 128,000-token context, and is listed by 1 seller.

Compared against 145 rated, priced models on this lens: 80 models dominate it and give up nothing, 5 more dominate it but give something up. GPT-4 Turbo scores 1271.7 at $15.00 per million tokens for this mix.

Envelope-safe replacements

Each row scores at least as high, costs no more, and covers this model’s context, output, modalities, tools, and reasoning. A cheaper narrower model is not listed here.

ModelLMArenaEffective $/MYou save
Claude Opus 5.51511.7 +240.0$8.00/M47%
Claude Opus 4.61503.2 +231.5$10.00/M33%
Claude Opus 51502.1 +230.4$10.00/M33%
Gemini 3.8 Flash1497.0 +225.3$3.00/M80%
MiMo-V2.6-Pro1491.0 +219.3$0.540/M96%
Muse Spark 1.31489.9 +218.2$2.00/M87%
Claude Opus 4.71489.9 +218.2$10.00/M33%
Gemini 3.7 Flash1487.0 +215.3$3.00/M80%
Muse Spark 1.21483.2 +211.5$2.00/M87%
Gemini 3.1 Pro Preview1480.2 +208.5$4.50/M70%
Gemini 3.6 Flash1479.5 +207.8$1.50/M90%
Muse Spark 1.11479.2 +207.5$2.00/M87%
Gemini 3.5 Flash1477.7 +206.0$3.38/M78%
Kimi K31475.9 +204.2$3.76/M75%
GLM 5.3 Flash1469.6 +197.9$0.238/M98%
GPT-5.51466.8 +195.1$11.25/M25%
Claude Sonnet 5.51466.6 +194.9$4.00/M73%
DeepSeek V4.1 Flash1462.5 +190.8$0.182/M99%
Claude Opus 4.81461.0 +189.3$10.00/M33%
Gemini 2.5 Pro1457.8 +186.1$3.44/M77%
Claude Sonnet 4.61457.8 +186.1$6.00/M60%
MiMo-V2.6-Flash1456.4 +184.7$0.175/M99%
GPT-5.6 Sol1456.3 +184.6$8.00/M47%
Kimi K2.61455.5 +183.8$0.961/M94%
Qwen3.7 Plus1454.6 +182.9$0.560/M96%
GPT-5.41452.2 +180.5$5.63/M63%
Claude Opus 4.51451.0 +179.3$10.00/M33%
Grok 4.51448.1 +176.4$3.00/M80%
GPT-5.6 Terra1446.3 +174.6$4.50/M70%
GPT-6.1 Sol1445.6 +173.9$4.00/M73%
Kimi K2.51445.2 +173.5$1.20/M92%
Gemma 4 31B1443.4 +171.7$0.205/M99%
Claude Sonnet 51442.9 +171.2$4.00/M73%
Inkling1441.8 +170.1$1.73/M89%
Qwen3.8 27B1440.7 +169.0$0.562/M96%
Claude Sonnet 4.51438.6 +166.9$6.00/M60%
Qwen3.5 397B A17B1438.0 +166.3$0.878/M94%
GLM 5V Turbo1437.2 +165.5$1.90/M87%
Qwen3.6 Plus1436.7 +165.0$0.731/M95%
Gemini 3.5 Flash Lite1434.7 +163.0$0.850/M94%
Gemma 4 26B A4B 1433.9 +162.2$0.127/M99%
MiniMax M31432.1 +160.4$0.485/M97%
GPT-5.6 Luna1431.0 +159.3$0.450/M97%
MiMo-V2.51427.5 +155.8$0.175/M99%
Grok 4.61427.4 +155.7$3.00/M80%
GPT-5.11422.2 +150.5$3.44/M77%
Mistral Medium 3.51421.2 +149.5$3.00/M80%
Qwen3 VL 235B A22B Instruct1420.1 +148.4$0.370/M98%
Qwen3.5-122B-A10B1417.1 +145.4$0.715/M95%
Gemini 2.5 Flash1417.0 +145.3$0.850/M94%
Gemini 3.1 Flash Lite Preview1415.7 +144.0$0.563/M96%
Inkling Small1413.7 +142.0$0.638/M96%
GPT-5.21412.3 +140.6$4.81/M68%
GPT-5.4 Mini1411.4 +139.7$1.69/M89%
o31409.9 +138.2$3.50/M77%
Qwen3.5-27B1408.9 +137.2$0.536/M96%
GPT-51406.1 +134.4$3.44/M77%
Qwen3 VL 235B A22B Thinking1400.7 +129.0$1.30/M91%
Grok 4.71399.9 +128.2$3.00/M80%
Qwen3.5-Flash1397.0 +125.3$0.114/M99%
Grok 4.31396.7 +125.0$1.56/M90%
Claude Haiku 4.51396.1 +124.4$2.00/M87%
GPT-6 Sol1395.4 +123.7$4.00/M73%
Qwen3.5-35B-A3B1394.4 +122.7$0.355/M98%
GPT-6 Luna1391.5 +119.8$0.200/M99%
GPT-4.11383.0 +111.3$3.50/M77%
GLM 4.6V1376.8 +105.1$0.450/M97%
GPT-5 Mini1373.2 +101.5$0.688/M95%
GPT-5.4 Nano1372.1 +100.4$0.463/M97%
Nova 2 Lite1361.8 +90.1$0.850/M94%
Gemma 3 27B1357.8 +86.1$0.100/M99%
o4 Mini1353.2 +81.5$1.93/M87%
GPT-4.1 Mini1340.4 +68.7$0.700/M95%
Claude Sonnet 41339.1 +67.4$6.00/M60%
Gemma 3 12B1334.2 +62.5$0.075/M100%
GPT-5 Nano1320.0 +48.3$0.138/M99%
GPT-4o (2024-05-13)1300.4 +28.7$7.50/M50%
GPT-4o-mini (2024-07-18)1286.4 +14.7$0.263/M98%
GPT-4.1 Nano1284.8 +13.1$0.175/M99%
GPT-4o (2024-08-06)1282.5 +10.8$4.38/M71%

Higher score, lower price, named losses

Not a drop-in. The loss is why these are not a recommendation.

ModelLMArenaEffective $/MYou give up
GLM 5.31471.4$1.00/Mimage
GLM 5.21470.4$1.12/Mimage
MiMo-V2.5-Pro1465.0$0.544/Mimage
GLM 5.11461.5$1.83/Mimage
Qwen3.6 Max Preview1446.8$2.31/Mimage
Claude Opus 5.51511.7 · $8.00/MGPT-4 Turbo1271.7 · $15.00/MLMArena EloEffective $/M · Balanced
145 rated, priced standard models. Chartreuse is the frontier. Cobalt is the model you named, when it is not on the staircase.

Open the frontier

The link keeps the model and mix, using the current catalogue. Monthly spend and switching cost stay in this browser tab and are left out of the link.

Saved decision references

Save the model, workload, catalogue date and capability-preservation rule in this browser. Spend, switching cost and bill contents are not saved. No account or notifications.

Constraint: replacements must preserve the model’s capabilities; any losses remain named trade-offs.

Historical decisions cannot be fully replayed from saved references: past prices, scores and capability evidence are not stored.

Evidence & Ask