Granica opłacalności
Każdy model naniesiony według tego, ile kosztuje, i tego, jak jest dobry. Schodki to granica: żadnego z tych modeli nic nie bije na obu osiach naraz. Wszystko poniżej i na prawo to oferta ściśle gorsza — jest coś zarazem mądrzejszego i tańszego.
Changing the workload below moves the frontier less than you would expect: 91% of frontier models are on it under every workload (only Solar Pro 4 moves). What does change is the price — the same model swings up to 4× across these shapes. So the control is not about which models to consider. It is about what you would actually pay for them.
10 niezdominowane z 106 gorsze oferty: 96
Obciążenie
Mierzone na
- 1 Ling-3.0-flashinclusionAI 37.8 $0.032/M najtańszy szczebel
- 2 Solar Pro 4Upstage 41.6 $0.052/M +3.8 za 1.7× ceny
- 3 DeepSeek V4 Flash 0423DeepSeek 42.1 $0.061/M +0.5 za 1.2× ceny
- 4 DeepSeek V4 Flash 0731DeepSeek 51.8 $0.105/M +9.7 za 1.7× ceny
- 5 GPT-5.6 LunaOpenAI 52.3 $0.450/M +0.5 za 4.3× ceny
- 6 Gemini 3.7 FlashGoogle 56.0 $0.750/M +3.7 za 1.7× ceny
- 7 Muse Spark 1.2Meta 56.8 $2.00/M +0.8 za 2.7× ceny
- 8 GLM 5.3Z.ai 59.5 $2.15/M +2.7 za 1.1× ceny
- 9 Grok 4.6xAI 60.9 $3.00/M +1.4 za 1.4× ceny
- 10 Claude Opus 5Anthropic 63.1 $10.00/M +2.2 za 3.3× ceny