가치 프런티어

모든 모델을 가격과 성능 두 축에 찍은 그래프입니다. 계단선이 프런티어입니다. 두 축 모두에서 이 모델들을 동시에 이기는 모델은 없습니다. 그 아래쪽과 오른쪽에 있는 것은 전부 무조건 손해인 선택입니다. 더 똑똑하면서 더 싼 모델이 존재합니다.

Changing the workload below moves the frontier less than you would expect: 91% of frontier models are on it under every workload (only Solar Pro 4 moves). What does change is the price — the same model swings up to across these shapes. So the control is not about which models to consider. It is about what you would actually pay for them.

10 106개 중 비열위 손해인 선택 96개
작업 유형
측정 기준
$0.030$0.100$0.300$1.00$3.00$10.00102030405060100만 토큰당 실효 $ — 로그 스케일일반 성능
  1. 1 Ling-3.0-flashinclusionAI 37.8 $0.032/M 최저가 단
  2. 2 Solar Pro 4Upstage 41.6 $0.052/M 가격 1.7배에 +3.8
  3. 3 DeepSeek V4 Flash 0423DeepSeek 42.1 $0.061/M 가격 1.2배에 +0.5
  4. 4 DeepSeek V4 Flash 0731DeepSeek 51.8 $0.105/M 가격 1.7배에 +9.7
  5. 5 GPT-5.6 LunaOpenAI 52.3 $0.450/M 가격 4.3배에 +0.5
  6. 6 Gemini 3.7 FlashGoogle 56.0 $0.750/M 가격 1.7배에 +3.7
  7. 7 Muse Spark 1.2Meta 56.8 $2.00/M 가격 2.7배에 +0.8
  8. 8 GLM 5.3Z.ai 59.5 $2.15/M 가격 1.1배에 +2.7
  9. 9 Grok 4.6xAI 60.9 $3.00/M 가격 1.4배에 +1.4
  10. 10 Claude Opus 5Anthropic 63.1 $10.00/M 가격 3.3배에 +2.2
Markdown for LLMs