Check base rates for this workload Cheaper alternatives Family MODEL PROOF
gpt-oss-20b OpenAI · 05 Aug 2025
Balanced Summarise Chat Code gen Agentic
FRONTIER
Nothing is both better and cheaper under this workload. $0.036 / 1M tokens Balanced · 3 tokens in per 1 out
Serving precision differs between offers.
Independent LMArena score 1,287.3 ± 6.39 · 10,393 votes
Context window 131.07K
Maximum output 32.77K
Input modalities text
Output modalities text
Published input price $0.018 / 1M tokens
Published output price $0.090 / 1M tokens
Pricing kind fixed
Inspect complete billing conditions and endpoint terms below.
Every price dimension · USD per million tokens Input $0.018/M Output $0.090/M Cached input Unknown/M Cache write Unknown/M Cache write 1h Unknown/M Reasoning Unknown/M
Who sells it The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.029/M from DekaLLM.
12 provider offers · rates, limits and conditions Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
Seller Input Output Precision Uptime (1d) Seller Darkbloom
darkbloom/fp8 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 32,768 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.68% Input $0.018 Output $0.090 Precision fp8 Uptime (1d) 99.75% Seller AkashML
akashml/fp4 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.93% Input $0.020 Output $0.100 Precision fp4 Uptime (1d) 98.32% Seller DekaLLM
dekallm/bf16 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.60% Input $0.029 Output $0.140 Precision bf16 Uptime (1d) 99.72% Seller DeepInfra
deepinfra/bf16 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.98% Input $0.030 Output $0.140 Precision bf16 Uptime (1d) 99.96% Seller CoreWeave
coreweave/fp4 Endpoint terms
Cached input /M $0.030
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.99% Input $0.030 Output $0.130 Precision fp4 Uptime (1d) 100.00% Seller Parasail
parasail/fp4 Endpoint terms
Cached input /M $0.020
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 99.93% Input $0.030 Output $0.150 Precision fp4 Uptime (1d) 99.83% Seller SiliconFlow
siliconflow/fp8 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 8,192 tokens
Tools Not listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 93.31% Input $0.040 Output $0.180 Precision fp8 Uptime (1d) 96.92% Seller Novita
novita/fp4 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 32,768 tokens
Tools Not listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 100.00% Input $0.040 Output $0.150 Precision fp4 Uptime (1d) 99.90% Seller Amazon Bedrock
amazon-bedrock/eu-west-1 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes Unknown Input $0.070 Output $0.150 Precision undeclared Uptime (1d) 100.00% Seller Amazon Bedrock
amazon-bedrock Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 117,964 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 100.00% Input $0.070 Output $0.150 Precision undeclared Uptime (1d) 98.60% Seller Google
google-vertex/us-central1 Endpoint terms
Cached input /M Unknown
Context limit 131,072 tokens
Output limit 32,768 tokens
Tools Not listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 100.00% Input $0.070 Output $0.250 Precision undeclared Uptime (1d) 98.51% Seller Groq
groq Endpoint terms
Cached input /M $0.037
Context limit 131,072 tokens
Output limit 65,536 tokens
Tools Listed by endpoint
Reasoning Listed by endpoint
Promotional discount None reported
Uptime · last 30 minutes 98.63% Input $0.075 Output $0.300 Precision undeclared Uptime (1d) 99.45%
Independent scores LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Capability Context window 131K Max output 33K Input modes text Tool use yes Reasoning always on Knowledge cutoff 2024-06-30 Open weights Unknown
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Questions this page answers What does gpt-oss-20b cost? $0.018/M in, $0.090/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does gpt-oss-20b have an independent quality score? gpt-oss-20b has an LMArena score in this catalogue.
What beats gpt-oss-20b? Nothing is both better and cheaper than gpt-oss-20b.
Does this page use Artificial Analysis scores for gpt-oss-20b? No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest gpt-oss-20b endpoint the same product? Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.029/M from DekaLLM.
MODEL MONUMENT
Save or share this proof Download proof card · 1200 × 630 Preview and more sharing options OpenAI: gpt-oss-20b. Balanced workload. Frontier. Evidence as of 23 Sept 2026. OpenAI: gpt-oss-20b. Balanced workload. Frontier. Evidence as of 23 Sept 2026. gpt-oss-20b OpenAI SOURCE ID · openai/gpt-oss-20b BALANCED WORKLOAD · 3 tokens in per 1 out $0.036 USD · EFFECTIVE COST PER 1M WEIGHTED TOKENS 1,287.3 LMArena · cc-by-4.0 CONTEXT 131.07K · INPUT TEXT FRONTIER EVIDENCE AS OF 23 Sept 2026 · undominated.ai/models/openai__gpt-oss-20b/ OpenAI: gpt-oss-20b. Balanced workload. Frontier. Evidence as of 23 Sept 2026.
Copy model link Download square card · 1080 × 1080 Copy landscape image Share landscape card
Badge The verdict as one SVG, for a README or a docs page. It states this model's dominance
status at the balanced workload on the LMArena lens, with the date it was computed, and
it is rebuilt with the catalogue — so it changes when the verdict changes, including to
one you would rather it did not.
Markdown Copy
[](https://undominated.ai/models/openai__gpt-oss-20b/)HTML Copy
<a href="https://undominated.ai/models/openai__gpt-oss-20b/"><img src="https://undominated.ai/badge/openai__gpt-oss-20b.svg" alt="gpt-oss-20b — Undominated.ai dominance verdict" height="36"></a> Direct file: https://undominated.ai/badge/openai__gpt-oss-20b.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never
older than the last deploy.