MODEL PROOF

Llama 3.2 3B Instruct

BETTER VALUE OPTION

A stored alternative has an equal-or-higher measured score and an equal-or-lower price.

$0.120 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

DeepSeek V4 Flash 0423 · Δ score 322.6 points · 6% lower measured price · $0.113 / 1M tokens

Qwen3.5-Flash is a further stored option with these losses: 118K → 66K max output

Independent LMArena score
1,109.5 ± 7.55 · 7,936 votes
Context window
131.07K
Maximum output
117.96K
Input modalities
text
Output modalities
text
Published input price
$0.050 / 1M tokens
Published output price
$0.330 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from Parasail: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (bf16), at its standard rate. The maker’s own precision is not known.

Every price dimension · USD per million tokens
Input$0.050/M
Output$0.330/M
Cached inputUnknown/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.050/M from Parasail.

2 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Parasailparasail/bf16
Endpoint terms
Cached input /M
Unknown
Context limit
131,072 tokens
Output limit
117,964 tokens
Tools
Not listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.050 $0.330 bf16 99.99%
Cloudflarecloudflare
Endpoint terms
Cached input /M
Unknown
Context limit
80,000 tokens
Output limit
72,000 tokens
Tools
Not listed by endpoint
Reasoning
Not listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.051 $0.335 undeclared 98.11%

Independent scores

LMArena Elo1109.5default

LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.

Capability

Context window131K
Max output117K
Input modestext
Tool useno
Reasoningno
Knowledge cutoff2023-12-31
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataverified

People search Llama 3.2 3B Instruct under spaced, hyphenated, and OpenRouter-id forms. This is that API row, not the 1B instruct SKU.

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

Questions this page answers

What does Llama 3.2 3B Instruct cost?

$0.050/M in, $0.330/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Llama 3.2 3B Instruct have an independent quality score?

Llama 3.2 3B Instruct has an LMArena score in this catalogue.

What beats Llama 3.2 3B Instruct?

DeepSeek V4 Flash 0423 has an equal-or-higher measured score and an equal-or-lower price.

Does this page use Artificial Analysis scores for Llama 3.2 3B Instruct?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Is the cheapest Llama 3.2 3B Instruct endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.050/M from Parasail.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Llama 3.2 3B Instruct — Undominated.ai dominance verdict

Markdown
[![Llama 3.2 3B Instruct — Undominated.ai dominance verdict](https://undominated.ai/badge/meta-llama__llama-3.2-3b-instruct.svg)](https://undominated.ai/models/meta-llama__llama-3.2-3b-instruct/)
HTML
<a href="https://undominated.ai/models/meta-llama__llama-3.2-3b-instruct/"><img src="https://undominated.ai/badge/meta-llama__llama-3.2-3b-instruct.svg" alt="Llama 3.2 3B Instruct — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/meta-llama__llama-3.2-3b-instruct.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask