MODEL PROOF
Qwen3 30B A3B Instruct 2507
BETTER VALUE OPTION
A stored alternative has an equal-or-higher measured score and an equal-or-lower price.
Serving precision differs between offers.
Gemma 4 26B A4B · Δ score 49.8 points · 11% lower measured price · $0.127 / 1M tokens
Qwen3.5-Flash is a further stored option with these losses: 236K → 66K max output
- Independent LMArena score
- 1,384.1 ± 4.89 · 23,174 votes
- Context window
- 262.14K
- Maximum output
- 235.93K
- Input modalities
- text
- Output modalities
- text
- Published input price
- $0.090 / 1M tokens
- Published output price
- $0.300 / 1M tokens
- Pricing kind
- fixed
Inspect complete billing conditions and endpoint terms below.
Price from SiliconFlow: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (fp8), at its standard rate. The maker’s own precision is not known.
Cheaper right now: $0.084/M at StreamLake — promotion 55% · precision not disclosed It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
| Input | $0.090/M |
|---|---|
| Output | $0.300/M |
| Cached input | Unknown/M |
| Cache write | Unknown/M |
| Cache write 1h | Unknown/M |
| Reasoning | Unknown/M |
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.090/M from SiliconFlow.
5 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
| Seller | Input | Output | Precision | Uptime (1d) |
|---|---|---|---|---|
| StreamLakestreamlake Endpoint terms
| $0.048 | $0.193 | undeclared | 99.65% |
| SiliconFlowsiliconflow/fp8 Endpoint terms
| $0.090 | $0.300 | fp8 | 98.37% |
| DekaLLMdekallm Endpoint terms
| $0.090 | $0.300 | undeclared | 98.84% |
| Nebiusnebius/fp8 Endpoint terms
| $0.100 | $0.300 | fp8 | 92.10% |
| Alibabaalibaba Endpoint terms
| $0.130 | $0.520 | undeclared | 99.75% |
Independent scores
| LMArena Elo | 1384.1default |
|---|
LMArena Elo under CC BY 4.0. Absence of another board is not a score of zero.
Capability
| Context window | 262K |
|---|---|
| Max output | 235K |
| Input modes | text |
| Tool use | yes |
| Reasoning | no |
| Knowledge cutoff | 2025-06-30 |
| Open weights | Unknown |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-10-06 |
| Quality data | verified |
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Questions this page answers
What does Qwen3 30B A3B Instruct 2507 cost?
$0.090/M in, $0.300/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Qwen3 30B A3B Instruct 2507 have an independent quality score?
Qwen3 30B A3B Instruct 2507 has an LMArena score in this catalogue.
What beats Qwen3 30B A3B Instruct 2507?
Gemma 4 26B A4B has an equal-or-higher measured score and an equal-or-lower price.
Does this page use Artificial Analysis scores for Qwen3 30B A3B Instruct 2507?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
Is the cheapest Qwen3 30B A3B Instruct 2507 endpoint the same product?
Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.090/M from SiliconFlow.
MODEL MONUMENT