MODEL PROOF

Nemotron 3.5 Lightning

NOT INDEPENDENTLY RATED

No independent LMArena score is published for this model.

$0.085 / 1M tokens Balanced · 3 tokens in per 1 out

Serving precision differs between offers.

Independent LMArena score
Not independently rated
Context window
262.14K
Maximum output
32.77K
Input modalities
text
Output modalities
text
Published input price
$0.060 / 1M tokens
Published output price
$0.160 / 1M tokens
Pricing kind
fixed

Inspect complete billing conditions and endpoint terms below.

Price from DeepInfra: the cheapest offer that is serving, at standard delivery and at a declared precision that is not 4-bit (bf16), at its standard rate. The maker’s own precision is not known.

Cheaper right now: $0.072/M at Io Net — promotion 30% · precision not disclosed It does not pass the like-for-like test, so it never sets a rank.

Every price dimension · USD per million tokens
Input$0.060/M
Output$0.160/M
Cached input$0.030/M
Cache writeUnknown/M
Cache write 1hUnknown/M
ReasoningUnknown/M

Who sells it

The same weights, different shops. Cheapest is not like-for-like when serving precision differs.

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.060/M from DeepInfra.

5 provider offers · rates, limits and conditions

Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.

SellerInputOutputPrecisionUptime (1d)
Darkbloomdarkbloom/int4
Endpoint terms
Cached input /M
$0.020
Context limit
262,144 tokens
Output limit
32,768 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.84%
$0.039 $0.180 int4 99.34%
Io Netio-net
Endpoint terms
Cached input /M
$0.025
Context limit
262,144 tokens
Output limit
131,072 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
30% · already included in these rates
Uptime · last 30 minutes
99.20%
$0.049 $0.140 undeclared 99.66%
DeepInfradeepinfra/bf16
Endpoint terms
Cached input /M
$0.030
Context limit
262,144 tokens
Output limit
32,768 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.96%
$0.060 $0.160 bf16 99.98%
CoreWeavecoreweave/bf16
Endpoint terms
Cached input /M
$0.040
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
$0.070 $0.200 bf16 100.00%
Phalaphala
Endpoint terms
Cached input /M
$0.040
Context limit
262,144 tokens
Output limit
235,929 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
99.89%
$0.070 $0.200 undeclared 99.83%

Independent scores

No independent benchmark has measured this model. Unrated is not a score of zero.

Capability

Context window262K
Max output32K
Input modestext
Tool useyes
Reasoningoptional
Open weightsUnknown

Provenance

Price sourceopenrouter.ai
Fetched2026-10-06
Quality dataunrated

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Questions this page answers

What does Nemotron 3.5 Lightning cost?

$0.060/M in, $0.160/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.

Does Nemotron 3.5 Lightning have an independent quality score?

Nemotron 3.5 Lightning has no independent quality score in this catalogue. Unrated is not a score of zero.

What beats Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning has no independent quality score. Unrated is not a score of zero.

Does this page use Artificial Analysis scores for Nemotron 3.5 Lightning?

No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.

Does a missing score mean Nemotron 3.5 Lightning scored zero?

Unrated is not a score of zero.

Is the cheapest Nemotron 3.5 Lightning endpoint the same product?

Headline cheapest is a lower precision. Like-for-like at the best declared precision is $0.060/M from DeepInfra.

MODEL MONUMENT

Save or share this proof

Badge

The verdict as one SVG, for a README or a docs page. It states this model's dominance status at the balanced workload on the LMArena lens, with the date it was computed, and it is rebuilt with the catalogue — so it changes when the verdict changes, including to one you would rather it did not.

Nemotron 3.5 Lightning — Undominated.ai dominance verdict

Markdown
[![Nemotron 3.5 Lightning — Undominated.ai dominance verdict](https://undominated.ai/badge/nvidia__nemotron-3.5-lightning.svg)](https://undominated.ai/models/nvidia__nemotron-3.5-lightning/)
HTML
<a href="https://undominated.ai/models/nvidia__nemotron-3.5-lightning/"><img src="https://undominated.ai/badge/nvidia__nemotron-3.5-lightning.svg" alt="Nemotron 3.5 Lightning — Undominated.ai dominance verdict" height="36"></a>

Direct file: https://undominated.ai/badge/nvidia__nemotron-3.5-lightning.svg — a static SVG, written by scripts/build-badges.mjs on every build, so the copy you embed is never older than the last deploy.

Evidence & Ask