Nvidia
Nvidia
EfficientAUTO PROFILELIVE PRICING

Nemotron 3.5 Lightning

text->text

Nemotron 3.5 Lightning is a general-purpose chat-tuned model. Available via Nvidia. 1,000,000-token context. Budget pricing.

Nemotron 3.5 Lightning is a efficient AI model from Nvidia. It costs $0.080 per million input tokens and $0.200 per million output tokens (blended $0.116/M), with a 1,000,000-token context window.

INPUT
$0.080/M
per million input tokens
OUTPUT
$0.200/M
per million output tokens
BLENDED 70/30
$0.116/M

0.0%over 42 days · hover to read

Nemotron 3.5 Lightning — blended price

Reconstructed from 42 days of recorded rate-card changes. Prices hold flat between changes because that is what a posted price does — no value here is interpolated.

CONTEXT
1,000,000
tokens
What it is good at
  • Solid general chat performance
  • Reasonable instruction following
Typical use cases
  • General chat
  • Drafting and rewriting
  • Personal productivity

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from Nvidia

See all 6 →

Price history

8 recorded changes
Recorded price changes for Nemotron 3.5 Lightning, newest first
DateRateWasNowChange
24 Sept 2026Input$0.070$0.080+14%
19 Sept 2026Input$0.080$0.070-12%
29 Aug 2026Input$0.100$0.080-20%
29 Aug 2026Output$0.250$0.200-20%
28 Aug 2026Output$0.200$0.250+25%
28 Aug 2026Input$0.080$0.100+25%
17 Aug 2026Input$0.100$0.080-20%
17 Aug 2026Output$0.250$0.200-20%

Rates per million tokens, as published by the provider or its router listing. Tracking runs since 3 May; anything earlier is not on record. See every tracked price change.

Frequently asked questions

How much does Nemotron 3.5 Lightning cost?

Nemotron 3.5 Lightning costs $0.080 per million input tokens and $0.200 per million output tokens, for a blended reference rate of $0.116 per million tokens.

What is Nemotron 3.5 Lightning's context window?

Nemotron 3.5 Lightning supports up to 1,000,000 tokens of context in a single request.

What is Nemotron 3.5 Lightning best for?

Nemotron 3.5 Lightning is well suited to Solid general chat performance, Reasonable instruction following and General chat.

Who makes Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is developed and served by Nvidia.

Terms used on this page