Nemotron 3.5 Lightning
Nemotron 3.5 Lightning is a general-purpose chat-tuned model. Available via Nvidia. 1,000,000-token context. Budget pricing.
Nemotron 3.5 Lightning is a efficient AI model from Nvidia. It costs $0.080 per million input tokens and $0.200 per million output tokens (blended $0.116/M), with a 1,000,000-token context window.
0.0%over 42 days · hover to read
- Solid general chat performance
- Reasonable instruction following
- General chat
- Drafting and rewriting
- Personal productivity
Benchmarks
More from Nvidia
See all 6 →Price history
8 recorded changes| Date | Rate | Was | Now | Change |
|---|---|---|---|---|
| 24 Sept 2026 | Input | $0.070 | $0.080 | +14% |
| 19 Sept 2026 | Input | $0.080 | $0.070 | -12% |
| 29 Aug 2026 | Input | $0.100 | $0.080 | -20% |
| 29 Aug 2026 | Output | $0.250 | $0.200 | -20% |
| 28 Aug 2026 | Output | $0.200 | $0.250 | +25% |
| 28 Aug 2026 | Input | $0.080 | $0.100 | +25% |
| 17 Aug 2026 | Input | $0.100 | $0.080 | -20% |
| 17 Aug 2026 | Output | $0.250 | $0.200 | -20% |
Rates per million tokens, as published by the provider or its router listing. Tracking runs since 3 May; anything earlier is not on record. See every tracked price change.
Frequently asked questions
How much does Nemotron 3.5 Lightning cost?
Nemotron 3.5 Lightning costs $0.080 per million input tokens and $0.200 per million output tokens, for a blended reference rate of $0.116 per million tokens.
What is Nemotron 3.5 Lightning's context window?
Nemotron 3.5 Lightning supports up to 1,000,000 tokens of context in a single request.
What is Nemotron 3.5 Lightning best for?
Nemotron 3.5 Lightning is well suited to Solid general chat performance, Reasonable instruction following and General chat.
Who makes Nemotron 3.5 Lightning?
Nemotron 3.5 Lightning is developed and served by Nvidia.