Nemotron 3.5 Lightning Pricing
FastNVIDIA · 262K tokens context
Nemotron 3.5 Lightning from NVIDIA costs $0.100 per 1 million input tokens and $0.250 per 1 million output tokens as of August 2026 (live OpenRouter data). The model supports a 262,144-token context window (approximately 196,608 words) with a 262K-token maximum output. A typical 1,000-token request costs $0.0001 in input charges; a 10,000-token request costs $0.0010.
| Input price | $0.100 / 1M tokens |
|---|---|
| Output price | $0.250 / 1M tokens |
| Output / input ratio | 2.5× |
| Context window | 262,144 tokens (~196,608 words) |
| Maximum output | 262,144 tokens |
| Cost per 1K tokens (input) | $0.0001 |
| Tier | Fast |
| Last verified |
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total.
Input Price
$0.100
per 1 million tokens
Output Price
$0.250
per 1 million tokens
Context Window
262K tokens
max 262K output
Cost Examples
| Request Type | Tokens | Input Cost | Output Cost |
|---|---|---|---|
| 1,000 word article | 1,333 | $0.000133 | $0.0001 |
| 10-page document (2,500 words) | 3,333 | $0.000333 | $0.00025 |
| 1,000 lines of code | 5,000 | $0.0005 | $0.000375 |
| 100K token document | 100,000 | $0.01 | $0.0075 |
Output cost estimated at 30% of input token count. Use the calculator for exact figures.
Strengths
- ✓Extremely cheap at $0.100/1M input tokens
- ✓Large 262K-token context window
- ✓Low latency for high-throughput workloads
Limitations
- –Less capable than flagship models on complex reasoning
- –Quality and availability can vary by hosting provider
Best Use Cases
Calculate Nemotron 3.5 Lightning Costs
Use the TokenRate calculator to convert any budget, token count, or text into exact Nemotron 3.5 Lightning costs — and compare across all models.
Open Calculator →Nemotron 3.5 Lightning — FAQ
Related Models
Related Guides