TokenRate

Nemotron 3.5 Content Safety Pricing

Fast

NVIDIA · 131K tokens context

Nemotron 3.5 Content Safety from NVIDIA costs $0.200 per 1 million input tokens and $0.200 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 131,072-token context window (approximately 98,304 words) with a 118K-token maximum output. A typical 1,000-token request costs $0.0002 in input charges; a 10,000-token request costs $0.0020.

Nemotron 3.5 Content Safety pricing and capability summary
Input price$0.200 / 1M tokens
Output price$0.200 / 1M tokens
Output / input ratio1.0×
Context window131,072 tokens (~98,304 words)
Maximum output117,964 tokens
Cost per 1K tokens (input)$0.0002
TierFast
Last verified

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B.

Live pricing from OpenRouter

Input Price

$0.200

per 1 million tokens

Output Price

$0.200

per 1 million tokens

Context Window

131K tokens

max 118K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.000267$0.00008
10-page document (2,500 words)3,333$0.000667$0.0002
1,000 lines of code5,000$0.001$0.0003
100K token document100,000$0.02$0.006

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • Extremely cheap at $0.200/1M input tokens
  • Low latency for high-throughput workloads
  • Multimodal: understands images as well as text

Limitations

  • Less capable than flagship models on complex reasoning
  • Quality and availability can vary by hosting provider

Best Use Cases

High-volume classification
Text extraction and summarization
Simple chat and Q&A
Cost-sensitive pipelines

Calculate Nemotron 3.5 Content Safety Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact Nemotron 3.5 Content Safety costs — and compare across all models.

Open Calculator →

Nemotron 3.5 Content Safety — FAQ

Related Models

Related Guides