NVIDIA: Nemotron 3.5 Content Safety
by NVIDIANVIDIA: Nemotron 3.5 Content Safety is a large language model from NVIDIA. It costs $0.200 per million input tokens and $0.200 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.200
- Output / 1M tokens
- $0.200
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
1 hosts serve Nemotron 3.5 Content Safety. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestbf16 | $0.200 | $0.200 | 131K | - | 100.0% |
About Nemotron 3.5 Content Safety
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Specifications
| Model ID | nvidia/nemotron-3.5-content-safety |
|---|---|
| Provider | NVIDIA |
| Context window | 131K tokens |
| Max output | 118K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - nvidia/Nemotron-3.5-Content-Safety |
| Released | June 4, 2026 |
Cheaper alternatives
Models that cost less than Nemotron 3.5 Content Safety while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does NVIDIA: Nemotron 3.5 Content Safety cost?
$0.200 per million input tokens and $0.200 per million output tokens.
What is the context window of NVIDIA: Nemotron 3.5 Content Safety?
131K tokens, with up to 118K tokens of output per request.
Confirm against the source: NVIDIA official pricing.