Skip to content
LLMs
NVIDIA logo

NVIDIA: Nemotron 3.5 Content Safety

by NVIDIA

NVIDIA: Nemotron 3.5 Content Safety is a large language model from NVIDIA. It costs $0.200 per million input tokens and $0.200 per million output tokens. Its context window is 131K tokens.

Free tier availableReasoningStructured outputOpen weightsimage input
Input / 1M tokens
$0.200
Output / 1M tokens
$0.200
Cached input / 1M
-

Not supported

Context window
131K

tokens

Who serves it cheapest

1 hosts serve Nemotron 3.5 Content Safety. Same weights, same API - the price difference is pure margin and routing.

Providers serving NVIDIA: Nemotron 3.5 Content Safety, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DeepInfraCheapestbf16$0.200$0.200131K-100.0%

About Nemotron 3.5 Content Safety

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

Specifications

NVIDIA: Nemotron 3.5 Content Safety specifications
Model IDnvidia/nemotron-3.5-content-safety
ProviderNVIDIA
Context window131K tokens
Max output118K tokens
Input modalitiestext, image
Output modalitiestext
Knowledge cutoff-
Open weightsYes - nvidia/Nemotron-3.5-Content-Safety
ReleasedJune 4, 2026

Cheaper alternatives

Models that cost less than Nemotron 3.5 Content Safety while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does NVIDIA: Nemotron 3.5 Content Safety cost?

$0.200 per million input tokens and $0.200 per million output tokens.

What is the context window of NVIDIA: Nemotron 3.5 Content Safety?

131K tokens, with up to 118K tokens of output per request.

Confirm against the source: NVIDIA official pricing.