NVIDIA: Nemotron 3 Nano Omni (free)
by NVIDIANVIDIA: Nemotron 3 Nano Omni (free) is a large language model from NVIDIA. It costs Free per million input tokens and Free per million output tokens. Its context window is 256K tokens.
- Input / 1M tokens
- Free
- Output / 1M tokens
- Free
- Cached input / 1M
- -
- Context window
- 256K
Not supported
tokens
Benchmarks
Independent scores published alongside the catalogue.
- Coding index
- 13.8
About Nemotron 3 Nano Omni (free)
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Specifications
| Model ID | nvidia/nemotron-3-nano-omni-30b-a3b-reasoning |
|---|---|
| Provider | NVIDIA |
| Context window | 256K tokens |
| Max output | 66K tokens |
| Input modalities | text, audio, image, video |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 |
| Released | April 28, 2026 |
Cheaper alternatives
Models that cost less than Nemotron 3 Nano Omni (free) while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does NVIDIA: Nemotron 3 Nano Omni (free) cost?
Free per million input tokens and Free per million output tokens.
What is the context window of NVIDIA: Nemotron 3 Nano Omni (free)?
256K tokens, with up to 66K tokens of output per request.
Confirm against the source: NVIDIA official pricing.