Nemotron 3 Nano Omni (free) pricing & benchmarks
Released Apr 28, 2026 · served by 1 provider
What Nemotron 3 Nano Omni (free) is
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
- Model ID:
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free - Modality: text+image+audio+video->text
- Tool calling: supported · Structured output: not supported
- Weights:
nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16on Hugging Face
Providers and prices for Nemotron 3 Nano Omni (free)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Nvidiacheapest | Free | Free | 256K | — | 96.4% | — |
FAQ
How much does the Nemotron 3 Nano Omni (free) API cost?
Free per 1M input tokens and Free per 1M output tokens.
Is Nemotron 3 Nano Omni (free) free?
Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.
Which provider is cheapest for Nemotron 3 Nano Omni (free)?
Nvidia at Free input / Free output per 1M tokens.
Can I self-host Nemotron 3 Nano Omni (free)?
Yes — weights are published as nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 on Hugging Face, so you can run it on your own hardware.