nvidia Open weights Vision Reasoning Free tier

Nemotron 3 Nano Omni (free) pricing & benchmarks

Released Apr 28, 2026 · served by 1 provider

Input / 1M tokensFree
Output / 1M tokensFree
Context window256K
Cached input
Coding index13.8
1M in + 300K out$0.00

What Nemotron 3 Nano Omni (free) is

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Providers and prices for Nemotron 3 Nano Omni (free)

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Nvidiacheapest Free Free 256K 96.4%

FAQ

How much does the Nemotron 3 Nano Omni (free) API cost?

Free per 1M input tokens and Free per 1M output tokens.

Is Nemotron 3 Nano Omni (free) free?

Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.

Which provider is cheapest for Nemotron 3 Nano Omni (free)?

Nvidia at Free input / Free output per 1M tokens.

Can I self-host Nemotron 3 Nano Omni (free)?

Yes — weights are published as nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 on Hugging Face, so you can run it on your own hardware.