nvidia Open weights Vision Reasoning Free tier

Nemotron Nano 12B 2 VL (free) pricing & benchmarks

Released Oct 28, 2025 · served by 1 provider

Input / 1M tokensFree
Output / 1M tokensFree
Context window128K
Cached input
1M in + 300K out$0.00

What Nemotron Nano 12B 2 VL (free) is

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

Providers and prices for Nemotron Nano 12B 2 VL (free)

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Nvidiacheapest Free Free 128K 88.2%

FAQ

How much does the Nemotron Nano 12B 2 VL (free) API cost?

Free per 1M input tokens and Free per 1M output tokens.

Is Nemotron Nano 12B 2 VL (free) free?

Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.

Which provider is cheapest for Nemotron Nano 12B 2 VL (free)?

Nvidia at Free input / Free output per 1M tokens.

Can I self-host Nemotron Nano 12B 2 VL (free)?

Yes — weights are published as nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16 on Hugging Face, so you can run it on your own hardware.