Nemotron Nano 12B 2 VL (free) pricing & benchmarks
Released Oct 28, 2025 · served by 1 provider
What Nemotron Nano 12B 2 VL (free) is
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
- Model ID:
nvidia/nemotron-nano-12b-v2-vl:free - Modality: text+image+video->text
- Tool calling: supported · Structured output: not supported
- Weights:
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16on Hugging Face
Providers and prices for Nemotron Nano 12B 2 VL (free)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Nvidiacheapest | Free | Free | 128K | — | 88.2% | — |
FAQ
How much does the Nemotron Nano 12B 2 VL (free) API cost?
Free per 1M input tokens and Free per 1M output tokens.
Is Nemotron Nano 12B 2 VL (free) free?
Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.
Which provider is cheapest for Nemotron Nano 12B 2 VL (free)?
Nvidia at Free input / Free output per 1M tokens.
Can I self-host Nemotron Nano 12B 2 VL (free)?
Yes — weights are published as nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16 on Hugging Face, so you can run it on your own hardware.