Nemotron Nano 9B V2 (free) pricing & benchmarks
Released Sep 5, 2025 · served by 1 provider
What Nemotron Nano 9B V2 (free) is
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
- Model ID:
nvidia/nemotron-nano-9b-v2:free - Modality: text->text
- Knowledge cutoff: 2025-03-31
- Tool calling: supported · Structured output: supported
- Weights:
nvidia/NVIDIA-Nemotron-Nano-9B-v2on Hugging Face
Providers and prices for Nemotron Nano 9B V2 (free)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Nvidiacheapest | Free | Free | 128K | bf16 | 90.3% | — |
FAQ
How much does the Nemotron Nano 9B V2 (free) API cost?
Free per 1M input tokens and Free per 1M output tokens.
Is Nemotron Nano 9B V2 (free) free?
Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.
Which provider is cheapest for Nemotron Nano 9B V2 (free)?
Nvidia at Free input / Free output per 1M tokens (bf16 quantization).
Can I self-host Nemotron Nano 9B V2 (free)?
Yes — weights are published as nvidia/NVIDIA-Nemotron-Nano-9B-v2 on Hugging Face, so you can run it on your own hardware.