nvidia Open weights Reasoning Free tier

Nemotron Nano 9B V2 (free) pricing & benchmarks

Released Sep 5, 2025 · served by 1 provider

Input / 1M tokensFree
Output / 1M tokensFree
Context window128K
Cached input
1M in + 300K out$0.00

What Nemotron Nano 9B V2 (free) is

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

Providers and prices for Nemotron Nano 9B V2 (free)

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Nvidiacheapest Free Free 128K bf16 90.3%

FAQ

How much does the Nemotron Nano 9B V2 (free) API cost?

Free per 1M input tokens and Free per 1M output tokens.

Is Nemotron Nano 9B V2 (free) free?

Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.

Which provider is cheapest for Nemotron Nano 9B V2 (free)?

Nvidia at Free input / Free output per 1M tokens (bf16 quantization).

Can I self-host Nemotron Nano 9B V2 (free)?

Yes — weights are published as nvidia/NVIDIA-Nemotron-Nano-9B-v2 on Hugging Face, so you can run it on your own hardware.