google Open weights

Gemma 3n 4B pricing & benchmarks

Released May 20, 2025 · served by 1 provider

Input / 1M tokens$0.060
Output / 1M tokens$0.120
Context window33K
Cached input
Coding index3.2
1M in + 300K out$0.10

What Gemma 3n 4B is

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

Providers and prices for Gemma 3n 4B

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Togethercheapest $0.060 $0.120 33K 100.0%

FAQ

How much does the Gemma 3n 4B API cost?

$0.060 per 1M input tokens and $0.120 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.10.

Is Gemma 3n 4B free?

No. It is a paid model starting at $0.060 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Gemma 3n 4B?

Together at $0.060 input / $0.120 output per 1M tokens.

Can I self-host Gemma 3n 4B?

Yes — weights are published as google/gemma-3n-E4B-it on Hugging Face, so you can run it on your own hardware.