Gemma 3n 4B pricing & benchmarks
Released May 20, 2025 · served by 1 provider
What Gemma 3n 4B is
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
- Model ID:
google/gemma-3n-e4b-it - Modality: text->text
- Knowledge cutoff: 2024-08-31
- Tool calling: not supported · Structured output: supported
- Weights:
google/gemma-3n-E4B-iton Hugging Face
Providers and prices for Gemma 3n 4B
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Togethercheapest | $0.060 | $0.120 | 33K | — | 100.0% | — |
FAQ
How much does the Gemma 3n 4B API cost?
$0.060 per 1M input tokens and $0.120 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.10.
Is Gemma 3n 4B free?
No. It is a paid model starting at $0.060 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Gemma 3n 4B?
Together at $0.060 input / $0.120 output per 1M tokens.
Can I self-host Gemma 3n 4B?
Yes — weights are published as google/gemma-3n-E4B-it on Hugging Face, so you can run it on your own hardware.