Gemma 4 31B (free) pricing & benchmarks
Released Apr 2, 2026 · served by 1 provider · ranked #58 by intelligence in our index
What Gemma 4 31B (free) is
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- Model ID:
google/gemma-4-31b-it:free - Modality: text+image+video->text
- Tool calling: supported · Structured output: not supported
- Weights:
google/gemma-4-31B-iton Hugging Face
Providers and prices for Gemma 4 31B (free)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Google AI Studiocheapest | Free | Free | 262K | — | 99.7% | — |
FAQ
How much does the Gemma 4 31B (free) API cost?
Free per 1M input tokens and Free per 1M output tokens.
Is Gemma 4 31B (free) free?
Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.
Which provider is cheapest for Gemma 4 31B (free)?
Google AI Studio at Free input / Free output per 1M tokens.
Can I self-host Gemma 4 31B (free)?
Yes — weights are published as google/gemma-4-31B-it on Hugging Face, so you can run it on your own hardware.