google Vision Reasoning

Gemini 3.1 Flash Lite Preview pricing & benchmarks

Released Mar 3, 2026 · served by 1 provider

Input / 1M tokens$0.250
Output / 1M tokens$1.50
Context window1.0M
Cached input$0.025
Intelligence index25.0
Coding index34.7
Agentic index6.2
1M in + 300K out$0.70

What Gemini 3.1 Flash Lite Preview is

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Providers and prices for Gemini 3.1 Flash Lite Preview

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Google AI Studiocheapest $0.450 $2.70 1.0M 99.0%

Gemini 3.1 Flash Lite Preview benchmark results

Design Arena head-to-head results, as reported through the OpenRouter model API.

CategoryArenaRankEloWin rate
asciiart models #19 1205 50.7%
svg models #53 1102 42.5%
uicomponent models #77 1108 37.7%
3d models #81 1105 38.8%
codecategories models #84 1102 36.4%
gamedev models #86 1085 33.7%
website models #86 1107 36.6%
dataviz models #87 1074 33.3%

Cheaper models in the same class

Models scoring within 4 points of Gemini 3.1 Flash Lite Preview on the Intelligence Index, but with a lower output price.

FAQ

How much does the Gemini 3.1 Flash Lite Preview API cost?

$0.250 per 1M input tokens and $1.50 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.70.

Is Gemini 3.1 Flash Lite Preview free?

No. It is a paid model starting at $0.250 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Gemini 3.1 Flash Lite Preview?

Google AI Studio at $0.450 input / $2.70 output per 1M tokens.

Can I self-host Gemini 3.1 Flash Lite Preview?

No. Gemini 3.1 Flash Lite Preview is closed-weight and only available through hosted APIs.