google Vision Reasoning

Gemini 3 Flash Preview pricing & benchmarks

Released Dec 17, 2025 · served by 2 providers

Input / 1M tokens$0.500
Output / 1M tokens$3.00
Context window1.0M
Cached input$0.050
1M in + 300K out$1.40

What Gemini 3 Flash Preview is

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

Providers and prices for Gemini 3 Flash Preview

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Googlecheapest $0.900 $5.40 1.0M 97.2%
Google AI Studio $0.900 $5.40 1.0M 99.7%

Gemini 3 Flash Preview benchmark results

Design Arena head-to-head results, as reported through the OpenRouter model API.

CategoryArenaRankEloWin rate
agenticslides agents #9 1073 39.3%
agenticslides(python-pptx) agents #9 1075 39.3%
godotgamedev agents #15 1161 50.6%
python-pptxslides agents #18 1013 38.3%
fullstack agents #21 1111 47.1%
mobileapps agents #21 1180 49.1%
webapps agents #24 1168 49.2%
androidnative agents #29 1039 48%
3d models #36 1241 62.7%
codecategories models #40 1220 57.6%
website models #40 1221 57%
gamedev models #43 1222 58.3%

FAQ

How much does the Gemini 3 Flash Preview API cost?

$0.500 per 1M input tokens and $3.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $1.40.

Is Gemini 3 Flash Preview free?

No. It is a paid model starting at $0.500 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Gemini 3 Flash Preview?

Google at $0.900 input / $5.40 output per 1M tokens.

Can I self-host Gemini 3 Flash Preview?

No. Gemini 3 Flash Preview is closed-weight and only available through hosted APIs.