google Vision Reasoning

Gemini 2.5 Flash pricing & benchmarks

Released Jun 17, 2025 · served by 2 providers

Input / 1M tokens$0.300
Output / 1M tokens$2.50
Context window1.0M
Cached input$0.030
1M in + 300K out$1.05

What Gemini 2.5 Flash is

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Providers and prices for Gemini 2.5 Flash

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Googlecheapest $0.300 $2.50 1.0M 91.4%
Google AI Studio $0.540 $4.50 1.0M 99.8%

Gemini 2.5 Flash benchmark results

Design Arena head-to-head results, as reported through the OpenRouter model API.

CategoryArenaRankEloWin rate
svg models #61 1070 43.1%
dataviz models #66 1156 48.4%
uicomponent models #72 1129 48.9%
3d models #75 1128 47.4%
website models #78 1140 47.1%
codecategories models #79 1135 46.9%
gamedev models #79 1122 44.3%

FAQ

How much does the Gemini 2.5 Flash API cost?

$0.300 per 1M input tokens and $2.50 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $1.05.

Is Gemini 2.5 Flash free?

No. It is a paid model starting at $0.300 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Gemini 2.5 Flash?

Google at $0.300 input / $2.50 output per 1M tokens.

Can I self-host Gemini 2.5 Flash?

No. Gemini 2.5 Flash is closed-weight and only available through hosted APIs.