Gemini 3.5 Flash vs Kimi K3

Price, context and benchmark scores side by side — every number from the live model catalogue.

Gemini 3.5 Flashgoogle Kimi K3moonshotai
Input / 1M tokens $1.50 $3.00
Output / 1M tokens $9.00 $15.00
1M in + 300K out $4.20 $7.50
Context window 1.0M 1.0M
Intelligence index 50.2 57.1
Coding index 70.1 76.2
Agentic index 37.4 50.1
Cheapest provider Google BaseTen
Providers serving it 2 6
Weights Closed Open
Image input Yes Yes
Tool calling Yes Yes
Reasoning mode Yes Yes
Released May 19, 2026 Jul 16, 2026

Which one to pick

Pick Gemini 3.5 Flash if…

  • Token cost matters: $9.00 per 1M output vs $15.00.

Pick Kimi K3 if…

  • You want the higher general-intelligence score (57.1).
  • Code quality is the priority — it leads on the coding index (76.2).
  • You are running agent loops — higher agentic index (50.1).
  • You need the bigger context window (1.0M).
  • You want the option to self-host — weights are public.
  • You want provider redundancy: 6 providers serve it.

FAQ

Is Gemini 3.5 Flash better than Kimi K3?

Kimi K3 scores higher on general intelligence (50.2 vs 57.1), while Gemini 3.5 Flash is 1.7× cheaper per output token ($9.00 vs $15.00). For code, Kimi K3 leads (76.2 coding index). Kimi K3 takes the larger context at 1.0M.

Which is cheaper, Gemini 3.5 Flash or Kimi K3?

Gemini 3.5 Flash, at $9.00 per 1M output tokens versus $15.00. On a 1M input + 300K output workload that is $4.20 against $7.50.

Can I self-host Gemini 3.5 Flash or Kimi K3?

Gemini 3.5 Flash: closed weights, API only. Kimi K3: open weights (moonshotai/Kimi-K3).

Related comparisons