z-ai Vision Reasoning

GLM 5V Turbo pricing & benchmarks

Released Apr 1, 2026 · served by 1 provider

Input / 1M tokens$1.20
Output / 1M tokens$4.00
Context window203K
Cached input$0.240
1M in + 300K out$2.40

What GLM 5V Turbo is

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Providers and prices for GLM 5V Turbo

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Z.AIcheapest $1.20 $4.00 203K fp8 100.0%

GLM 5V Turbo benchmark results

Design Arena head-to-head results, as reported through the OpenRouter model API.

CategoryArenaRankEloWin rate
androidnative agents #5 1267 54.8%
agenticslides agents #6 1171 51.6%
agenticslides(python-pptx) agents #6 1183 53.9%
pptxslides agents #6 1164 52.1%
agentichtmlslides agents #8 1134 41.7%
agenticslides(html) agents #8 1137 41.8%
python-pptxslides agents #13 1165 51.9%
mobileapps agents #15 1207 50.5%
fullstack agents #16 1199 52.7%
htmlslides agents #17 1145 43.5%
agenticgamedev agents #18 1111 41.9%
webapps agents #21 1179 45.6%
gamedev models #25 1276 54.1%
3d models #28 1267 54.4%
svg models #28 1201 51.1%
godotgamedev agents #30 993 26.5%
codecategories models #32 1260 51.9%
website models #32 1255 50.7%
uicomponent models #33 1249 50.6%
dataviz models #41 1225 48%
asciiart models #46 1137 42.6%

FAQ

How much does the GLM 5V Turbo API cost?

$1.20 per 1M input tokens and $4.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $2.40.

Is GLM 5V Turbo free?

No. It is a paid model starting at $1.20 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for GLM 5V Turbo?

Z.AI at $1.20 input / $4.00 output per 1M tokens (fp8 quantization).

Can I self-host GLM 5V Turbo?

No. GLM 5V Turbo is closed-weight and only available through hosted APIs.