z-ai Reasoning

GLM 5 Turbo pricing & benchmarks

Released Mar 15, 2026 · served by 1 provider

Input / 1M tokens$1.20
Output / 1M tokens$4.00
Context window203K
Cached input$0.240
1M in + 300K out$2.40

What GLM 5 Turbo is

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

Providers and prices for GLM 5 Turbo

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Z.AIcheapest $1.20 $4.00 203K 99.9%

GLM 5 Turbo benchmark results

Design Arena head-to-head results, as reported through the OpenRouter model API.

CategoryArenaRankEloWin rate
svg models #10 1258 57.3%
3d models #14 1309 58.8%
dataviz models #14 1294 57.6%
gamedev models #14 1314 59.1%
codecategories models #16 1300 56.7%
uicomponent models #16 1300 57.5%
website models #17 1296 55.7%
asciiart models #24 1191 50%

FAQ

How much does the GLM 5 Turbo API cost?

$1.20 per 1M input tokens and $4.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $2.40.

Is GLM 5 Turbo free?

No. It is a paid model starting at $1.20 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for GLM 5 Turbo?

Z.AI at $1.20 input / $4.00 output per 1M tokens.

Can I self-host GLM 5 Turbo?

No. GLM 5 Turbo is closed-weight and only available through hosted APIs.