GLM 5V Turbo pricing & benchmarks
Released Apr 1, 2026 · served by 1 provider
What GLM 5V Turbo is
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
- Model ID:
z-ai/glm-5v-turbo - Modality: text+image+video->text
- Tool calling: supported · Structured output: not supported
Providers and prices for GLM 5V Turbo
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Z.AIcheapest | $1.20 | $4.00 | 203K | fp8 | 100.0% | — |
GLM 5V Turbo benchmark results
Design Arena head-to-head results, as reported through the OpenRouter model API.
| Category | Arena | Rank | Elo | Win rate |
|---|---|---|---|---|
| androidnative | agents | #5 | 1267 | 54.8% |
| agenticslides | agents | #6 | 1171 | 51.6% |
| agenticslides(python-pptx) | agents | #6 | 1183 | 53.9% |
| pptxslides | agents | #6 | 1164 | 52.1% |
| agentichtmlslides | agents | #8 | 1134 | 41.7% |
| agenticslides(html) | agents | #8 | 1137 | 41.8% |
| python-pptxslides | agents | #13 | 1165 | 51.9% |
| mobileapps | agents | #15 | 1207 | 50.5% |
| fullstack | agents | #16 | 1199 | 52.7% |
| htmlslides | agents | #17 | 1145 | 43.5% |
| agenticgamedev | agents | #18 | 1111 | 41.9% |
| webapps | agents | #21 | 1179 | 45.6% |
| gamedev | models | #25 | 1276 | 54.1% |
| 3d | models | #28 | 1267 | 54.4% |
| svg | models | #28 | 1201 | 51.1% |
| godotgamedev | agents | #30 | 993 | 26.5% |
| codecategories | models | #32 | 1260 | 51.9% |
| website | models | #32 | 1255 | 50.7% |
| uicomponent | models | #33 | 1249 | 50.6% |
| dataviz | models | #41 | 1225 | 48% |
| asciiart | models | #46 | 1137 | 42.6% |
FAQ
How much does the GLM 5V Turbo API cost?
$1.20 per 1M input tokens and $4.00 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $2.40.
Is GLM 5V Turbo free?
No. It is a paid model starting at $1.20 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for GLM 5V Turbo?
Z.AI at $1.20 input / $4.00 output per 1M tokens (fp8 quantization).
Can I self-host GLM 5V Turbo?
No. GLM 5V Turbo is closed-weight and only available through hosted APIs.