GLM 5 Turbo
Z.ai · Proprietary · Released Mar 15, 2026 · z-ai/glm-5-turbo
Best $1.20 in / $4.00 out via Z.AI
203K contextToolsReasoningStructured outputPrompt caching
Available from 1 providerUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 6 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $1.20 | $4.00 | $0.24 | 203K | — | ✕ | ✕ | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 6 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$1.20 / $4.00
per MTok in / out
CACHE READ $0.24
About
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
Specifications
Context window202,752
Max output131,072
Modalities intext
Modalities outtext
LicenseProprietary
ReleasedMar 15, 2026
TokenizerOther
Related
Embed badge
[](https://modelindex.ai/models/z-ai/glm-5-turbo)Sources
OpenRouter API23 minutes ago
AGGREGATORModelIndex benchmarks6 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.