Models/Z.ai/GLM 5.3 Prime

GLM 5.3 Prime

NEW
Z.ai · Proprietary · Released Sep 23, 2026 · z-ai/glm-5.3-prime
Best $2.80 in / $8.80 out via Alibaba
CompareEstimate cost
1M contextToolsReasoningStructured outputPrompt caching
Available from 1 providerUSD / MTOK · PRICES 14 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$2.80 / $8.80
per MTok in / out
CACHE READ $0.56
Recent changesRSS ↗
Sep 24, 2026model added
About

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

Specifications
Context window1,000,000
Max output131,072
Modalities intext
Modalities outtext
LicenseProprietary
ReleasedSep 23, 2026
TokenizerOther
Embed badge
GLM 5.3 Prime price and speed badge
[![GLM 5.3 Prime](https://modelindex.ai/badge/z-ai/glm-5.3-prime)](https://modelindex.ai/models/z-ai/glm-5.3-prime)
Sources
OpenRouter API14 minutes ago
AGGREGATOR
Verified 14 minutes ago · list prices, not negotiated rates.