GLM 5.3 FlashX
NEWZ.ai · Proprietary · Released Sep 18, 2026 · z-ai/glm-5.3-flashx
Best $0.37 in / $1.25 out via Z.AI · fp8
1.05M contextVisionToolsReasoningStructured outputPrompt caching
Available from 1 providerUSD / MTOK · PRICES 16 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.37 | $1.25 | $0.075 | 1.05M | fp8 | — | — | 100.0% |
Pricing variants · USD / MTok in / out
STANDARD
$0.37 / $1.25
per MTok in / out
CACHE READ $0.075
Recent changesRSS ↗
Sep 19, 2026model added
About
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Specifications
Context window1,048,576
Max output131,072
Modalities intext, image, video
Modalities outtext
LicenseProprietary
ReleasedSep 18, 2026
TokenizerOther
Related
Embed badge
[](https://modelindex.ai/models/z-ai/glm-5.3-flashx)Sources
OpenRouter API16 minutes ago
AGGREGATORVerified 16 minutes ago · list prices, not negotiated rates.