Models/Z.ai/GLM 5.3 FlashX

GLM 5.3 FlashX

NEW
Z.ai · Proprietary · Released Sep 18, 2026 · z-ai/glm-5.3-flashx
Best $0.37 in / $1.25 out via Z.AI · fp8
CompareEstimate cost
1.05M contextVisionToolsReasoningStructured outputPrompt caching
Available from 1 providerUSD / MTOK · PRICES 16 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
Pricing variants · USD / MTok in / out
STANDARD
$0.37 / $1.25
per MTok in / out
CACHE READ $0.075
Recent changesRSS ↗
Sep 19, 2026model added
About

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Specifications
Context window1,048,576
Max output131,072
Modalities intext, image, video
Modalities outtext
LicenseProprietary
ReleasedSep 18, 2026
TokenizerOther
Embed badge
GLM 5.3 FlashX price and speed badge
[![GLM 5.3 FlashX](https://modelindex.ai/badge/z-ai/glm-5.3-flashx)](https://modelindex.ai/models/z-ai/glm-5.3-flashx)
Sources
OpenRouter API16 minutes ago
AGGREGATOR
Verified 16 minutes ago · list prices, not negotiated rates.