Z.ai: GLM 5.3 FlashX
Z AI · 1 configuration
Intelligence—
Input / 1M$0.37
Output / 1M$1.25
Context1M
Tariff: Kilo Gateway · Source: models-dev · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
◌
No pair of measurements available
Change the filters or pick another benchmark. Missing values are not estimated.
Selected: thinking default · Click a point to switch configuration
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.
Specifications from the catalog
Inputtext, image, video
Outputtext
Tool callingYes
Structured outputNo
SpeedNot available
First tokenNot available
Maximum output131.1K
Catalog listing date2026-09-18