Find a model

Search models and open their dedicated profile. Use the arrow keys to navigate and Enter to open.

glm-5.2-fast

Z AI · 1 configuration

Intelligence—
Input / 1M$2.10
Output / 1M$6.60
Context1M

Tariff: Requesty · Source: models-dev · USD per million tokens

Thinking configuration

What does more thinking buy you?

Compare this model’s measured thinking configurations.

◌

No pair of measurements available

Change the filters or pick another benchmark. Missing values are not estimated.

Selected: thinking default · Click a point to switch configuration

GLM-5.2 introduces a robust 1M-token context and advanced, multi-effort coding capabilities to significantly enhance performance on long-horizon tasks. Its new IndexShare architecture and improved MTP layer simultaneously boost efficiency by reducing per-token FLOPs and increasing speculative decoding lengths. A 743B-parameter model in Zhipu AI's GLM series, designed to plan, execute, and iterate autonomously on extended, engineering-grade tasks.

This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.

Specifications from the catalog

Inputtext
Outputtext
Tool callingYes
Structured outputNo
SpeedNot available
First tokenNot available
Maximum output131.1K
Catalog listing date2026-07-13