Zhipu AI

GLM 5.2

Flagship model for long-horizon tasks

GLM-5.2 targets project-scale engineering workflows, with a truly usable 1M context window, stable long-horizon execution, and reliable adherence to engineering conventions. Best for coding, agents, and long-running development tasks.

Context window1M
Released2026-06
Relays60 sites
1M context windowAgents and tool useStructured outputCoding and long workflows

Zhipu AI Official Pricing

CNY
Updated: 2026-06-17T10:42:15.998+08:00Source

Input

¥8/ 1M tokens

Output

¥10/ 1M tokens

Cache read

¥2/ 1M tokens

Relay Comparison

Compare token, per-request, or per-second pricing by relay channel.

How should GLM 5.2 relay pricing be compared?

This GLM 5.2 pricing page compares official pricing with public prices from 60 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/21/2026, 10:14.

Data sources
Public price catalogs, official pricing records, and monitoring results.
Metric definitions
Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
Risk note
Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.