Zhipu AI

GLM 5.3

Next-generation GLM for reasoning, tools, and long-horizon coding

GLM 5.3 is Zhipu's next-generation flagship for reasoning, tool use, and long-horizon coding, with open weights and about 1M context. Best for engineering work, agent orchestration, and production jobs that need stable long-context execution; verify exact capabilities and endpoints with the official channel.

1M-token context1M
Released2026-08
Relays63 sites
ReasoningToolsOpen weights1M context

Zhipu AI Official Pricing

CNY
Updated: 2026-08-19T12:04:48.443+08:00Source

observed_hvoyweb · Input

~¥8/ 1M tokens

Observed in the hvoyweb active model catalog on 2026-08-18; verify the Zhipu official pricing page before publishing.

observed_hvoyweb · Output

~¥28/ 1M tokens

Observed in the hvoyweb active model catalog on 2026-08-18; verify the Zhipu official pricing page before publishing.

observed_hvoyweb · Cache read

~¥2/ 1M tokens

Observed in the hvoyweb active model catalog on 2026-08-18; verify the Zhipu official pricing page before publishing.

Observed in the hvoyweb active model catalog on 2026-08-18; verify the Zhipu official pricing page before publishing.

Relay Comparison

Compare token, per-request, or per-second pricing by relay channel.

Price range ¥1.88 / 1M tokens

How should GLM 5.3 relay pricing be compared?

This GLM 5.3 pricing page compares official pricing with public prices from 63 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 15:34.

Data sources
Public price catalogs, official pricing records, and monitoring results.
Metric definitions
Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
Risk note
Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.