| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| hvoy | Token | ¥0.8 | ¥2.8 | ¥0.16 | - |
Zhipu AI
GLM 5.1
A flagship model for agents and complex coding
GLM-5.1 is Zhipu AI's next-generation flagship text model, with stronger thinking, coding, and agent-task capabilities. It supports long context, context caching, structured output, and function calling, making it suitable for complex coding, tool use, multi-step reasoning, and long-running agent workflows.
Zhipu AI Official Pricing
CNYInput
Output
Cache read
Relay Comparison
Compare token, per-request, or per-second pricing by relay channel.
Price range ¥3.6 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产官方模型 | Token | ¥3.6 | ¥14.4 | ¥0.78 | - |
Price range ¥3 - ¥4.2 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| GLM智谱开源 | Token | ¥3 | ¥12 | ¥0.65 | - |
| GLM智谱官转 | Token | ¥4.2 | ¥16.8 | ¥0.91 | - |
Price range ¥4 - ¥6.4 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| glm-sale | Token | ¥4 | ¥16 | ¥0.8 | - |
| default | Token | ¥6.4 | ¥25.6 | ¥1.28 | - |
Price range ¥0.89 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| default | Token | ¥0.89 | ¥3.48 | ¥0.19 | - |
Price range ¥5.92 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Alibaba | Token | ¥5.92 | ¥23.8 | ¥1.29 | - |
| official | Token | ¥5.92 | ¥23.8 | ¥1.29 | - |
Price range ¥1.4 - ¥2.1 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Self-Deployed-2 | Token | ¥1.4 | ¥4.4 | ¥0.308 | - |
| Alibaba-2 | Token | ¥2.1 | ¥6.6 | ¥0.462 | - |
| Self-Deployed-3 | Token | ¥2.1 | ¥6.6 | ¥0.462 | - |
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| marketplace | Token | ¥0.478 | ¥1.68 | ¥0.1217 | ¥0.1217 |
Price range ¥3.2 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产大模型(Claude 协议) | Token | ¥3.2 | ¥11.2 | ¥0.8 | ¥0.4 |
| 国产大模型(OpenAI 协议) | Token | ¥3.2 | ¥11.2 | ¥0.8 | ¥0.4 |
Price range ¥1.88 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国模2折福利 | Token | ¥1.88 | ¥5.89 | ¥0.3483 | ¥0 |
| 国模福利-chat端点 | Token | ¥1.88 | ¥5.89 | ¥0.3483 | ¥0 |
| 国模福利-messages端点 | Token | ¥1.88 | ¥5.89 | ¥0.3483 | ¥0 |
How should GLM 5.1 relay pricing be compared?
This GLM 5.1 pricing page compares official pricing with public prices from 43 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 23:14.
- Data sources
- Public price catalogs, official pricing records, and monitoring results.
- Metric definitions
- Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
- Risk note
- Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.