| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| hvoy | Token | ¥10.22 | ¥52.2 | ¥1.02 | - |
Moonshot AI
Kimi K3
Kimi flagship with 2.8T parameters, native vision, and 1M context
Kimi K3 is Moonshot AI's most capable flagship, with 2.8 trillion parameters, Kimi Delta Attention, Attention Residuals, sparse MoE, native vision, and a 1M-token context window. It targets long-horizon coding, knowledge work, and reasoning, with tool calling, structured output, and automatic context caching.
Moonshot AI Official Pricing
CNYInput
Official flat pricing with no context-length tiers: CNY 20 uncached input, CNY 2 cached input, and CNY 100 output per 1M tokens; the context window is 1,048,576 tokens.
Output
Official flat pricing with no context-length tiers: CNY 20 uncached input, CNY 2 cached input, and CNY 100 output per 1M tokens; the context window is 1,048,576 tokens.
Cache read
Official flat pricing with no context-length tiers: CNY 20 uncached input, CNY 2 cached input, and CNY 100 output per 1M tokens; the context window is 1,048,576 tokens.
Official flat pricing with no context-length tiers: CNY 20 uncached input, CNY 2 cached input, and CNY 100 output per 1M tokens; the context window is 1,048,576 tokens.
Relay Comparison
Compare token, per-request, or per-second pricing by relay channel.
Price range ¥12.5 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| kimi-stable | Token | ¥12.5 | ¥62.49 | ¥1.25 | - |
Price range ¥10.08 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| kimi 官方满血线路 | Token | ¥10.08 | ¥50.39 | ¥1.01 | - |
Price range ¥20 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| default | Token | ¥20 | ¥100 | ¥2 | - |
Price range ¥5 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| DeepSeek | Token | ¥5 | ¥25 | ¥0.5 | ¥0 |
| glm | Token | ¥5 | ¥25 | ¥0.5 | ¥0 |
| kimi | Token | ¥5 | ¥25 | ¥0.5 | ¥0 |
Price range ¥10 - ¥14 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Kimi开源 | Token | ¥10 | ¥50 | ¥1 | - |
| Kimi官转 | Token | ¥14 | ¥70 | ¥1.4 | - |
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| default | Token | ¥12 | ¥60 | ¥1.2 | ¥0 |
Price range ¥10.5 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| kimi-core | Token | ¥10.5 | ¥52.5 | ¥1.05 | - |
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型聚合 | Token | ¥12 | ¥60 | ¥1.2 | - |
Price range ¥10 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型聚合专区 | Token | ¥10 | ¥50 | ¥1 | - |
How should Kimi K3 relay pricing be compared?
This Kimi K3 pricing page compares official pricing with public prices from 75 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 23:14.
- Data sources
- Public price catalogs, official pricing records, and monitoring results.
- Metric definitions
- Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
- Risk note
- Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.