| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| hvoy | Token | ¥1.26 | ¥3.78 | ¥0.2521 | ¥1.58 |
Alibaba Cloud
Qwen3.8 Max
Qwen 3.8 flagship model
Qwen3.8 Max is Alibaba Cloud Model Studio's Qwen 3.8 flagship model, available on Token Plan personal and team tiers; qwen3.8-max-preview has been retired and auto-routes here.
Relay Comparison
Compare token, per-request, or per-second pricing by relay channel.
Price range ¥4.4 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国模分组-openai协议 | Token | ¥4.4 | ¥13.4 | ¥0.5 | - |
Price range ¥6 - ¥8.4 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Qwen开源 | Token | ¥6 | ¥18 | ¥0.75 | ¥6 |
| Kimi官转 | Token | ¥8.4 | ¥25.2 | ¥1.05 | ¥8.4 |
Price range ¥9.26 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| default | Token | ¥9.26 | ¥27.77 | ¥1.16 | - |
Price range ¥6 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型-oai | Token | ¥6 | ¥18 | ¥1.2 | ¥7.5 |
Price range ¥1.93 - ¥13.8 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Qwen高效优选特惠 | Token | ¥1.93 | ¥5.8 | ¥0.1932 | ¥2.42 |
| default | Token | ¥13.8 | ¥41.4 | ¥1.38 | ¥17.25 |
Price range ¥3.7 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型-openai协议 | Token | ¥3.7 | ¥11.2 | ¥0.4 | - |
Price range ¥510 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| official | Token | ¥510 | ¥510 | - | - |
Price range ¥1.2 - ¥3 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Self-Deployed-1 | Token | ¥1.2 | ¥3.6 | ¥0.15 | - |
| Self-Deployed-2 | Token | ¥2 | ¥6 | ¥0.25 | - |
| Alibaba-2 | Token | ¥3 | ¥9 | ¥0.375 | - |
Price range ¥9.6 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| group-55 | Token | ¥9.6 | ¥28.8 | ¥1.2 | - |
How should Qwen3.8 Max relay pricing be compared?
This Qwen3.8 Max pricing page compares official pricing with public prices from 33 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 23:14.
- Data sources
- Public price catalogs, official pricing records, and monitoring results.
- Metric definitions
- Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
- Risk note
- Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.