| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| hvoy | Token | ¥0.1 | ¥0.3 | ¥0.01 | ¥0.125 |
Alibaba Cloud
Qwen3.8 Flash
Qwen 3.8 fast model
Qwen3.8 Flash is the fast variant in the Qwen 3.8 series, available on Token Plan personal and team tiers for high-frequency and lighter workloads.
1M input context1M
Released2026-08
Relays18 sites
FastAgent1M context
Relay Comparison
Compare token, per-request, or per-second pricing by relay channel.
Unit: token prices are ¥ / 1M tokens; per-request and per-second rows use their shown unit
Price range ¥0.4 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Qwen开源 | Token | ¥0.4 | ¥1.35 | ¥0.05 | ¥0.4 |
Price range ¥0.64 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| default | Token | ¥0.64 | ¥2.16 | ¥0.08 | - |
Price range ¥0.5 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型-oai | Token | ¥0.5 | ¥1.5 | - | - |
Price range ¥0.1546 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Qwen高效优选特惠 | Token | ¥0.1546 | ¥0.454 | ¥0.0168 | ¥0.0001 |
Price range ¥0.225 - ¥0.33 / 1M tokens
Alibaba-2
BillingToken
Input¥0.225
Output¥0.705
Cache hit¥0.024
Cache write-
Alibaba-3
BillingToken
Input¥0.33
Output¥1.03
Cache hit¥0.0352
Cache write-
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| Alibaba-2 | Token | ¥0.225 | ¥0.705 | ¥0.024 | - |
| Alibaba-3 | Token | ¥0.33 | ¥1.03 | ¥0.0352 | - |
Price range ¥0.64 / 1M tokens
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| group-55 | Token | ¥0.64 | ¥2.16 | ¥0.08 | - |
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国产模型聚合 | Token | ¥0.48 | ¥1.62 | ¥0.06 | - |
国模全家桶-openai接口
BillingToken
Input¥0.33
Output¥0.96
Cache hit¥0.06
Cache write¥0.03
国模高缓存-openai接口
BillingToken
Input¥0.66
Output¥1.92
Cache hit¥0.12
Cache write¥0.06
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国模全家桶-openai接口 | Token | ¥0.33 | ¥0.96 | ¥0.06 | ¥0.03 |
| 国模高缓存-openai接口 | Token | ¥0.66 | ¥1.92 | ¥0.12 | ¥0.06 |
国模分组
BillingToken
Input¥0.44
Output¥1.28
Cache hit¥0.08
Cache write¥0.04
国模高缓分组
BillingToken
Input¥0.66
Output¥1.92
Cache hit¥0.12
Cache write¥0.06
| Channel | Billing | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|---|
| 国模分组 | Token | ¥0.44 | ¥1.28 | ¥0.08 | ¥0.04 |
| 国模高缓分组 | Token | ¥0.66 | ¥1.92 | ¥0.12 | ¥0.06 |
How should Qwen3.8 Flash relay pricing be compared?
This Qwen3.8 Flash pricing page compares official pricing with public prices from 18 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 23:14.
- Data sources
- Public price catalogs, official pricing records, and monitoring results.
- Metric definitions
- Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
- Risk note
- Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.