Google

Gemini 3.7 Flash

Google multimodal model for fast agents, coding, and complex multi-step reasoning

Gemini 3.7 Flash is a Google multimodal model released on 2026-08-13 for fast agentic workflows, coding, tool use, and complex multi-step reasoning, with a 1,048,576-token context window.

1,048,576-token context1M
Released2026-08-13
Relays63 sites
High speedMultimodalCodingReasoningTool use1M context

Google Official Pricing

CNY
Updated: 2026-08-19T12:34:41.697+08:00Source

observed_hvoyweb · Input

~¥5.4/ 1M tokens

Observed from the hvoyweb active catalog and LLM metadata on 2026-08-19; USD prices are converted with the app_config FX rate. Verify Google official pricing before publication.

observed_hvoyweb · Output

~¥27/ 1M tokens

Observed from the hvoyweb active catalog and LLM metadata on 2026-08-19; USD prices are converted with the app_config FX rate. Verify Google official pricing before publication.

observed_hvoyweb · Cache read

~¥0.54/ 1M tokens

Observed from the hvoyweb active catalog and LLM metadata on 2026-08-19; USD prices are converted with the app_config FX rate. Verify Google official pricing before publication.

Observed from the hvoyweb active catalog and LLM metadata on 2026-08-19; USD prices are converted with the app_config FX rate. Verify Google official pricing before publication.

Relay Comparison

Compare token, per-request, or per-second pricing by relay channel.

How should Gemini 3.7 Flash relay pricing be compared?

This Gemini 3.7 Flash pricing page compares official pricing with public prices from 63 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 09/20/2026, 15:34.

Data sources
Public price catalogs, official pricing records, and monitoring results.
Metric definitions
Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
Risk note
Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.