Google

Gemini 3.1 Pro Preview

Multimodal model refined for thinking, factual consistency, software engineering, and reliable tool use

Gemini 3.1 Pro Preview accepts text, image, video, audio, and PDF input and produces text output. It supports up to 1,048,576 input tokens and 65,536 output tokens, plus caching, code execution, function calling, Search and Maps grounding, structured output, thinking, and URL context for software engineering and multi-step agentic workflows.

1,048,576 input tokens1M
Released2026-02
Relays20 sites
Native multimodalLong contextSoftware engineeringAgentic workflowsTool useStructured output

Google Official Pricing

CNY
Updated: 2026-07-17T10:50:25.324+08:00Source

official-lte-200k · Input

~¥13.7/ 1M tokens

Official Standard pricing for prompts up to 200K tokens: $2 input, $0.20 context caching, and $12 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

official-lte-200k · Output

~¥82.2/ 1M tokens

Official Standard pricing for prompts up to 200K tokens: $2 input, $0.20 context caching, and $12 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

official-lte-200k · Cache read

~¥1.37/ 1M tokens

Official Standard pricing for prompts up to 200K tokens: $2 input, $0.20 context caching, and $12 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

official-gt-200k · Input

~¥27.4/ 1M tokens

Official Standard long-context pricing for prompts above 200K tokens: $4 input, $0.40 context caching, and $18 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

official-gt-200k · Output

~¥123.3/ 1M tokens

Official Standard long-context pricing for prompts above 200K tokens: $4 input, $0.40 context caching, and $18 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

official-gt-200k · Cache read

~¥2.74/ 1M tokens

Official Standard long-context pricing for prompts above 200K tokens: $4 input, $0.40 context caching, and $18 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

Official Standard pricing for prompts up to 200K tokens: $2 input, $0.20 context caching, and $12 output including thinking tokens per 1M tokens; cache storage is an additional $4.50 per 1M tokens per hour. Converted at about USD 1 = CNY 6.85.

Relay Comparison

Compare token, per-request, or per-second pricing by relay channel.

How should Gemini 3.1 Pro Preview relay pricing be compared?

This Gemini 3.1 Pro Preview pricing page compares official pricing with public prices from 20 listed AI gateways. Token prices are shown in CNY per 1M tokens, while per-request, per-second, and per-character rows use the unit shown in the table. Last updated: 08/06/2026, 23:13.

Data sources
Public price catalogs, official pricing records, and monitoring results.
Metric definitions
Uptime means successful probe response rate, fake-rate signals possible model mismatch or abnormal output risk, and latency is average API response time.
Risk note
Relay gateways are third-party services. Pricing, billing, privacy, and stability can change; start with a small top-up and verify reliability before continued use.