TechCompare LogoTechCompare

Gemini 3.1 Pro API Pricing & Cost Calculator

Gemini 3.1 Pro provides top-tier intelligence at highly competitive prices. Its 2-million token context window is incredibly powerful when paired with 75% prompt caching.

Gemini 3.1 Pro is Google's flagship reasoning engine, offering an unmatched 2-million token context window.

By TechCompare · Updated

Input tokens
100,000
per request
Output tokens
5,000
per request
Volume
100 / monthly
Standard API

Calculator

Cost Comparison

Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.

How this is calculated

Gemini 3.1 Pro (<=200k) is priced at $2.00 per million input tokens and $12.00 per million output tokens, supporting a 75% caching discount ($0.50 per million) and a 50% batch discount ($1.00 per million inputs, $6.00 per million outputs).

Verdict

At $2/M input and $12/M output, with cached input at $0.50/M and batch at $1.00/$6.00, Gemini 3.1 Pro sits below the Anthropic flagship pricing on every row. The standout is the 2M-token context window, which is what makes the 75% cache discount worth chasing: drop a large persistent context block into the prompt (a corpus, a codebase, a long doc set) and the cached input rate turns an otherwise huge per-call cost into a small fixed fee. The 25% cache discount gap against Anthropic's 90% is the trade-off, but the dollar math still favors Gemini because the base rates are lower.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

Does Gemini 3.1 Pro support prompt caching?
Yes, Google offers a 75% discount on cached input tokens, reducing the input price to $0.50 per million for matches.
How much does Gemini 3.1 Pro cost per million output tokens?
Gemini 3.1 Pro (<=200k) charges $12.00 per million output tokens at standard pricing. With the 50% batch discount, that drops to $6.00 per million. The 75% prompt caching discount brings cached input to $0.50 per million. Above 200K tokens of context, output rises to $18/M and input to $4/M.
Is Gemini 3.1 Pro cheaper than Claude Sonnet 4.6?
Yes, comfortably. At $2/M input and $12/M output versus Sonnet 4.6's $3/$15, Gemini wins on every base row. The 75% cache discount leaves cached input at $0.50/M, slightly above Sonnet's $0.30/M (90% off $3), but Gemini's lower base rates keep standard calls cheaper. Gemini also has the 2M context window, which Sonnet 4.6's 500K can't match.