TechCompare LogoTechCompare

Gemini 3.6 Flash API Pricing & Cost Calculator

Gemini 3.6 Flash is the multimodal value-pick: $1.50/$7.50 with full image, audio, and video input support makes it uniquely cost-effective for media-heavy pipelines. Combine with the batch API for half-off workloads.

Gemini 3.6 Flash succeeds Gemini 3.5 Flash as Google's fast multimodal model, with a 1-million-token context window and native image, audio, and video input. It targets the high-volume workhorse tier.

By TechCompare · Updated

Input tokens
30,000
per request
Output tokens
3,000
per request
Volume
1,000 / monthly
Standard API

Calculator

Cost Comparison

Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.

How this is calculated

Gemini 3.6 Flash is priced at $1.50 per million input tokens and $7.50 per million output tokens, with a 90% prompt caching discount ($0.15 per million cached reads) and a 50% batch discount ($0.75 per million inputs, $3.75 per million outputs) via Google's Batch API. That's a touch cheaper than Gemini 3.5 Flash on output.

Verdict

Gemini 3.6 Flash sits at $1.50/M input and $7.50/M output, with the 90% cache discount dropping cached input to $0.15/M and the 50% batch discount dropping both to $0.75/M input and $3.75/M output. The multimodal case is what makes the per-token value stack: image, audio, and video inputs route through a single endpoint at the same per-million-token rate, so a video or audio transcription pipeline that would otherwise need transcription-plus-classification models collapses into one bill. The deeper 90% cache discount eclipses Gemini 3.5 Flash on cached workloads, and the batch-API 50% halves both rows for offline media processing jobs.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

Is Gemini 3.6 Flash cheaper than Gemini 3.5 Flash?
Yes, slightly. 3.6 Flash drops output cost from $9.00 to $7.50 per million tokens while keeping input at $1.50, with the same 90% cache discount. A small but real improvement for high-output workloads.
How much does Gemini 3.6 Flash cost per million output tokens?
Gemini 3.6 Flash charges $7.50 per million output tokens, down from Gemini 3.5 Flash's $9.00 rate while input stays at $1.50 per million. The 90% prompt caching discount brings cached input to $0.15 per million based on the $1.50/M input price. The 50% batch discount further drops output to $3.75 per million for asynchronous bulk work.