TechCompare LogoTechCompare

Gemini 3.7 Flash API Pricing & Cost Calculator

Gemini 3.7 Flash at $0.75/$3.75 promo is the cheapest big-lab frontier model available - half of GLM-5.3 on both rows. Lock in workloads before January 2027 when it doubles back to $1.50/$7.50.

Gemini 3.7 Flash is Google's default fast model for the back half of 2026, shipping with a promo rate of $0.75/$3.75 that cuts Gemini 3.6 Flash's card in half until January 2027.

By TechCompare · Updated

Input tokens
100,000
per request
Output tokens
5,000
per request
Volume
100 / monthly
Standard API

Calculator

Cost Comparison

Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.

How this is calculated

Gemini 3.7 Flash is priced at $0.75 per million input tokens and $3.75 per million output tokens under Google's promotion through December 31, 2026, then reverts to $1.50/$7.50 standard. Cached input is $0.075 per million (90% off, plus $0.50/M-hour storage) and the batch tier is 50% off both rows. The context window is 1M tokens.

Verdict

At the promo rate a 100K + 5K call costs $0.094, undercutting GLM-5.3 ($0.162) and Muse Spark 1.2 ($0.146) while keeping Google's strongest cache story (90% discount, cached input $0.075/M) and a 50% batch tier. The catch is the sunset: on January 1, 2027 the card doubles to $1.50/$7.50 and GLM-5.3 becomes the cheaper option again. For greenfield builds through the end of 2026 this is the price leader among the big-three labs.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

When does the Gemini 3.7 Flash promo price end?
December 31, 2026. Google lists $0.75/M input and $3.75/M output as a promotion. From January 1, 2027 the standard rate of $1.50/M input and $7.50/M output applies. Cache (90% off) and batch (50% off) percentages stay the same.
How much does Gemini 3.7 Flash cost per million output tokens?
$3.75 per million through the end of 2026 promo, $7.50 per million after. Batch mode halves both to $1.875/M and $3.75/M. Cached input lands at $0.075 per million under the 90% discount, plus $0.50 per million tokens per hour for cache storage on persistent contexts.
Gemini 3.7 Flash or GLM-5.3 on price?
3.7 Flash wins through December 2026 at $0.75/$3.75 vs GLM-5.3's $1.40/$4.40. After the promo sunsets, GLM-5.3 flips cheaper. Flash also has the stronger cache discount (90% vs 81%) and a published 50% batch tier GLM lacks, so cache-heavy and batch workloads favor Flash year-round.