Gemini 3.7 Flash API Pricing & Cost Calculator
Gemini 3.7 Flash at $0.75/$3.75 promo is the cheapest big-lab frontier model available - half of GLM-5.3 on both rows. Lock in workloads before January 2027 when it doubles back to $1.50/$7.50.
Gemini 3.7 Flash is Google's default fast model for the back half of 2026, shipping with a promo rate of $0.75/$3.75 that cuts Gemini 3.6 Flash's card in half until January 2027.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.
How this is calculated
Gemini 3.7 Flash is priced at $0.75 per million input tokens and $3.75 per million output tokens under Google's promotion through December 31, 2026, then reverts to $1.50/$7.50 standard. Cached input is $0.075 per million (90% off, plus $0.50/M-hour storage) and the batch tier is 50% off both rows. The context window is 1M tokens.
Verdict
At the promo rate a 100K + 5K call costs $0.094, undercutting GLM-5.3 ($0.162) and Muse Spark 1.2 ($0.146) while keeping Google's strongest cache story (90% discount, cached input $0.075/M) and a 50% batch tier. The catch is the sunset: on January 1, 2027 the card doubles to $1.50/$7.50 and GLM-5.3 becomes the cheaper option again. For greenfield builds through the end of 2026 this is the price leader among the big-three labs.
More API Standalones scenarios
Related guides
Frequently asked questions
When does the Gemini 3.7 Flash promo price end?
How much does Gemini 3.7 Flash cost per million output tokens?
Gemini 3.7 Flash or GLM-5.3 on price?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with syntax highlighting.
Use tool ➜