Qwen3.7 Max API Pricing & Cost Calculator

Qwen3.7 Max is the value-pick for Qwen-lineage buyers. With a real 1M context window and mid-tier pricing, it competes head-on with GLM-5.2 and Gemini Flash for high-volume document work.

Qwen3.7 Max is Alibaba's flagship reasoning model under the Qwen line. It sits in the mid-tier workhorse pricing band alongside GLM-5.2 and Gemini 3.6 Flash.

By TechCompare · Updated

Input tokens
100,000
50% cached
Output tokens
5,000
per request
Volume
100 / monthly
Standard API

Calculator

Cost Comparison

Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.

How this is calculated

Qwen3.7 Max is priced at $1.475 per million input tokens and $4.425 per million output tokens, with an 80% prompt caching discount ($0.295 per million cached reads). That puts it slightly costlier than GLM-5.2 ($1.40/$4.40) but cheaper than Gemini 3.6 Flash on output.

Verdict

Qwen3.7 Max is the value-pick for Qwen-lineage buyers. With a real 1M context window and mid-tier pricing, it competes head-on with GLM-5.2 and Gemini Flash for high-volume document work.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

How does Qwen3.7 Max compare to GLM-5.2?
GLM-5.2 wins on raw price ($1.40/$4.40 vs Qwen3.7 Max's $1.475/$4.425) and cache depth (~81% vs 80% discount). Qwen3.7 Max is the pick if you're already in the Qwen ecosystem or prefer Alibaba's tooling.