TechCompare LogoTechCompare

Qwen 3.8 27B API Pricing & Cost Calculator

Qwen 3.8 27B at $0.575/$3.45 is the cheapest model in the 2026 Qwen flagship family with a real 1M context. For utility pipelines it's the pick over Qwen 3.8 Max when the reasoning ceiling doesn't matter.

Qwen 3.8 27B is the open-weights mid-size option in the Qwen 3.8 family, hosted by Alibaba Cloud International at $0.575/$3.45 with a 1M context window - and downloadable for self-hosting under open terms.

By TechCompare · Updated

Input tokens
30,000
per request
Output tokens
3,000
per request
Volume
1,000 / monthly
Standard API

Calculator

Cost Comparison

Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.

How this is calculated

Qwen 3.8 27B is priced at $0.575 per million input tokens and $3.45 per million output tokens on Alibaba Cloud International, the proprietary hoster, with an 80% prompt caching discount ($0.115 per million cached reads) and a 50% batch discount ($0.2875/$1.725). Third-party hosts on OpenRouter undercut that to about $0.45/$3.20, but with weaker latency and uptime - Alibaba's own endpoint runs about 41 tokens/sec at 99.92% uptime, roughly double the throughput of the cheaper resellers. As with the rest of Alibaba's catalog, cache and batch discounts cannot apply simultaneously.

Verdict

At $0.575/M input and $3.45/M output a 100K + 5K call costs about $0.075, under a third of Qwen 3.8 Max's $0.23 and roughly a seventh of GPT-5.6 Sol's $0.50. Cached input drops to $0.115/M, which drives stable-prompt utility workloads (classification, routing, formatting) almost to zero on the input side. Cheaper third-party OpenRouter hosts list the same weights around $0.45/$3.20, but they post 15-24 tokens/sec and 97-99% uptime against Alibaba's 41 tokens/sec and 99.92%, so the 20-28% price cut buys a measurable reliability drop. Self-hosting is the other alternative: the weights run on a 48 GB GPU at about 45 GB of VRAM for 128K context.

More API Standalones scenarios

MiMo-V2.6-Pro Pricing
Cost calculator for this model
View details ➜
MiMo-V2.6-Flash Pricing
Cost calculator for this model
View details ➜
GLM-5.3-Flash Pricing
Cost calculator for this model
View details ➜

Frequently asked questions

Is Qwen 3.8 27B open source?
Yes, the weights are downloadable and you can self-host, which is unusual for a model with a 1M native context. The hosted Alibaba Cloud International endpoint is $0.575/$3.45 per million tokens if you'd rather not run 45 GB of VRAM yourself, and cheaper third-party OpenRouter hosts list it around $0.45/$3.20 with weaker uptime and throughput.
Qwen 3.8 27B API or self-host?
Self-host once your monthly spend crosses a modest GPU rental bill. At $0.575/$3.45 a month of 50M input and 5M output runs $46 on the API and a single 48 GB card rents for roughly the same per day. For bursty or low-volume work the API wins. For steady 24/7 load with 128K-or-less context, self-hosting pays back fast.