TechCompare LogoTechCompare

Qwen 3.8 27B API Pricing & Cost Calculator

Qwen 3.8 27B at $0.575/$3.45 is the cheapest model in the 2026 Qwen flagship family with a real 1M context. For utility pipelines it's the pick over Qwen 3.8 Max when the reasoning ceiling doesn't matter.

Qwen 3.8 27B is the open-weights mid-size option in the Qwen 3.8 family, hosted by Alibaba Cloud International at $0.575/$3.45 with a 1M context window - and downloadable for self-hosting under open terms.

By TechCompare · Updated

Input tokens
30,000
per request
Output tokens
3,000
per request
Volume
1,000 / monthly
Standard API

Calculator

Cost Comparison

Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.

How this is calculated

Qwen 3.8 27B is priced at $0.575 per million input tokens and $3.45 per million output tokens on Alibaba Cloud International, the proprietary hoster, with an 80% prompt caching discount ($0.115 per million cached reads) and a 50% batch discount ($0.2875/$1.725). Third-party hosts on OpenRouter undercut that to about $0.45/$3.20, but with weaker latency and uptime - Alibaba's own endpoint runs about 41 tokens/sec at 99.92% uptime, roughly double the throughput of the cheaper resellers. As with the rest of Alibaba's catalog, cache and batch discounts cannot apply simultaneously.

Verdict

At $0.575/M input and $3.45/M output a 100K + 5K call costs about $0.075, under a third of Qwen 3.8 Max's $0.23 and roughly a tenth of GPT-5.6 Sol's $0.65. Cached input drops to $0.115/M, which drives stable-prompt utility workloads (classification, routing, formatting) almost to zero on the input side. Cheaper third-party OpenRouter hosts list the same weights around $0.45/$3.20, but they post 15-24 tokens/sec and 97-99% uptime against Alibaba's 41 tokens/sec and 99.92%, so the 20-28% price cut buys a measurable reliability drop. Self-hosting is the other alternative: the weights run on a 48 GB GPU at about 45 GB of VRAM for 128K context.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

Is Qwen 3.8 27B open source?
Yes, the weights are downloadable and you can self-host, which is unusual for a model with a 1M native context. The hosted Alibaba Cloud International endpoint is $0.575/$3.45 per million tokens if you'd rather not run 45 GB of VRAM yourself, and cheaper third-party OpenRouter hosts list it around $0.45/$3.20 with weaker uptime and throughput.
Qwen 3.8 27B API or self-host?
Self-host once your monthly spend crosses a modest GPU rental bill. At $0.575/$3.45 a month of 50M input and 5M output runs $46 on the API and a single 48 GB card rents for roughly the same per day. For bursty or low-volume work the API wins. For steady 24/7 load with 128K-or-less context, self-hosting pays back fast.