Qwen 3.8 27B API Pricing & Cost Calculator
Qwen 3.8 27B at $0.575/$3.45 is the cheapest model in the 2026 Qwen flagship family with a real 1M context. For utility pipelines it's the pick over Qwen 3.8 Max when the reasoning ceiling doesn't matter.
Qwen 3.8 27B is the open-weights mid-size option in the Qwen 3.8 family, hosted by Alibaba Cloud International at $0.575/$3.45 with a 1M context window - and downloadable for self-hosting under open terms.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.
How this is calculated
Qwen 3.8 27B is priced at $0.575 per million input tokens and $3.45 per million output tokens on Alibaba Cloud International, the proprietary hoster, with an 80% prompt caching discount ($0.115 per million cached reads) and a 50% batch discount ($0.2875/$1.725). Third-party hosts on OpenRouter undercut that to about $0.45/$3.20, but with weaker latency and uptime - Alibaba's own endpoint runs about 41 tokens/sec at 99.92% uptime, roughly double the throughput of the cheaper resellers. As with the rest of Alibaba's catalog, cache and batch discounts cannot apply simultaneously.
Verdict
At $0.575/M input and $3.45/M output a 100K + 5K call costs about $0.075, under a third of Qwen 3.8 Max's $0.23 and roughly a tenth of GPT-5.6 Sol's $0.65. Cached input drops to $0.115/M, which drives stable-prompt utility workloads (classification, routing, formatting) almost to zero on the input side. Cheaper third-party OpenRouter hosts list the same weights around $0.45/$3.20, but they post 15-24 tokens/sec and 97-99% uptime against Alibaba's 41 tokens/sec and 99.92%, so the 20-28% price cut buys a measurable reliability drop. Self-hosting is the other alternative: the weights run on a 48 GB GPU at about 45 GB of VRAM for 128K context.
More API Standalones scenarios
Related guides
Frequently asked questions
Is Qwen 3.8 27B open source?
Qwen 3.8 27B API or self-host?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with syntax highlighting.
Use tool ➜