Grok 4.6 API Pricing & Cost Calculator
Grok 4.6 at $2/$6 is priced like Qwen 3.8 Max, but xAI's 200K length tier is the trap: one long call bills entirely at $4/$12. Keep prompts under 200K and it's competitive with the Chinese flagships.
Grok 4.6 is xAI's current flagship API, keeping Grok 4.5's $2/$6 token rate but introducing a prompt-length tier: everything doubles once a request crosses 200K tokens, including cache reads.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.
How this is calculated
Grok 4.6 is priced at $2.00 per million input tokens and $6.00 per million output tokens for prompts under 200K tokens, and $4.00/$12.00 for prompts at or above 200K. Cache reads are $0.50/M under 200K and $1.00/M above. There is no published batch discount, and the full context window is 500K.
Verdict
A 100K + 5K call costs $0.23, matching Qwen 3.8 Max exactly at the low tier. Cross the 200K prompt threshold and the whole request bills at $4/$12, so a 250K + 5K call runs $1.06 versus $0.53 on a tierless $2/$6 model. Cache reads at $0.50/M (under 200K) are weaker than Alibaba's $0.20 and DeepSeek's $0.044, so stable-prompt loops cost more on xAI than any of those. The 500K window is the real differentiator if your workload sits between 200K and 1M.
More API Standalones scenarios
Related guides
Frequently asked questions
How does the 200K threshold billing work on Grok 4.6?
How does Grok 4.6 compare to Qwen 3.8 Max?
How much does Grok 4.6 cost per million output tokens?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with syntax highlighting.
Use tool ➜