Claude Sonnet 5.5 API pricing and cost calculator
A Sonnet request with 100K input and 5K output costs $0.25 uncached or $0.16 with half the input cached. Prompt growth raises token volume without adding a long-context price tier.
Claude Sonnet 5.5 charges $2 per million input tokens and $10 per million output tokens at standard rates. Its $0.20 cache-read price reduces repeated-context spending, while standard token rates stay flat within its 1M context window. Those properties are useful when budgeting document processing or ongoing conversations. The bill still depends on how much new input and output each turn adds, so start with actual usage rather than a model's context capacity.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.
How this is calculated
At 100,000 input and 5,000 output tokens, Sonnet costs $0.20 for uncached input plus $0.05 for output, totaling $0.25. The starting calculator example reads half the input from cache, giving $0.10 uncached input, $0.01 cached input, and $0.05 output. That is $0.16 per request or $16 for 100 monthly requests. A 300,000-input, 5,000-output prompt costs $0.65 with no cache hits or $0.38 with half the input cached. A longer prompt doesn't introduce a separate token-rate tier within the advertised window. These examples exclude cache writes and tool charges. The embedded estimate uses OpenRouter's live provider rates, so check its routing assumptions alongside Anthropic's standard price list.
Verdict
Sonnet is worth testing when its results meet your quality targets and a predictable token rate helps you budget long prompts. Keep output concise when the task allows it, and compare retrieval with sending the full document on every turn. Prompt caching is valuable for reusable prefixes, but count the writes and misses too. If another model completes a task in fewer calls, its total session cost can matter more than a small difference in input rates.
More API Standalones scenarios
Related guides
Frequently asked questions
What is the cost of 10,000 Sonnet 5.5 requests?
How much does Sonnet 5.5 prompt caching cost?
Is Sonnet 5.5 cheaper than GPT-6.1 Sol?
Does Sonnet's 1M context window mean I should use it all?
What does Sonnet 5.5 cost in batch mode?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with readable output and error detection.
Use tool ➜