GPT-6 Luna API pricing and cost calculator
Luna's 20K-input, 2K-output example costs $0.003 uncached or $0.0021 with half the input cached. At 100,000 monthly requests, those token totals are $300 and $210.
GPT-6 Luna's standard base rates are $0.10 per million input tokens and $0.50 per million output tokens, with cache reads at $0.01 per million. Those prices make it worth evaluating for repeated classification, extraction, and routing work. At high request volumes, small per-call differences add up. The relevant questions are whether it completes the task reliably and how much spending remains once you include volume, retries, tools, and cache creation.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 20,000 input tokens (50% cached), 2,000 output tokens, and 100,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.
How this is calculated
This calculator starts with 20,000 input tokens, 2,000 output tokens, half the input read from cache, and 100,000 monthly requests. Uncached input costs $0.001, cache-read input costs $0.0001, and output costs $0.001, giving $0.0021 per request or $210 per month. Without cache hits, the same workload is $0.003 per call or $300 monthly. Above 272,000 input tokens, Luna applies double input and cache prices and 1.5 times output prices to the whole request. At 300,000 input and 5,000 output tokens, that is $0.06375 without cache hits or $0.03675 with half the input cached. Cache writes and tools are separate. The embedded chart loads live OpenRouter provider prices, while these examples use OpenAI's standard rates.
Verdict
Evaluate Luna directly on the small tasks that dominate your application's request count. Keep outputs short when the task only needs a label or a few extracted fields, and track failures rather than assuming cheap calls make retries harmless. GPT-6.1 Sol is a useful next comparison when you need to decide which tasks deserve a larger reasoning budget. A standalone Luna estimate is already useful for seeing whether volume or tools dominate the bill.
More API Standalones scenarios
Related guides
Frequently asked questions
What would one million example Luna requests cost?
Does a cheap model mean the whole agent is cheap?
How much does Luna prompt caching save?
Can GPT-6 Luna use large prompts at its base price?
What does Luna cost through the Batch API?
What should GPT-6 Luna be compared with?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with readable output and error detection.
Use tool ➜