GPT-5.6 Luna API Pricing & Cost Calculator
Luna is built for scale: at $0.20/$1.20 per million list price you can route millions of classification or formatting requests without breaking the budget. Pair with 90% caching for near-zero per-token cost on repetitive workloads.
GPT-5.6 Luna is the ultra-cheap tier of OpenAI's 5.6 lineup, designed for high-volume classification, extraction, and routing tasks. It sits an order of magnitude below Sol and Terra on price.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.
How this is calculated
GPT-5.6 Luna is priced at $0.20 per million input tokens and $1.20 per million output tokens at OpenAI's published list price, with a 90% prompt caching discount ($0.02 per million cached reads). That's 10x cheaper than GPT-5.6 Terra ($2/$12) on input and 10x cheaper on output, putting it in the high-volume utility tier alongside MiniMax-M3. Note: OpenRouter currently runs a limited-time 50% discount that can show Luna at $0.10/$0.60 in live calculators.
Verdict
Luna is the floor of the GPT-5.6 lineup at $0.20/M input and $1.20/M output. The 90% cache discount takes cached input to $0.02/M, which on a stable system prompt reused across millions of daily classification or formatting requests is essentially free input cost. A representative 30K input + 1K output utility call lands at roughly $0.0072 before discounts and $0.0011 with a 90% cache hit, so the per-call cost on the high-volume utility tier falls below what most pipelines cost to manage. The 10x gap above Luna in the lineup (Terra at $2/$12) is what makes the price ceiling for routing: anything more complex than a one-shot classifier belongs upsold.
More API Standalones scenarios
Related guides
Frequently asked questions
Is GPT-5.6 Luna too cheap to be good?
How much does GPT-5.6 Luna cost per million output tokens?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with readable output and error detection.
Use tool ➜