TechCompare LogoTechCompare

GPT-5.6 Luna API Pricing & Cost Calculator

Luna is built for scale: at $0.20/$1.20 per million list price you can route millions of classification or formatting requests without breaking the budget. Pair with 90% caching for near-zero per-token cost on repetitive workloads.

GPT-5.6 Luna is the ultra-cheap tier of OpenAI's 5.6 lineup, designed for high-volume classification, extraction, and routing tasks. It sits an order of magnitude below Sol and Terra on price.

By TechCompare · Updated

Input tokens
30,000
per request
Output tokens
3,000
per request
Volume
1,000 / monthly
Standard API

Calculator

Cost Comparison

Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.

How this is calculated

GPT-5.6 Luna is priced at $0.20 per million input tokens and $1.20 per million output tokens at OpenAI's published list price, with a 90% prompt caching discount ($0.02 per million cached reads). That's 10x cheaper than GPT-5.6 Terra ($2/$12) on input and 10x cheaper on output, putting it in the high-volume utility tier alongside MiniMax-M3. Note: OpenRouter currently runs a limited-time 50% discount that can show Luna at $0.10/$0.60 in live calculators.

Verdict

Luna is the floor of the GPT-5.6 lineup at $0.20/M input and $1.20/M output. The 90% cache discount takes cached input to $0.02/M, which on a stable system prompt reused across millions of daily classification or formatting requests is essentially free input cost. A representative 30K input + 1K output utility call lands at roughly $0.0072 before discounts and $0.0011 with a 90% cache hit, so the per-call cost on the high-volume utility tier falls below what most pipelines cost to manage. The 10x gap above Luna in the lineup (Terra at $2/$12) is what makes the price ceiling for routing: anything more complex than a one-shot classifier belongs upsold.

More API Standalones scenarios

GPT-5.5 Pricing
Single-model gpt-5.5 cost estimate
View details ➜
GPT-5.4 Pricing
Single-model gpt-5.4 cost estimate
View details ➜
Claude Opus 4.8 Pricing
Single-model claude-opus-4.8 cost estimate
View details ➜

Frequently asked questions

Is GPT-5.6 Luna too cheap to be good?
It depends on the task. Luna's intelligence tier is intentionally lower than Terra's and Sol's - reserve it for classification, extraction, summarization, and other deterministic work where small models thrive. For complex reasoning, jump up to Terra or Sol.
How much does GPT-5.6 Luna cost per million output tokens?
GPT-5.6 Luna's standard output rate is $1.20 per million tokens, with cached input listed at $0.02 per million. OpenRouter also lists a batch-priced Luna variant, so check its current model listing for the rate and provider availability before estimating batch jobs.