LLMCostLab

Prices → GPT-5.6 Luna

GPT-5.6 Luna API pricing

GPT-5.6 Luna from OpenAI is priced at $0.100 per 1M input tokens and $0.600 per 1M output tokens, with a 400k-token context window. Output costs 6.0x what input costs, which is the number that decides whether this model is expensive for you.

Verified 2026-08-19
advertisement
Input /1M
$0.100
Output /1M
$0.600
Blended /1M (3:1)
$0.225

What GPT-5.6 Luna costs on real workloads

Monthly cost at each workload's default volume, shown with and without prompt caching. The right-hand column is what caching is worth on this specific model.

WorkloadCalls/moNo cacheWith cacheSaved
Customer support chatbot50,000$19.50$14.6425%
RAG document Q&A20,000$23.20$19.6016%
Coding agent4,000$17.20$10.4539%
Bulk summarization100,000$144.00$138.604%
Structured data extraction500,000$250.00$187.0025%
Long-form content generation10,000$26.00$25.462%
Email triage agent200,000$68.00$43.2536%
Translation pipeline150,000$166.50$164.471%

On caching

Cached input reads at $0.010 per 1M, a 90% discount on fresh input. On a workload that re-sends the same system prompt or the same documents every call, that discount is usually the single largest lever available — larger than switching models.

advertisement

Closest alternatives by price

The six models nearest GPT-5.6 Luna on blended rate — the realistic swap candidates.

ModelIn /1MOut /1MBlendedvs GPT-5.6 Luna
Gemini 2.5 Flash-Lite$0.100$0.400$0.175-22%compare
Gemini 3.1 Flash-Lite$0.250$1.500$0.562+150%
Gemini 3.5 Flash-Lite$0.300$2.500$0.850+278%
Gemini 2.5 Flash$0.300$2.500$0.850+278%
Gemini 3.7 Flash$0.750$3.750$1.500+567%compare
Gemini 3.6 Flash$0.750$3.750$1.500+567%

Price it against your own volume

Cost calculator

Enter your own numbers, or start from a workload preset. Costs are monthly and update as you type.

Cheapest model
Cheapest → priciest spread
ModelIn /1MOut /1MMonthlyvs best

Cutting this bill in practice

If a cheaper model on the table above would work for your task, the blocker is usually integration effort rather than the price difference.

Switch models without rewriting your integration

The savings on this page are only real if you can actually move traffic to the cheaper model. Aggregators and gateways sit in front of multiple providers so that switch is a configuration change.