LLMCostLab

PricesCompare → GPT-5.6 Sol vs GPT-5.6 Luna

GPT-5.6 Sol vs GPT-5.6 Luna: API cost compared

GPT-5.6 Sol lists at $2.500 in / $15.00 out per 1M tokens. GPT-5.6 Luna lists at $0.100 in / $0.600. Sticker price only tells you so much — below is what each actually costs across eight production workloads.

advertisement
Cheaper on more workloads
GPT-5.6 Luna
wins 8 of 8
Blended rate gap
25.0x
at 3:1 in:out

Side by side

GPT-5.6 SolGPT-5.6 Luna
ProviderOpenAIOpenAI
Input per 1M$2.500$0.100
Output per 1M$15.00$0.600
Cached input per 1M$0.250$0.010
Context window400k400k
Blended per 1M (3:1)$5.625$0.225

Monthly cost by workload

Caching applied at each workload's typical hit rate. This is where the two models actually separate.

WorkloadGPT-5.6 SolGPT-5.6 LunaCheaperDifference
Customer support chatbot$366.00$14.64GPT-5.6 Luna$351.36 (96%)
RAG document Q&A$490.00$19.60GPT-5.6 Luna$470.40 (96%)
Coding agent$261.25$10.45GPT-5.6 Luna$250.80 (96%)
Bulk summarization$3,465$138.60GPT-5.6 Luna$3,326 (96%)
Structured data extraction$4,675$187.00GPT-5.6 Luna$4,488 (96%)
Long-form content generation$636.50$25.46GPT-5.6 Luna$611.04 (96%)
Email triage agent$1,081$43.25GPT-5.6 Luna$1,038 (96%)
Translation pipeline$4,112$164.47GPT-5.6 Luna$3,947 (96%)
advertisement

How to read this

GPT-5.6 Luna is cheaper on the majority of these workloads, but "majority" is not the decision. The output:input price ratio is 6.0 for GPT-5.6 Sol and 6.0 for GPT-5.6 Luna. If your traffic is output-heavy — long-form drafting, translation — the model with the lower output price wins regardless of what the input price suggests. If you are input-heavy — RAG, bulk summarization, agents re-reading a codebase — the cached input rate matters more than either headline number.

Cost is also not the only axis. This site does not benchmark quality, latency, or rate limits, and a model that is 3x cheaper but needs two attempts per task is not cheaper. Use these numbers to size a bill, then validate on your own evals.

Run your own numbers

Cost calculator

Enter your own numbers, or start from a workload preset. Costs are monthly and update as you type.

Cheapest model
Cheapest → priciest spread
ModelIn /1MOut /1MMonthlyvs best