Model
GPT-6 Luna API pricing
Compare the published GPT-6 Luna (gpt-6-luna) list rate with corouter's outlet rate, and estimate a monthly token bill. GPT-6 Luna keeps the 1.05M window and lists above GPT-6 Sol on output, without the Code tag. Send gpt-6-luna when you want that ID rather than gpt-6-sol or gpt-6.1-sol.
Where this model fits
- Long-context chat that should bill the Luna output rate
- Vision passes that do not need a Code tag
- Clients that name gpt-6-luna rather than gpt-6-sol
Estimates only. Actual billing depends on your real token counts, cache hit rate, and the prices in effect at request time.
From the blog
- OpenAI vs Anthropic API pricingBoth vendors publish per-million-token rates, but they price context, output, and caching differently enough that the cheaper choice depends entirely on your workload shape.Read the post
- How to reduce your LLM API costsMost advice about cutting inference spend is vague. This is the concrete version: where the money goes, which five levers move it, and how much each one is worth on a real monthly volume.Read the post

