Model
Claude Opus 4.8 API pricing
Compare the published Claude Opus 4.8 (claude-opus-4-8) list rate with corouter's outlet rate, and estimate a monthly token bill. Claude Opus 4.8 matches the Opus 5 list rate and window, without the Opus 5.5 cache-read cut. Send claude-opus-4-8 to keep an Opus 4.8 client unchanged.
Where this model fits
- Opus 4.8 clients you do not want to rename
- Reasoning at the same list rate as Opus 5
- A controlled comparison against claude-opus-5-5
Estimates only. Actual billing depends on your real token counts, cache hit rate, and the prices in effect at request time.
From the blog
- OpenAI vs Anthropic API pricingBoth vendors publish per-million-token rates, but they price context, output, and caching differently enough that the cheaper choice depends entirely on your workload shape.Read the post
- How to reduce your LLM API costsMost advice about cutting inference spend is vague. This is the concrete version: where the money goes, which five levers move it, and how much each one is worth on a real monthly volume.Read the post

