Use case
GPT-6 Luna streaming
streaming notes for gpt-6-luna: the rate card, a worked volume, and when to send this ID. A sample of 1 million input tokens plus 200 thousand output tokens is about $0.440 on corouter (input $0.200, output $0.240).
How streaming is billed
A streamed response is still billed on input and output tokens, not on how long the socket stays open. Output for gpt-6-luna is $1.20 per million tokens.
A sample of 1 million input tokens plus 200 thousand output tokens is about $0.440 on corouter (input $0.200, output $0.240).
A field you can set
A familiar chat-completions shape often sets "stream": true next to "model": "gpt-6-luna". On a messages shape, follow the stream field your client actually sends.
GPT-6 Luna keeps the 1.05M window and lists above GPT-6 Sol on output, without the Code tag. Send gpt-6-luna when you want that ID rather than gpt-6-sol or gpt-6.1-sol.
Questions
- Which model ID should the request send?
- Send gpt-6-luna. Do not put a different name in the model field.
- What is the minimum top-up?
- Prepaid credits, $5 minimum top-up, no subscription tier.

