Third-party dispatches, not graded receipts
OpenAI slashes GPT-5.6 Luna price by 80%, making it cheaper than Gemini Flash-Lite
GPT-5.6 Sol was used to optimize inference kernels and load balancing, enabling a 20% cost reduction for Terra and an 80% cut for Luna. At $0.20 per million input tokens, Luna is now one-fifth the input price of Anthropic's cheapest model, Claude Haiku 4.5.
“OpenAI credit 5.6 Sol with enabling this: in How GPT‑5.6 fuses frontier intelligence with frontier efficiency they describe using 5.6 Sol to optimize load balancing, and more impressively to optimize inference”