[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization
Latent Space ● Covered by 12 sources
OpenAI cut GPT-5.6 model prices by 20–80% and introduced a faster inference tier, driven by self-optimization improvements in kernel rewriting, speculative decoding, and caching that reduced serving costs. GPT-5.4 intelligence now costs roughly one-thirteenth of its March price at the same performance level, representing an annualized rate of approximately 2,000x cost reduction. The price cuts shift downstream AI workflows toward cheaper models, with tools like ChatGPT auto-review and Codex moving from GPT-5.4 to Luna at roughly 10x lower cost.
Why it matters
Distillation is all you need!
Also covered by
- TLDR — OpenAI cuts prices for two of its GPT-5.6 AI models as companies grow sensitive to costs
- Simon Willison — Advancing the price-performance frontier with GPT‑5.6
- Simon Willison — llm 0.32rc2
- The New Stack — Chinese AI competitors may have forced OpenAI’s hand on pricing
- Simon Willison — llm 0.32rc1
- TLDR Dev — How GPT-5.6 fuses frontier intelligence with frontier efficiency
- TLDR — How ChatGPT Optimizes its Agent Loop: Harness, API, and Inference
- OpenAI Blog — Advancing the price-performance frontier with GPT-5.6
- The New Stack — OpenAI fixed GPT-5.6 Sol’s most frustrating flaw: Burning limits while it waits
- The New Stack — GPT-5.6 kernel of truth: Sol can cut its own costs, says OpenAI
- OpenAI Blog — How GPT-5.6 fuses frontier intelligence with frontier efficiency