TLDRocket
22 September 2026
The big throughline today is cost control—AI getting cheaper to run without loosening the safety ties that companies use to keep customers from panicking. OpenAI rolled out GPT-6 Sol and GPT-6 Luna, cutting token prices by half to $2 per million input tokens for Sol (and $0.10 per million output), while pairing the move with better prompt caching. The practical promise is visible in the plumbing: GPT-6’s caching improvements aim for higher hit rates and lower prompt latency, with diagnostics, breakpoints, and controls that make repeated work less expensive and more observable. Anthropic’s answer, Claude Opus 5.5, leans into the same theme—about 40% lower running cost on typical default workloads—and even tweaks model behavior and billing, with API pricing at $4/$20 per million input/output tokens.
Read the full briefing →