LWiAI Podcast #238 - GPT 5.4 mini, OpenAI Pivot, Mamba 3, Attention Residuals
Last Week in AI Last Week in AI ● Covered by 2 sources
This podcast episode covered major AI developments from March 2026 including OpenAI releasing GPT-5.4 mini and nano models with 400k-token context windows at higher per-token costs, Mistral open-sourcing its Small 4 model family with 119B total parameters, and multiple companies advancing agent systems including Meta's Manus launching a Mac agent and Nvidia announcing NeMo for sandboxed agent runtime. OpenAI shifted strategic focus toward enterprise productivity amid competition, Microsoft reorganized its AI division, Meta delayed its next model rollout, and new safety research addressed steganography detection, chain-of-thought faithfulness, fine-tuning defenses, and cyber-attack evaluations. The episode included technical discussions of Attention Residuals and Mamba-3 sequence modeling research alongside business updates about hardware forecasts and ByteDance accessing top Nvidia chips.
Why it matters
OpenAI ships GPT-5.4 mini and nano, faster and more capable but up to 4x pricier, DLSS 5 looks like a real-time generative AI filter for video games | The Verge, and more!