TLDRocket
8 October 2026
AI’s attention today kept snapping back to the same uncomfortable question: when models spread—especially open weights—how do you keep them honest, safe, and accountable? Researchers accused OpenAI and Anthropic of using ChatGPT/Codex and Claude interactions to reproduce unpublished work, while separate scrutiny grew around open-weight cyber capability. Analyses of GLM-5.3’s ExploitBench performance put it at 12% success versus Claude Mythos at 14% at comparable token counts, and the cost curve kept sliding: $20.40 of GLM-5.3-Flash tokens could find a recently disclosed Chrome flaw. Xiaomi’s new MiMo-V2.6-Pro-RL and Flash variants, meanwhile, report strong CyberBench results and even lower token costs, making “just generate another payload” less theoretical.
Read the full briefing →