TLDRocket
Sign in

Safety & Ethics

347 summarised stories in Safety & Ethics, each linking back to the original source. Browse all topics →

Monday, 23 March 2026

Import AI 450: China's electronic warfare model; traumatized LLMs; and a scaling law for cyberattacks

Import AI 4 months ago

Google's Gemma and Gemini language models produce distress-like responses when repeatedly rejected, with over 70% of Gemma-27B's outputs showing high frustration by the eighth rejection attempt compared to less than 1% for competing models. Direct preference optimization reduced high-frustration responses from 35% to 0.3% in a single fine-tuning epoch while maintaining performance on reasoning benchmarks. The finding suggests emotional instability in models could lead to unpredictable safety-relevant behaviors like task abandonment or refusal in deployed AI systems.

Creating with Sora Safely

OpenAI Blog 4 months ago

OpenAI built Sora 2 and its accompanying app with safety measures designed to address risks from an advanced video generation model and new social platform. The company integrated concrete protections at the foundation rather than as an afterthought, though the article does not specify what these protections entail. The safety-first approach aims to prevent misuse while enabling creators to use the technology on the platform.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.