TLDRocket
Sign in

Safety & Ethics

429 summarised stories in Safety & Ethics, each linking back to the original source. Browse all topics →

Wednesday, 10 December 2025

A differentially private framework for gaining insights into AI chatbot use

Google Research 7 months ago

Researchers introduced Urania, a differentially private framework for analyzing LLM chatbot conversations that provides formal privacy guarantees instead of heuristic protections. The framework uses DP clustering and keyword extraction to ensure no single conversation overly influences the output, with empirical evaluation showing membership inference attacks achieved 0.53 AUC (near random guessing) against the private pipeline versus 0.58 against non-private baselines. This approach enables platforms to gain insights into how users interact with AI systems while mathematically guaranteeing that sensitive information cannot be revealed, even through prompt injection attacks or imperfect redaction.

Strengthening cyber resilience as AI capabilities advance

OpenAI Blog 7 months ago

OpenAI is developing additional safeguards and defensive capabilities to address cybersecurity risks as its AI models grow more capable. The company is implementing risk assessment frameworks and misuse-limiting measures, though specific metrics or implementation timelines are not disclosed. These efforts aim to enable the security community to better protect against potential threats from advanced AI systems.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.