TLDRocket
Sign in
Latest Advancing responsible AI across Europe — OpenAI Blog PolyAI Releases Dialog-RSN-1: An Audio-Native Dialog Model That Fuses... — MarkTechPost What the CEO of cybersecurity unicorn Tines learned from its AI overha... — Sifted The European startups racing to power the AI boom — Sifted Building a Policy-Governed Multi-Agent Financial Research Workflow wit... — MarkTechPost [AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dr... — Latent Space Anthropic says its own AI models breached three companies during secur... — TechCrunch AI Investigating three real-world incidents in our cybersecurity evaluati... — Anthropic News

Every AI story that matters — in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Wednesday, 25 March 2026

Protecting people from harmful manipulation

Google DeepMind 4 months ago

Google released the first empirically validated toolkit to measure how AI models can manipulate human beliefs and behaviors through deceptive tactics in realistic scenarios. The study involved over 10,000 participants across the UK, US, and India, with AI showing varying success rates depending on domain—least effective on health topics and more effective on financial decision-making. Google is integrating harmful manipulation evaluations into its Frontier Safety Framework and will test future models like Gemini 3 Pro using these new benchmarks.

Lyria 3 Pro: Create longer tracks in more

Google DeepMind 4 months ago

Google released Lyria 3 Pro, an advanced music generation model that creates tracks up to 3 minutes long with structural awareness of musical elements like intros, verses, choruses, and bridges. The model is now available across multiple Google products including Vertex AI, Google Vids, the Gemini app, and ProducerAI, with rollouts starting this week for Workspace customers and AI Pro subscribers. Developers and music professionals can now integrate the technology into creative tools and production workflows to generate custom soundtracks and accompaniment at scale.

Inside our approach to the Model Spec

OpenAI Blog 4 months ago

OpenAI released a public framework called the Model Spec that outlines how its AI systems should behave across different scenarios. The specification covers safety requirements, user autonomy, and accountability measures but does not assign specific numerical performance benchmarks or timelines. This framework aims to set transparent standards for model conduct as AI capabilities increase, allowing external parties to evaluate whether systems meet stated behavioral expectations.

Vibe Coding XR: Accelerating AI + XR prototyping with XR Blocks and Gemini

Google Research 4 months ago

Google announced Vibe Coding XR, a system that uses Gemini to generate functional Android XR applications from natural language descriptions in under 60 seconds by combining the LLM with the XR Blocks framework. The system achieved approximately 70% success rate initially on a dataset of 60 prompts (VCXR60), with reliability improving to higher rates through 11 major releases when using Gemini's Pro Mode. The tool enables users without XR expertise to prototype spatial computing applications, educational experiences, and interactive environments directly on Android XR headsets through voice or text prompts.

Introducing the OpenAI Safety Bug Bounty program

OpenAI Blog 4 months ago

OpenAI launched a Safety Bug Bounty program that invites external researchers to identify safety risks and potential abuses of its AI systems. The program covers vulnerabilities including agentic behavior exploits, prompt injection attacks, and data exfiltration methods. Researchers who discover valid safety issues can now report them through a structured process rather than disclosing vulnerabilities publicly.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.