TLDRocket
Sign in
Latest Meet the 82-year-old Kentucky grandma who turned down $26 million to t... — Fortune How Chevron became the AI darling of Big Oil — Fortune Mistral AI Releases Shieldstral 1.0 3B: An Open-Weights Policy-Adaptiv... — MarkTechPost [AINews] Zawinski's Law of MultiAgents — Latent Space Now we have a timeline of the OpenAI accidental attack against Hugging... — Simon Willison’s Weblog OpenAI says it slowed Astra model development over security concerns — TechCrunch Tencent Cloud Open-Sources TencentDB Agent Memory v2.0: A Team-Level M... — MarkTechPost Auto Mode will soon be the default in Claude Code — because humans can... — The New Stack

The AI intelligence platform

Every AI story that matters and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Thursday, 23 October 2025

Strengthening our Frontier Safety Framework

Google DeepMind 9 months ago 37

An AI company published the third version of its Frontier Safety Framework, expanding risk assessment categories and refining protocols for identifying severe hazards from advanced models. The update introduces a new Critical Capability Level focused on harmful manipulation and adds Tracked Capability Levels as of April 17, 2026 to catch less extreme risks earlier. The framework now requires safety case reviews before external launches and expands such reviews to large-scale internal deployments of models that could accelerate AI development.

Gemini Robotics 1.5 brings AI agents into the physical world

Google DeepMind 9 months ago 48 3 sources

Google released Gemini Robotics 1.5 and Gemini Robotics-ER 1.5, two models that enable robots to reason, plan, and execute complex multi-step physical tasks in the real world. Gemini Robotics-ER 1.5 achieved state-of-the-art performance across 15 academic embodied reasoning benchmarks, including Point-Bench and ERQA. The models work together as an agentic system where the reasoning model plans high-level tasks and the vision-language-action model executes them, while also learning to transfer skills across different robot embodiments without specialization.

Introducing CodeMender: an AI agent for code security

Google DeepMind 9 months ago 50

Researchers introduced CodeMender, an AI agent that automatically detects and fixes software vulnerabilities in code. Over six months, CodeMender generated 72 security patches for open-source projects, including some with 4.5 million lines of code, using reasoning capabilities to identify root causes and validate changes before human review. The tool enables developers to focus on building software while the system handles security patching and proactively rewrites existing code to use more secure APIs and data structures.

Bringing AI to the next generation of fusion energy

Google DeepMind 9 months ago 44

Google DeepMind is partnering with Commonwealth Fusion Systems to use artificial intelligence for controlling plasma in fusion reactors, focusing on optimizing the SPARC tokamak machine designed to achieve net energy gain. The collaboration centers on three areas: developing TORAX, a fast plasma simulator built in JAX; using reinforcement learning to find optimal operating parameters; and creating AI-based real-time control systems to manage heat distribution and plasma stability. The partnership aims to accelerate the timeline toward delivering fusion energy to the grid by running millions of virtual experiments before SPARC operation begins.

Try Deep Think in the Gemini app

Google DeepMind 9 months ago 40 3 sources

Google released Deep Think in the Gemini app for AI Ultra subscribers, a reasoning model variant that achieved gold-medal performance at the International Mathematical Olympiad. The version available to subscribers reaches Bronze-level performance on the 2025 IMO benchmark while operating faster than the competition model, with users limited to a fixed number of prompts per day. The tool is designed to help with complex problem-solving tasks including web development, scientific research, and coding challenges through extended reasoning time.

Rethinking how we measure AI intelligence

Google DeepMind 9 months ago 5

Google DeepMind and Kaggle launched Game Arena, an open-source benchmarking platform where AI models compete in strategic games to evaluate their capabilities. The platform will host its inaugural chess tournament on August 5 with eight frontier models competing, with final rankings determined by over one hundred matches between each model pair. Game Arena addresses limitations of current benchmarks by using games with clear winning conditions to measure reasoning and planning rather than memorization, with plans to expand to Go, poker, and video games.

Introducing Gemma 3 270M: The compact model for hyper-efficient AI

Google DeepMind 9 months ago 5 4 sources

Google released Gemma 3 270M, a 270-million-parameter compact model designed for task-specific fine-tuning on edge devices and resource-constrained systems. Internal testing on a Pixel 9 Pro showed the INT4-quantized model consumed just 0.75% battery for 25 conversations, making it Google's most power-efficient Gemma model. The release enables developers to build specialized, production-ready AI systems for specific tasks like text classification and data extraction without requiring large general-purpose models.

Image editing in Gemini just got a major upgrade

Google DeepMind 9 months ago 13

Google integrated a new image editing model called Nano Banana into the Gemini app, enabling users to edit photos while maintaining consistent appearance of people and pets across multiple versions. The model is described as the top-rated image editing model available, with capabilities including multi-turn editing, photo blending, and style transfer. Users can now apply costume changes, background modifications, and complex edits while preserving specific elements of their photos, with all generated images marked by visible and invisible watermarks.

VaultGemma: The world's most capable differentially private LLM

Google DeepMind 9 months ago 17

Google DeepMind released VaultGemma, a 1-billion-parameter language model trained with differential privacy, accompanied by new scaling laws that describe how privacy, compute, and data budgets interact during training. The model was trained with a sequence-level privacy guarantee of (ε ≤ 2.0, δ ≤ 1.1e-10) and showed no detectable memorization of training data in empirical tests. VaultGemma performs comparably to non-private models from approximately five years ago, establishing a baseline for measuring progress as privacy-preserving training methods improve.

Introducing the Gemini 2.5 Computer Use model

Google DeepMind 9 months ago 4 3 sources

Google released Gemini 2.5 Computer Use, a specialized AI model that can interact with graphical user interfaces by clicking, typing, and scrolling through web pages and applications like humans do. The model outperforms leading alternatives on multiple web and mobile control benchmarks while maintaining lower latency. Developers can now build AI agents that automate tasks requiring direct UI interaction, such as form completion and workflow automation, through the Gemini API.

OpenAI acquires Software Applications Incorporated, maker of Sky

OpenAI 9 months ago 14

OpenAI has acquired Software Applications Incorporated, the developer of Sky, a natural language interface for Mac. Sky will be integrated into ChatGPT to enable users to perform actions directly on their desktop through natural language commands. The acquisition allows ChatGPT to access deeper macOS capabilities and provide more contextual, action-oriented responses on Apple computers.

Consensus accelerates research with GPT-5 and Responses API

OpenAI 9 months ago 13

Consensus has integrated GPT-5 and OpenAI's Responses API into its research assistant tool to enable automated reading and synthesis of scientific evidence. The platform serves over 8 million researchers who can now complete evidence analysis tasks in minutes rather than hours or days. This reduces the time researchers spend on literature review and synthesis, allowing them to focus more quickly on experimental design or manuscript preparation.

Google Earth AI: Unlocking geospatial insights with foundation models and cross-modal reasoning

Google Research 9 months ago 44 2 sources

Google introduced Earth AI, which combines foundation models for satellite imagery, population data, and environmental forecasting with a Gemini-powered reasoning agent to answer complex geospatial questions. The Geospatial Reasoning Agent achieved 0.82 accuracy on a Q&A benchmark compared to 0.50 for Gemini 2.5 Pro, and combining population and landscape embeddings improved FEMA's National Risk Index predictions by an average of 11% R² across 20 hazards. Organizations including the UN, GiveDirectly, and insurance companies are using these capabilities to predict natural disasters, assess damage, and identify vulnerable communities for aid distribution.

Work smarter with your company knowledge in ChatGPT

OpenAI 9 months ago 31

OpenAI has added a company knowledge feature to ChatGPT that pulls information from business applications to provide organization-specific answers with citations. The feature is available immediately to Business, Enterprise, and Edu plan subscribers. Users can now access work-relevant data within ChatGPT while maintaining security controls and administrative oversight of what information gets shared.

AI in South Korea—OpenAI’s Economic Blueprint

OpenAI 9 months ago 23 2 sources

OpenAI published an economic blueprint for South Korea describing how the country can develop trusted AI systems through domestic capabilities and strategic partnerships. The proposal emphasizes building sovereign AI infrastructure while maintaining international collaboration on technology development. South Korea would use this approach to position itself as a regional AI hub and generate economic growth through domestic innovation.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.