Simon Willison's Weblog·4 weeks ago·
28
● 14 sources
LLM 0.32rc2 release updated the default model from GPT-4o mini to GPT-5.6 Luna and added an openai endpoint command for running prompts against OpenAI-compatible endpoints without prior configuration. GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens, compared to GPT-4o mini's $0.15/$0.60. Users can now test prompts against local or remote compatible endpoints directly via CLI without installing LLM.
Google DeepMind released Gemini Robotics 2, a suite of three AI models for controlling robots across whole-body movement, multi-finger dexterity, and multi-robot collaboration. The system achieves success rates ranging from 32% to 92% depending on task complexity, with one model checkpoint controlling multiple robot embodiments including the Apptronik Apollo 2 and Franka Duo. Robots can now adapt to unpredictable environments, transfer skills between different bodies, and coordinate with other robots on shared tasks rather than relying on pre-programming or teleoperation.
Simon Willison's Weblog·1 month ago·
14
● 14 sources
LLM released version 0.32 release candidate, introducing a new database schema that uses content-addressable hash IDs for stored messages to enable deduplication and support for forked conversations. The update adds support for three new model variants: gpt-5.6-sol, gpt-5.6-terra, and gpt-5.6-luna. Existing data is preserved but users should back up their logs.db file before upgrading due to the significant schema changes.
OpenAI released GPT-5.6, a model family designed to balance performance with cost efficiency through load balancing and caching techniques that process more tokens effectively. The model includes a variant called GPT-5.6 Sol that optimizes its own code execution and resource allocation. This allows the system to handle increased workloads while reducing computational costs.
Google launched Gemini 3.5 Flash, a faster mid-tier model with improved agentic capabilities and visual understanding, but at three times the price of its predecessor. The model costs $1.50 per million input tokens (compared to $0.05 for Gemini 3 Flash) and topped benchmarks like APEX-Agents-AA with 47.1% accuracy. Higher per-token prices across major AI providers mean the "Flash" designation no longer signals a cost advantage for developers building agents. Meanwhile, the European Union delayed high-risk AI regulations to December 2027 and December 2026 for other provisions, citing competitiveness concerns, though it strengthened rules banning non-consensual intimate imagery.
Every AI story that matters,
in your inbox by 8am.
TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the
day in two minutes. Follow companies and topics for alerts, or get the
briefing in Slack. Free, no spam, unsubscribe anytime.
Reading TLDRocket needs no cookies, and the readership counts we rely on come from
our own cookieless analytics. Google Analytics is the exception: it sets cookies and
reports to Google, so it stays switched off until you allow it. You can change your
mind any time from “Cookie settings” in the footer.
Strictly necessary
Session security and form protection (tldrocket-session,
XSRF-TOKEN, 2 hours). The site cannot work without them,
so they need no consent.
Always on
Google Analytics 4 (_ga,
_ga_<id>, up to 2 years). Measures which
stories and sections readers use. Google acts as a third-party processor and may
store the data outside the EU. No advertising, no profiling, no data sold.