TLDRocket
Sign in

Model Release

82 summarised stories about Model Release, each linking back to the original source. Browse all topics →

+ Follow this topic

Thursday, 30 July 2026

llm 0.32rc2

Simon Willison's Weblog 4 weeks ago 28 14 sources

LLM 0.32rc2 release updated the default model from GPT-4o mini to GPT-5.6 Luna and added an openai endpoint command for running prompts against OpenAI-compatible endpoints without prior configuration. GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens, compared to GPT-4o mini's $0.15/$0.60. Users can now test prompts against local or remote compatible endpoints directly via CLI without installing LLM.

Google DeepMind Ships Three Physical AI Models For Whole Body Control, Dexterity And Multi Robot Collaboration

MarkTechPost 1 month ago 36 7 sources

Google DeepMind released Gemini Robotics 2, a suite of three AI models for controlling robots across whole-body movement, multi-finger dexterity, and multi-robot collaboration. The system achieves success rates ranging from 32% to 92% depending on task complexity, with one model checkpoint controlling multiple robot embodiments including the Apptronik Apollo 2 and Franka Duo. Robots can now adapt to unpredictable environments, transfer skills between different bodies, and coordinate with other robots on shared tasks rather than relying on pre-programming or teleoperation.

llm 0.32rc1

Simon Willison's Weblog 1 month ago 14 14 sources

LLM released version 0.32 release candidate, introducing a new database schema that uses content-addressable hash IDs for stored messages to enable deduplication and support for forked conversations. The update adds support for three new model variants: gpt-5.6-sol, gpt-5.6-terra, and gpt-5.6-luna. Existing data is preserved but users should back up their logs.db file before upgrading due to the significant schema changes.

How GPT-5.6 fuses frontier intelligence with frontier efficiency

OpenAI 1 month ago 45 14 sources

OpenAI released GPT-5.6, a model family designed to balance performance with cost efficiency through load balancing and caching techniques that process more tokens effectively. The model includes a variant called GPT-5.6 Sol that optimizes its own code execution and resource allocation. This allows the system to handle increased workloads while reducing computational costs.

Gemini Flash Gets Pricey, AI Act Delays, Agents Drive Online Traffic

The Batch 30 2 sources

Google launched Gemini 3.5 Flash, a faster mid-tier model with improved agentic capabilities and visual understanding, but at three times the price of its predecessor. The model costs $1.50 per million input tokens (compared to $0.05 for Gemini 3 Flash) and topped benchmarks like APEX-Agents-AA with 47.1% accuracy. Higher per-token prices across major AI providers mean the "Flash" designation no longer signals a cost advantage for developers building agents. Meanwhile, the European Union delayed high-risk AI regulations to December 2027 and December 2026 for other provisions, citing competitiveness concerns, though it strengthened rules banning non-consensual intimate imagery.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.