TLDRocket
Sign in

Models & Research

1194 summarised stories in Models & Research, each linking back to the original source. Browse all topics →

Wednesday, 2 September 2026

Google DeepMind Releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber: One Core Model, Two Access Envelopes

MarkTechPost 2 days ago 8 3 sources

Google DeepMind released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber as two variants built on the same core model but separated by different safety access rules. Gemini 3.8 Flash stays priced at $0.75 per 1M input tokens and $3.75 per 1M output tokens through December 31, 2026. Gemini 3.8 Flash is broadly deployable via Google’s platforms while Flash Cyber is restricted case by case through the Fairwind Program, and the 3.8 Flash API no longer supports MINIMAL (it returns a validation error).

Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

The Verge 2 days ago 27 5 sources

Google launched Gemini 3.8 Flash, claiming it performs more reasoning steps and uses iterative tool calls compared with Gemini 3.7 Flash. Pricing starts at $0.75 per million input tokens and $3.75 per million output tokens, but Google warns it may use more tokens at higher effort levels. Developers can stay on Gemini 3.7 Flash to minimize token usage.

The Sequence Learning Loop - Issue 925: Learn About Fable and Mythos 5.1, GLM-5.3-Flash, and Qwen 3.8

TheSequence 2 days ago 49

The Sequence Learning Loop newsletter highlighted three recent model releases: Anthropic’s Claude Fable 5.1 and Mythos 5.1, Zhipu’s GLM-5.3-Flash, and Alibaba’s Qwen 3.8 family. Zhipu’s GLM-5.3-Flash is a 320B model that activates 18B parameters and ran for a week serving anonymous traffic on Chinese chips. Together, the releases push toward cheaper long-hours self-running use by mixing different approaches to model size, activation, and deployment constraints.

REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs

Apple Machine Learning Research 2 days ago 40

Researchers introduced REFACTOR-VLA to turn monolithic vision-language-action behavior into reusable typed motor programs learned with a wake/sleep loop. The system clusters motor-program segments in the sleep phase using a Behavioral-Equivalence Kernel driven by rollouts in a learned latent world model trained via a three-phase schedule, and it reports NMI scores for n=3 multi-seeding including 0.915 ± 0.013 on the Goal suite. Performance shifts as a bigger world model (188M to 430M parameters) worsens 4 out of 4 LIBERO benchmark suites while adding an auxiliary InfoNCE contrastive loss in Phase A improves the quality of skill clustering in Phase C.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.