TLDRocket
Sign in

[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model

Latent Space

Black Forest Labs released FLUX 3, a unified multimodal model generating video, audio, images, and controlling robot actions, with capabilities claimed to match or exceed Gemini Omni and Grok Imagine. The model supports text-to-video, image-to-video, video-to-video, multilingual dialogue, and keyframe transitions, with an open-weights developer version coming soon. FLUX-mimic, built on FLUX 3, enables robot control on single GPUs by transferring video world modeling to dexterity tasks, now being tested in factory settings with Audi.

Why it matters

A HUGE win for BFL!

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.