TLDRocket
Sign in

[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model

Latent Space Covered by 3 sources

Black Forest Labs released FLUX 3, a unified multimodal model generating video, audio, images, and controlling robot actions, with capabilities claimed to match or exceed Gemini Omni and Grok Imagine. The model supports text-to-video, image-to-video, video-to-video, multilingual dialogue, and keyframe transitions, with an open-weights developer version coming soon. FLUX-mimic, built on FLUX 3, enables robot control on single GPUs by transferring video world modeling to dexterity tasks, now being tested in factory settings with Audi.

Why it matters

A HUGE win for BFL!

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.