[AINews] Fal’s H3 Max Live breaks the infinite videogen barrier
Latent Space
Fal turned MiniMax H3 Max into a live video stream that reacts to viewers. It’s messy slop, but it shows near-real-time video is no longer theoretical.
Based on reporting by Latent Space — read the original for the full story.
Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error
Fal took MiniMax’s H3 release from last month, post-trained it for cost and quality, then tuned it for its own inference engine. The company says that brought a 35x speedup over the official endpoint. That mattered because generative video has always been trapped by time: even when generation got faster, it still lagged far behind the pace of human attention.
The breakout moment was a live stream that kept generating continuously. Ethan Mollick noticed it first, and then fal employees turned the idea into an “infinite” Twitch stream. Twitch and YouTube quickly kicked Fal off the platform, so the company moved the experience to its own live video service, which it pitched as a kind of “twitch plays pokemon” for video generation.
The stream itself sounds more like a stress test than a product demo. Fal’s own framing is that viewers can upvote LLM-generated prompts, and the system is powered by H3 Max Director, an autoregressive continuous version of H3 Max with up to two minutes of context. In parallel, Fal also launched Reference-to-Video for MiniMax H3 Max, with early preview claims of real-time factor 1 at 768p.
What makes this interesting isn’t that the output looks good. The source is blunt that it doesn’t. It’s that live, audience-steered video generation now exists at all, even if the result is chaotic, low-quality, and barely watchable. That is still a line crossed, and the industry tends to notice those lines only after they’ve already been crossed.
My take — AI-written commentary, not fact-checked reporting
This is the usual AI trick: ship something ugly enough that nobody wants to watch it, then act as if that proves the future has arrived. It mostly proves that speed matters more than polish right now, and that the platforms are already nervous. The real story is not the slop stream; it’s how quickly “impossible” becomes a moderation problem.
Read more about this at: Latent Space
Related stories
MiniMax Releases MiniMax H3: An Omni-Modal Video Model That Generates 15-Second 2K Clips With Native Stereo Audio
MarkTechPost · 1 month ago ·
24
FLUX 3 Video creates clips up to 20 seconds with native audio from text and images
Black Forest Labs · 4 weeks ago ·
40
[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model
Latent Space · 1 month ago ·
9