Black Forest Labs unveils FLUX 3: one model for images, 20-second video with audio, and robot actions
Germany's Black Forest Labs announced FLUX 3 on 2026-07-23, a multimodal flow model jointly trained on images, video, audio and action prediction; it is BFL's first video model (clips up to 20 s with synced audio) and powers FLUX-mimic, a robot-manipulation model being tested by Audi. A 7B open-weights FLUX 3 Action followed on 2026-09-23.
Key facts
- Single architecture jointly trained on images, video, audio and action prediction
- FLUX 3 Video: clips up to 20 seconds with synchronized audio; aspect ratios 9:16 to 21:9; up to 10 image references (secondary sources)
- FLUX-mimic (with mimic robotics): fine-tunes to a task with ~30 minutes of robot data vs 30+ hours previously
- Audi testing FLUX-mimic for soft-body manipulation in production and logistics
- Launch partners/testers: Adobe Photoshop, Canva, Picsart, Krea, Burda, Magnific; Nous Research's Hermes Agent
- Video and Action in early access at launch; open-weight and faster versions promised later in 2026
- FLUX 3 Action: 7B open-weights robot-control model published 2026-09-23 (DataNorth)
What happened
Black Forest Labs (maker of FLUX image models) moved beyond still images with FLUX 3. The same backbone generates images, video with native audio, and robot action sequences. Its robotics application, FLUX-mimic, built with Swiss startup mimic robotics, is claimed to cut the robot data needed for a new manipulation task from 30+ hours to ~30 minutes; Audi is deploying it in pilots. FLUX 3 Video and Action launched in gated early access; on 2026-09-23 BFL published FLUX 3 Action as a 7B open-weights model.
Why it matters
FLUX 3 is a concrete instance of the "world model → robot policy" convergence: a generative video model doubling as a robot foundation model. It also makes BFL, a European lab, a full-stack video competitor.
Changelog
- 2026-09-29: created
Related events
Sources (5)
- officialGlobeNewswire: Black Forest Labs unveils FLUX 3
- officialBFL blog: FLUX 3 Video, Part 1: Generation
- pressVentureBeat: FLUX 3 generates images and 20-second video with audio
- pressMarkTechPost: FLUX 3 multimodal flow model
- pressDataNorth: FLUX 3 Action 7B robotics model
id: 2026-07-23-black-forest-labs-flux-3 · updated 2026-09-29 · open in the interactive timeline