Black Forest Labs Launches FLUX 3: Unifying Image, Video, Audio, and Robotics
pess_r · x · 2026-07-24
Black Forest Labs officially announces FLUX 3. Jointly trained in one unified architecture, the model handles Image, Video, Audio, and Action-Prediction.
FLUX 3 Video is now available in early access. The team also demonstrated its robotics potential: in collaboration with Mimic Robotics, they built a Video-Action Model (FLUX-mimic) on top of FLUX 3, which is already being tested on complex manufacturing tasks with Audi. This marks generative visual tech rapidly expanding into physical world control.
Related event: Black Forest Labs Launches FLUX 3: A Unified Multimodal Foundation Model(28 posts)→
More from Embodied
- Amazon and Google sold 600M+ smart speakers, so why no AGI-era successor? — julianlehr · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- Polish developers build iPhone app that detects nearby Meta smart glasses — Low-Honeydew6483 · 2026-09-11
- Ant's Afu health AI hits 150M users, unveils AI+hardware health alliance at Bund Summit — APPSO · 2026-09-11
- Johns Hopkins Launches Full-Stack Hands-on Robot Learning Class with SO-101 Arm Kits — _krishna_murthy · 2026-09-11
- SyncWorld: In-Context Robot World Model Simulates Unseen Views and Embodiments Zero-Shot — ChongZzZhang · 2026-09-11