FLUX 3 Multimodal Foundation Model Controls Factory Robots
thione · x · 2026-07-27
Black Forest Labs announced early access for FLUX 3, its new multimodal foundation model jointly trained on images, video, and audio. Video prediction accounts for over 95% of compute costs.
Additionally, BFL partnered with mimic to develop FLUX-mimic. This video-action model leverages FLUX 3's physical world understanding to control factory robots via demonstration learning, already tested and deployed at Audi.
Related event: FLUX 3 Multimodal Model Crosses Over to Robot Control(2 posts)→
More from Embodied
- GPT-6 tested on LIBERO robot task: turns on stove, fails to grasp moka pot — YuXiang_IRVL · 2026-09-23
- OpenRoboto Shift launches: decentralized egocentric video data network for robot brains — markjeffrey · 2026-09-23
- OpenRoboto launches decentralized rival to Figure's robot data network on Bittensor — markjeffrey · 2026-09-23
- Typesafe's decision model Jev drives a Go2 robot inside NVIDIA Isaac Sim via OM1 — paigeinsf · 2026-09-23
- Tesla pitches Optimus as the first general-purpose humanoid robot — XFreeze · 2026-09-23
- XPENG's IRON robot demos full-duplex speech with 9-mic array and lip reading — ChrisGPT · 2026-09-23