Black Forest Labs Launches FLUX 3: Unifying Image, Video, Audio, and Robotics

bennash · x · 2026-08-12

Black Forest Labs has officially announced FLUX 3, a new multimodal model that unifies image, video, audio generation, and robotics action prediction into a single architecture.

Key Features:

Availability: The model is available via API, open weights (for custom fine-tuning), and enterprise solutions. The web playground currently offers free video generation trials, allowing users to use storyboard images as starting frames for reference.

Related event: Black Forest Labs Launches FLUX 3 Multimodal Model(2 posts)→

Original post →

More from Embodied

Embodied channel →