Black Forest Labs Launches FLUX 3: Unified Multimodal Model for Image, Video, and Audio

bfl_ai · x · 2026-08-05

Black Forest Labs has officially released the FLUX 3 preview, a multimodal model trained uniformly across image, video, and audio. Its core highlight is the ability to generate video with synchronized audio via a single request shape (endpoint).

Core Features & Specs

Supported Modes

All requests run on the same flux-3-video endpoint, differentiated by a mode parameter:

The official blog notes that video editing and Omni Reference with images and videos will be available soon.

Related event: Black Forest Labs Launches FLUX 3 Multimodal Model(11 posts)→

Original post →

More from Multimodal

Multimodal channel →