FLUX 3 teaser hints at a multimodal generator with 20-second video clips
eyishazyer · x · 2026-07-23
Black Forest Labs’ apparent FLUX 3 teaser suggests a full multimodal generator with image, video, audio, and motion generation.
- The post says the model briefly surfaced as a preview and may be open if the usual pattern holds.
- It is described as supporting up to 20 seconds of video in a single generation.
- The author argues that, if released openly, it could give creators a major new capability boost.
Related event: Black Forest Labs Teases FLUX 3 with Multimodal Capabilities(2 posts)→
More from Multimodal
- Open-source node-based LoRA trainer puts captioning, checkpoints and VRAM stats in one graph — ashishsanu · 2026-07-23
- TERRA-129 debuts as an AI-animated sci-fi episode credited to Matygoo — Matygoo1 · 2026-07-23
- GPT Image 2 prompt aims to lock character identity and outfit consistency in ChatGPT — eyishazyer · 2026-07-23
- Alibaba launches Qwen-Audio-3.0-TTS with 16 languages and 3-minute one-pass audio — Alibaba_Qwen · 2026-07-23
- Kling AI is said to handle close-up facial expressions better — burny_tech · 2026-07-23
- FameGrid Krea 2 aims to generate more realistic social-media-style images — UltraMuseArt · 2026-07-23