FLUX 3 Video adds native audio, 20-second clips and video continuation
No_Winner_3807 · reddit · 2026-07-24
FLUX 3 Video looks promising because it goes beyond simple image-to-video motion and instead uses a unified image/video/audio architecture.
Key capabilities mentioned:
- native audio support
- generations up to 20 seconds
- image and video references
- keyframe transitions
- video continuation
- video-to-video generation
The poster is especially interested in how these controls might surface in ComfyUI as separate nodes for visual references, keyframes, audio conditioning, and continuation. The release is still in early access, so there is no public ComfyUI workflow yet. They are also waiting on API parameters, pricing, generation speed, and whether a future FLUX 3 Dev release will include usable open weights.
Related event: Black Forest Labs Launches FLUX 3: Towards Omni-modal and World Models(27 posts)→
More from Models
- User Cancels Claude Max After Weeks of Talking, Saying the Model Is Just Too Moralistic — breath_mirror · 2026-07-24
- Kimi K3 again posts clear benchmark wins over Inkling across five tests — echen · 2026-07-24
- Kimi K3 beats Inkling on five benchmarks, but the size gap is huge — echen · 2026-07-24
- Leaked Anthropic screenshot shows guardrails routing a chat from Fable 5 to Opus 5 — banteg · 2026-07-24
- Announcing EQ-Bench 4: Benchmarking LLM Emotional Intelligence via Multi-Turn Roleplay — sam_paech · 2026-07-24
- Polymarket prices a 11% chance of a Chinese top AI model this year — Polymarket · 2026-07-24