Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs
https://the-decoder.com/wp-content/uploads/2026/07/bfl_flux3_video.png" style="height: auto; margin-bottom: 10px;" width="1920" />
Black Forest Labs has released Flux 3, a multimodal foundation model that learns from images, video, and audio and can generate video with native sound for the first time.
BFL's own tests put it just ahead of market leader Seedance 2.0, though independent results aren't yet available.
The company ultimately wants to build a world model and is already testing Flux 3 on robotics tasks.
The article https://the-decoder.com/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-black-forest-labs/">Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs appeared first on https://the-decoder.com">The Decoder.