{"slug": "flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs", "title": "Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs", "summary": "Black Forest Labs released Flux 3, a multimodal foundation model that generates video with native audio for up to 20 seconds, a first for the company. The model outperforms Seedance 2.0 in internal tests, and the company plans to build a world model, already testing Flux 3 on robotics tasks.", "body_md": "# Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs\n\n[The Decoder](https://the-decoder.com)\n\nBlack Forest Labs has released Flux 3, a [multimodal](/glossary/multimodal) [foundation model](/glossary/foundation-model) that learns from images, video, and audio and can generate video with native sound for the first time. BFL's own tests put it just ahead of market leader Seedance 2.0, though independent results aren't yet available. The company ultimately wants to build a [world model](/glossary/world-model) and is already testing Flux 3 on [robotics](/category/robotics) tasks.\n\nThe article [Flux 3 generates videos with native audio up to 20 seconds long, a first for Black Forest Labs](https://the-decoder.com/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-black-forest-labs/) appeared first on [The Decoder](https://the-decoder.com).\n\nGet AI news in your inbox\n\nDaily digest of what matters in AI.\n\n## Key Terms Explained\n\n[Decoder](/glossary/decoder)\n\nThe part of a neural network that generates output from an internal representation.\n\n[Foundation Model](/glossary/foundation-model)\n\nA large AI model trained on broad data that can be adapted for many different tasks.\n\n[Multimodal](/glossary/multimodal)\n\nAI models that can understand and generate multiple types of data — text, images, audio, video.\n\n[World Model](/glossary/world-model)\n\nAn AI system's internal representation of how the world works — understanding physics, cause and effect, and spatial relationships.", "url": "https://wpnews.pro/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs", "canonical_source": "https://www.machinebrief.com/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-l-mfzs", "published_at": "2026-07-23 18:03:01+00:00", "updated_at": "2026-07-23 18:37:14.833462+00:00", "lang": "en", "topics": ["artificial-intelligence", "generative-ai", "ai-products"], "entities": ["Black Forest Labs", "Flux 3", "Seedance 2.0"], "alternates": {"html": "https://wpnews.pro/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs", "markdown": "https://wpnews.pro/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs.md", "text": "https://wpnews.pro/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs.txt", "jsonld": "https://wpnews.pro/news/flux-3-generates-videos-with-native-audio-up-to-20-seconds-long-a-first-for-labs.jsonld"}}