Black Forest Labs Releases FLUX 3: A Multimodal Flow Model for Image, Video, Audio and Robot Action Prediction
Black Forest Labs released FLUX 3, a multimodal foundation model that jointly learns from images, video, audio and robot action prediction within a single architecture, generating clips up to 20 seconds with native audio…