Black Forest Labs Unveils FLUX 3 Multimodal Foundation Model with Joint Image, Video, and Audio Generation
Black Forest Labs launches FLUX 3, a unified multimodal foundation model that jointly learns from images, video, and audio. Early access available now for video generation with native audio up to 20 seconds, with image synthesis and robotics action prediction following.
Read more →