Black Forest Labs Releases FLUX 3: Multimodal Model for Images, Video, Audio, and Robot Action Prediction
Black Forest Labs (BFL) has introduced FLUX 3, a multimodal foundation model trained on images, video, and audio within a single architecture. The model generates up to 20-second videos with synchronized audio and achieves high user preference results, surpassing Luma Ray 3.2 in 93% of cases and Runway Gen-4.5 in 77%.
Luma AI
Runway
xAI
Google/DeepMind
MarkTechPost26.07 · 21:02

