Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio — but in limited release to start

AI Summary
Black Forest Labs has launched FLUX 3, a multimodal AI model capable of generating images and 20-second audio/video clips from a single prompt, with plans to extend its applications to robotic vision and actions. The release is currently limited, as the AI lab aims to fine-tune its capabilities across different modalities.
From the source
Black Forest Labs (BFL) is expanding its FLUX family beyond image generation with today's launch of FLUX 3, a multimodal frontier model trained to understand and generate images, or combined audio/video clips up to 20 seconds from a single prompt — and to extend the same underlying architecture to robotic vision and actions. The Freiburg, Germany-based AI lab says FLUX 3 is jointly trained across those modalities rather than assembling separate image, video and audio models behind a common inter
The full text couldn't be loaded here (the source may require a subscription).
View original at VentureBeat AI