Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio — but in limited release to start
Black Forest Labs (BFL) is expanding its FLUX family beyond image generation with today's launch of FLUX 3 , a multimodal frontier model trained to understand and generate images, or combined audio/video clips up to 20 seconds from a single prompt — and to extend the same underlying architecture to robotic vision and actions. The Freiburg, Germany-based AI lab says FLUX 3 is jointly trained across those modalities rather than assembling separate image, video and audio models behind a common interface. That distinction is central to the company's pitch: BFL wants enterprises to think
Read the full story at VentureBeat →