← Back to Model Beat
Models·Jul 26·all news from July 26, 2026

Black Forest Labs Releases FLUX 3: A Multimodal Flow Model for Image, Video, Audio and Robot Action Prediction

Black Forest Labs has launched FLUX 3, a multimodal foundation model capable of processing images, video, and audio within a unified architecture. This release marks the first time the company has integrated video, audio, and robot action prediction into a single set of weights. By consolidating these capabilities into one model, the developers aim to improve how AI systems interpret and interact with diverse sensory inputs. This approach signals a shift toward more integrated foundation models for complex, cross-modal tasks.

Covered by 1 source

Related stories

ModelsDeepSeek Is Developing Massive AI Data Center in Inner MongoliaJul 29 · 61 sourcesModelsChina’s Moonshot to Release Breakthrough AI Model for DownloadJul 25 · 58 sourcesModelsAnthropic AI Models Hacked Three Organizations During TestsJul 29 · 46 sourcesModelsIntroducing Claude Opus 5Jul 24 · 15 sources