← Back to Model Beat
Models·Aug 1·all news from August 1, 2026

MiniMax Releases MiniMax H3: An Omni-Modal Video Model That Generates 15-Second 2K Clips With Native Stereo Audio

MiniMax has launched MiniMax H3, a multimodal model capable of generating 15-second, 2K resolution video clips accompanied by native stereo audio. Unlike systems that rely on modular components for sound, this model processes text, images, video, and audio within a unified framework. This integration aims to improve synchronization between visual content and its corresponding audio output.

Covered by 1 source

Related stories

ModelsDeepSeek Is Developing Massive AI Data Center in Inner MongoliaJul 29 · 61 sourcesModelsAlibaba’s Qwen3.8-Max AI Model Claims Benchmark Scores Rivaling AnthropicAug 3 · 82 sourcesModelsAnthropic AI Models Hacked Three Organizations During TestsJul 29 · 46 sourcesModelsGemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaborationJul 30 · 8 sources