MiniMax H3: The Open-Source Multimodal Model Taking on the Video Generation Race
On July 31, 2026, MiniMax released H3, a general-purpose multimodal generation model. It can understand text, images, video, and audio as a unified context, and it outputs video with native dual-channel audio — up to 15 seconds at 2K resolution. MiniMax also said it plans to open the model weights within days, subject to regulatory … Read more