LaunchAI ModelsAugust 1, 2026

MiniMax releases H3 omni-modal video model with native stereo audio

MiniMax H3 generates 15-second 2K clips with native stereo audio, reading text, images, video, and audio as one unified context. It is positioned as a general-purpose multimodal generation model, not a text-to-video model with add-ons.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed