MiniMax H3 releases open-weight omni-modal video model with native stereo audio
MiniMaxAI/MiniMax-H3
AISummary
MiniMax released H3, an open-weights omni-modal model that generates video with native stereo audio up to 2K and 15 seconds. The system combines H3-Context-IR preprocessing, the H3-Base generator at 768p, and H3-Regenerate-2K for 2K output, with the Context-IR and 2K modules available only through API.
AIWhy it matters
The source details a three-module pipeline and open weights with deployment paths, showing how a video model is served and reproduced locally.
Source: MiniMax · new models on Hugging Face · huggingface.co