Skip to content
Read the original: MiniMax · new models on Hugging Face·Published· Jul 28, 2026PickAI score76

MiniMax H3 releases open-weight omni-modal video model with native stereo audio

MiniMaxAI/MiniMax-H3

AISummary

MiniMax released H3, an open-weights omni-modal model that generates video with native stereo audio up to 2K and 15 seconds. The system combines H3-Context-IR preprocessing, the H3-Base generator at 768p, and H3-Regenerate-2K for 2K output, with the Context-IR and 2K modules available only through API.

AIWhy it matters

The source details a three-module pipeline and open weights with deployment paths, showing how a video model is served and reproduced locally.

Read the original huggingface.co

Source: MiniMax · new models on Hugging Face · huggingface.co