MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights
Your creative destiny, on your terms.
Now Live at HailuoAI.video & MiniMax API.
More ways to build with MiniMax-H3. 🐮
Great to see model optimization and AMD inference engineering come together to give creators a faster feedback loop.⚡️
5s of video, generated in 1.3s!
Nunchux brings @minimax’s MiniMax-H3 to @AMD MI355X with up to 26.7× faster inference than SGLang on 8 GPUs @AMDServer
And with streaming generation, you can change the prompt as the video plays, steering what happens next.
One stack, optimized
We're proud to contribute as an AI partner to the newly launched @Singtel AI Pass, supporting Singapore's SkillsFuture AI Subscription initiative. 🇸🇬
MiniMax H3 (@Hailuo_AI), MiniMax Agent (@MiniMaxAgent) and MiniMax Audio are all included, giving eligible learners across 200+
Three paths to faster video attention: compute the same interactions more efficiently, compute fewer in full, or change how information is mixed. Here’s a visual guide. 👇
Thanks to Nunchux AI and collaborators for VC-Attention, bringing training-free low-bit acceleration to
Introducing VC-Attention: fast and accurate low-bit attention without retraining.
On MiniMax-H3, VC-Attention speeds up attention by 1.6× on B200 and 1.5× on B300 over FlashAttention-4, with better fidelity than SageAttention2. It also works with existing sparse attention