H3 is the first model to reach sd-level quality and, most importantly, it’s open weights.
Personally I think this is one of the biggest steps forward for both the ecosystem and the tech as a whole.
From today on, we won’t be the only ones here for sure.
The weights are here. The nodes are native. The rest is up to your graph. 👀 H3 is in @ComfyUI on Day 0!!
one model, five workflows, endless ways to wire it. 😌
#MiniMaxH3#OpenWeights
MiniMax H3 is native in ComfyUI, day zero with the weights.
Model highlights:
→ Text-to-video, prompt only
→ Image-to-video
→ First-and-last-frame - control the opening frame, the closing frame, or both
→ Reference-to-video - carry a subject, a motion, or a voice
Open weights are step one. Making them easy to serve is step two. H3 now has Day 0 support in @vllm_project's vLLM-Omni, complete with an OpenAI-compatible video endpoint.
Huge thanks to the vLLM team for bringing H3 into the open inference stack. ⚡
#MiniMaxH3#OpenWeights
🎉 Congrats to @MiniMax_AI on releasing the open weights for MiniMax H3! Day-0 support in vLLM-Omni!
One model reads text, images, video, and audio as a single context and returns video with native stereo audio. Text-to-video, first/last-frame, and multi-reference generation, 4
@MiniMax_AI H3 is live in SGLang Diffusion, with day-0 serving support 🎬
This open model matches Seedance 2.0 at 1/3 the cost, or $0 if you run it locally on 2x 5090 or 1 RTX 6000.
With SGLang Diffusion, you can build visual concepts, motion design, e-commerce creatives,