MiniMax 盘点 H3 视频生成模型开源生态进展

MiniMax (official) · @MiniMax_AI · X·2026-09-14 14:14·1小时前
AI 导读

MiniMax 盘点其开源视频生成模型 H3(支持原生立体声音频和多模态参考控制)的社区优化成果,并附生态汇总仓库 https://github.com/MiniMax-AI/awesome-minimax-h3-integration。

MiniMax (official)@MiniMax_AI
61AI 编辑部评分,满分 100

MiniMax 盘点 H3 视频生成模型开源生态进展

2026-09-14 14:14· 1小时前
AI 导读

MiniMax 盘点其开源视频生成模型 H3(支持原生立体声音频和多模态参考控制)的社区优化成果,并附生态汇总仓库 https://github.com/MiniMax-AI/awesome-minimax-h3-integration。

Open weights. Shared progress. MiniMax H3 is moving fast.

We built H3 for video generation with native stereo audio and multimodal reference control. The open-source community is making that capability faster, more accessible, and easier to build on.

Recent highlights:

• FastH3 — FastVideo, Nuva Lab and NVIDIA: 4-step distillation, now running on DGX Spark and Apple Silicon.

• Sol-H3 — NVIDIA’s SANA team: 15 seconds of 768p video + audio in 6.6 seconds on 8×B300, in the team’s warm-inference benchmark.*

• VDN — Haocheng Xi and the OpenVDN team: rethinking attention for faster H3 inference, with weights, training and inference code released.

• PDD — NVIDIA’s distillation method, brought to H3 by Alibaba PAI as 8-step Acc-LoRAs, now supported in ComfyUI.

• LightX2V — 4- and 8-step Turbo LoRAs, with workflows for text, image and reference-conditioned video + audio.

Behind every release are people training, optimizing, quantizing, testing and sharing. Special thanks to: @haoailab @nuvalab @NVIDIAAI @xieenze_jr @HaochengXiUCB @ArashVahdat @julberner @LightX2V @ComfyUI And to the individual contributors pushing the work forward: @haozhangml @cxlcl1 @lawrence_cjs @yitongli165665 @haopengl33 @songhan_mit @shanasaimoe

Thank you for building with H3 and helping make it faster, more accessible, and more useful for the community. Powerful models go further when we build together. Keep pushing H3. Excited to see what comes next. 🚀

Explore the ecosystem: https://github.com/MiniMax-AI/awesome-minimax-h3-integration