AIAny
AI Video2026
Icon for item

MiniMax-H3_comfy

Provides ComfyUI-compatible conversions and LoRA adapters of the MiniMax‑H3 video+audio generative model, with example presets and demo videos to run short stereo audio+video inference inside ComfyUI workflows.

Introduction

MiniMax H3 packs multi‑modal video and native stereo audio generation into a single model pass; this repository makes those assets accessible inside ComfyUI by providing converted model files, LoRA adapters, example presets and demo outputs. The practical effect is a lower-friction path for ComfyUI users to experiment with MiniMax‑H3-style T2VA/FL2VA/Ref2VA workflows without hand-building node stacks from scratch.

What Sets It Apart
  • ComfyUI-ready conversions: packaged weights and processor/tokenizer files formatted for ComfyUI nodes so users can drop them into existing Comfy workflows.
  • LoRA adapters and recommended strengths: includes distilled LoRAs (example strengths noted) to tweak style or resource usage without retraining full weights.
  • Demo content and presets: example videos and preset node configurations help reproduce expected outputs and iterate quickly.
  • Focused on inference and integration: aimed at using MiniMax‑H3 within ComfyUI rather than providing new training pipelines or novel model research.
Who It's For and Tradeoffs

Great fit if you use ComfyUI and want to prototype short multimodal video+audio generations with ready-made MiniMax‑H3 conversions and LoRAs. Look elsewhere if you need official upstream training code, licensed original checkpoints bundled here, or lightweight CPU-only execution — running full generative video models typically requires a capable GPU and/or INT8 optimizations and may rely on obtaining original MiniMax weights or partner runtimes separately.

Information

Categories

More Items

Hugging Face
AI Video2026

Generates short videos with stereo audio from text prompts using a MiniMax‑H3 checkpoint; community‑uploaded on Hugging Face and distributed under Apache‑2.0. Tuned toward stylized creature and floral visuals and updated frequently per the model card.

Hugging Face
AI Video2026

Packaged diffusers checkpoint of MiniMax H3 for image/text-to-short-video generation with native stereo audio; provided for direct use in diffusers image-to-video pipelines and aimed at easy integration into prototyping and production workflows.

Hugging Face
AI Video2026

Provides ComfyUI-compatible pruned/curve-form LoRA conversions of the MiniMax‑H3 Turbo 4-step audio‑video generation preview, including further-trained ckpt500 EMA and non‑EMA variants and an example ComfyUI workflow for low-step experiments.