AIAny
AI Video2026
Icon for item

MiniMax-H3_comfy

Provides ComfyUI-compatible conversions and LoRA adapters of the MiniMax‑H3 video+audio generative model, with example presets and demo videos to run short stereo audio+video inference inside ComfyUI workflows.

Introduction

MiniMax H3 packs multi‑modal video and native stereo audio generation into a single model pass; this repository makes those assets accessible inside ComfyUI by providing converted model files, LoRA adapters, example presets and demo outputs. The practical effect is a lower-friction path for ComfyUI users to experiment with MiniMax‑H3-style T2VA/FL2VA/Ref2VA workflows without hand-building node stacks from scratch.

What Sets It Apart
  • ComfyUI-ready conversions: packaged weights and processor/tokenizer files formatted for ComfyUI nodes so users can drop them into existing Comfy workflows.
  • LoRA adapters and recommended strengths: includes distilled LoRAs (example strengths noted) to tweak style or resource usage without retraining full weights.
  • Demo content and presets: example videos and preset node configurations help reproduce expected outputs and iterate quickly.
  • Focused on inference and integration: aimed at using MiniMax‑H3 within ComfyUI rather than providing new training pipelines or novel model research.
Who It's For and Tradeoffs

Great fit if you use ComfyUI and want to prototype short multimodal video+audio generations with ready-made MiniMax‑H3 conversions and LoRAs. Look elsewhere if you need official upstream training code, licensed original checkpoints bundled here, or lightweight CPU-only execution — running full generative video models typically requires a capable GPU and/or INT8 optimizations and may rely on obtaining original MiniMax weights or partner runtimes separately.

Information

Categories

More Items

Hugging Face
AI Video2026

LoRA adapters for MiniMax H3 that sharpen and enhance videos in ComfyUI by conditioning on source clips via guide latents for pixel-level alignment. Designed mainly for ref2va as a second-pass sharpening tool, includes a ComfyUI workflow and example before/after clips; requires aligned guide clips at the target resolution and valid clip lengths.

Hugging Face
AI Video2026

Replaces a character in a video using a single repainted frame from the same clip and propagates that edit across the shot while preserving motion, camera and lighting; requires no pose estimator, segmentation, face tracker or text prompt. Key facts: a 33.1B MiniMax-H3 finetune, DMD-distilled to three forward passes, 124 frames in ~26s on one B200 GPU.

Hugging Face
AI Video2026

Generates short multimodal videos from text, images, or reference clips using a fine-tuned MiniMax‑H3 fusion model; improves HDR clarity, motion fluidity, distant-face fidelity and VFX while preserving MiniMax‑H3’s prompt/style behavior. Best used via ComfyUI.