AIAny
AI Video2026
Icon for item

MiniMax H3 Character Swap LoRA

Replaces a selected person in a source video with a character from a reference image via a MiniMax H3 LoRA adapter, aiming to preserve scene, camera, and background. Trained for 1,000 updates; intended for Ref2VA runtimes and ComfyUI. Short (≈4–5s) continuous shots work best; distributed under the MiniMax H3 Community License.

Introduction

Character swaps often fail because they either destroy the background or drift in timing; this adapter explicitly trades model capacity toward keeping the source scene while injecting a reference character style. It’s a lightweight LoRA edit designed to be applied at inference to a MiniMax H3 Ref2VA runtime rather than a standalone replacement model.

Key Capabilities
  • Targeted character replacement: conditions on a reference image to change a single person’s identity, outfit, and art style while attempting to preserve the original camera, lighting, background objects, and the movements of other people — so you can retarget a character without recreating the whole scene.
  • Adapter workflow and compatibility: published as a single .safetensors LoRA (final checkpoint after 1,000 updates). Apply at strength 1.0 to a Ref2VA-capable MiniMax H3 base (VAEs and base weights must be provided separately) and run in ComfyUI or compatible runtimes — so it plugs into existing MiniMax H3 pipelines.
  • Practical operating constraints: the training set is small (94 synthetic edit triplets, 40 unchanged clips; 76 edits used for optimization) and regularization clips are short, so short continuous shots (~4–5s at 24 fps) and conservative strength settings give the most reliable results.
  • Known failure modes: temporal drift across longer windows, imperfect hard-cut handling, and inconsistent facial-expression fidelity; audio is not reliably preserved by the model and often requires postprocessing.
Who it's for and trade-offs

Great fit if you need quick, scene-preserving single-character swaps in short video clips and you already run MiniMax H3 Ref2VA in ComfyUI or a similar runtime. It’s useful for experimentation, style transfers, and localized edits where keeping the original background matters.

Look elsewhere if you need robust multi-character replacements, guaranteed lip-sync or audio fidelity, long uninterrupted sequences (>~10–15s) with strict camera-cut timing, or an adapter-free end-to-end model. The adapter inherits MiniMax H3’s licensing constraints (MiniMax H3 Community License) and requires the base weights and training assistant artifacts that are not bundled here.

Information

Categories

More Items

Hugging Face
AI Video2026

LoRA adapters for MiniMax H3 that sharpen and enhance videos in ComfyUI by conditioning on source clips via guide latents for pixel-level alignment. Designed mainly for ref2va as a second-pass sharpening tool, includes a ComfyUI workflow and example before/after clips; requires aligned guide clips at the target resolution and valid clip lengths.

Hugging Face
AI Video2026

Replaces a character in a video using a single repainted frame from the same clip and propagates that edit across the shot while preserving motion, camera and lighting; requires no pose estimator, segmentation, face tracker or text prompt. Key facts: a 33.1B MiniMax-H3 finetune, DMD-distilled to three forward passes, 124 frames in ~26s on one B200 GPU.

Hugging Face
AI Video2026

Generates short multimodal videos from text, images, or reference clips using a fine-tuned MiniMax‑H3 fusion model; improves HDR clarity, motion fluidity, distant-face fidelity and VFX while preserving MiniMax‑H3’s prompt/style behavior. Best used via ComfyUI.