AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Category

Explore by categories

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All Categories

  • AI Leaderboard

  • AI Agent Tutorials

  • AI Coding Tutorials

  • AI Model

  • AI Agent Papers

  • Chatbot

  • AI Dataset

  • Machine Learning Foundation Books

  • AI Train

  • AI Deploy

  • AI Client

  • Machine Learning Foundation Papers

  • Machine Learning Foundation Tutorials

  • AI Image Demos

  • AI Agent

  • Large Language Model Tutorials

  • Large Language Model Papers

  • Machine Learning Engineering Papers

  • Computer Vision Tutorials

  • Computer Vision Papers

  • Natural Language Processing Papers

  • Reinforcement Learning Papers

  • Speech Technology Papers

  • AI API

  • AI Coding

  • AI Image

  • AI Video

  • MLOps

  • MCP Client

  • MCP Server

  • AI Video Papers

  • AI Audio

  • AI Others

  • AI Infra

  • Embodied AI

Hugging Face
AI Video·2026
Icon for item

MiniMax H3

MiniMaxAI

Generates synchronized stereo audio and video from multimodal inputs (text, images, video, audio), producing 4–15s clips at 24 FPS with a 768p base and an in‑context regeneration path to 2K; supports first/last‑frame and multi‑reference modes and ships as two task‑specific checkpoints.

#diffusers#multimodal#video#audio#ai-api+5
Hugging Face
AI Model·2026
Icon for item

MiniMax-H3 GGUFs

realrebelai

Provides GGUF-quantized, ComfyUI-ready MiniMax‑H3 model files (FL2VA/REF2VA, text encoder, audio/video VAEs) to enable local ComfyUI inference for short video + stereo audio generation; requires the official VAEs and sufficient VRAM.

#multimodal#video#audio#huggingface#diffusers+4
Hugging Face
AI Video·2026
Icon for item

TenStrip/10Eros-Max

TenStrip·TenStrip, MiniMaxAI

Experimental MiniMax H3 variant that injects learned stylistic and motion 'character' from LTX 2.3, Wan 2.2 and Krea 2 into H3 by surgically grafting attention and MLP components; preserves H3 modality routing while shifting t2v/i2v aesthetics, with limited audio impact and community-license constraints.

#video#ai-video#multimodal#huggingface#gemma+2
Hugging Face
AI Video·2026
Icon for item

SexGod1979/PinkCherry_MiniMax-H3

SexGod1979

Generates short videos with stereo audio from text prompts using a MiniMax‑H3 checkpoint; community‑uploaded on Hugging Face and distributed under Apache‑2.0. Tuned toward stylized creature and floral visuals and updated frequently per the model card.

#transformers#huggingface#ai-video#video#audio+1
Hugging Face
AI Video·2026
Icon for item

MiniMax-H3 Turbo LoRA

larryvrh

A LoRA adapter for MiniMax-H3 that enables joint video + synchronized stereo audio generation in as few as 4 sampler steps, cutting sampling time roughly ~5×; early prototype under-trained, so 6–8 steps or newer checkpoints give better sharpness.

#lora#ai-video#video#audio#multimodal+3
Hugging Face
AI Video·2026
Icon for item

MiniMax-H3 Turbo 4-Step — ComfyUI Pruned-Model LoRAs

drbaph

Provides ComfyUI-compatible pruned/curve-form LoRA conversions of the MiniMax‑H3 Turbo 4-step audio‑video generation preview, including further-trained ckpt500 EMA and non‑EMA variants and an example ComfyUI workflow for low-step experiments.

#lora#ai-video#audio#video#huggingface+4
Hugging Face
AI Video·2026
Icon for item

MiniMax-H3-Prompt-Rewriter-LoRA

lightx2v, ModelTC +2

Turns a short prompt plus aspect ratio and duration into a structured, shot-by-shot audio-video description for text-to-audio-video generation. A PEFT LoRA on Qwen3.6-27B that expands timing, camera motion, continuity, and synchronized diegetic/non‑diegetic sound; text-only and requires MiniMax-H3 + LightX2V to produce final AV.

#lora#qwen#ai-video#multimodal#prompt-engineering+3
Hugging Face
AI Video·2026
Icon for item

Minimax-h3-Turbo

lightx2v·ModelTC

Packaged diffusers checkpoint of MiniMax H3 for image/text-to-short-video generation with native stereo audio; provided for direct use in diffusers image-to-video pipelines and aimed at easy integration into prototyping and production workflows.

#diffusers#huggingface#ai-video#video#multimodal+3
Hugging Face
AI Video·2026
Icon for item

MiniMax-H3_comfy

Kijai

Provides ComfyUI-compatible conversions and LoRA adapters of the MiniMax‑H3 video+audio generative model, with example presets and demo videos to run short stereo audio+video inference inside ComfyUI workflows.

#ai-video#multimodal#audio#lora#huggingface+3
Hugging Face
AI Video·2026
Icon for item

MiniMax H3 Realism People LoRA

Lovis Odin·fal, MiniMaxAI

A LoRA adapter for MiniMax H3 that improves photorealistic rendering of people—preserving skin texture, coherent micro-expressions, film-style lighting and subtle handheld motion. Trigger word: r34l1sm; intended for text-to-video portrait and close-up shots.

#ai-video#video#multimodal#huggingface#ai-train
Hugging Face
AI Video·2026
Icon for item

Minimax H3 Latent Upscaler

LBH-123-AI

Upscales Minimax H3 24-channel VAE latents in-place to increase spatial resolution while preserving the time dimension. Replaces the decode→pixel-upscale→encode round-trip with a learned 2D/3D latent upscaler to save compute and avoid interpolation ghosting; supports 1.0–4.0× scaling.

#video#ai-video#safetensors#pytorch#huggingface+2
Hugging Face
AI Video·2026
Icon for item

MiniMax-H3-Fun-Controlnet-Union

Alibaba PAI

Conditions a MiniMax‑H3 video generator with a single ControlNet‑Union checkpoint to accept Canny, Depth, HED, MLSD or Pose control videos and run video inpainting. Guidance‑distilled for one‑pass inference; requires the base MiniMax‑H3 weights and specific control-branch config.

#ai-video#video#multimodal#huggingface#qwen+2
  • Previous
  • 1
  • 2
  • More pages
  • 6
  • 7
  • 8
  • Next