AIAny
AI Video2025
Icon for item

Y2A-Auto

Automatically transfers YouTube videos to AcFun and bilibili with an end-to-end pipeline: downloading, ASR, subtitle translation and QC, AI-generated metadata, content moderation, and automated uploads; includes a web dashboard and monitoring.

Introduction

This project addresses a common bottleneck for cross-posting video content at scale: stitching together downloading, speech-to-text, subtitle translation and QC, AI metadata generation, content moderation, and platform uploads into a single maintainable pipeline. It shifts manual reposting tasks into configurable automation while keeping human-in-the-loop controls for review and safety.

What Sets It Apart
  • Integrated AI steps so you can run ASR (Whisper/Voxtral), subtitle translation, AI-based subtitle QC, and AI-generated titles/descriptions/tags in one flow — so what: reduces repetitive manual editing and speeds up multilingual reposting.
  • Platform-specific automation: built-in AcFun/bilibili upload modules, QR/Cookie login support and CookieCloud sync — so what: lowers the friction of platform authentication and supports scheduled monitoring of channels/keywords.
  • Deployment-friendly design: Docker-first with optional local mode, hardware-accelerated encoding (HEVC fallback to H.264), and configurable concurrency — so what: fits small self-hosted servers to GPU-enabled instances without heavy rewrites.
  • Safety and observability: password protection, brute-force lockout, content-moderation hooks (Aliyun Green), and async notification channels — so what: makes unattended automation safer and auditable.
Who It's For and Tradeoffs

Great fit if you need to repeatedly repost multilingual YouTube content to Chinese video platforms and want automation for subtitles and metadata while preserving manual audit gates. Not ideal if you require strict platform compliance/legal guarantees from the upstream content owner or need a zero-dependency hosted SaaS; it’s self-hosted, requires valid cookies/API keys, and may need tuning for large-volume or highly regulated pipelines.

More Items

Hugging Face
AI Video2026

Turns a single photo into a geometry-consistent, frozen-time 360° camera orbit that returns to the exact start frame. Implemented as a LoRA for MiniMax‑H3 FL2VA — use identical first+last keyframes to produce seamless orbit clips; trained on a small human-centric square orbit dataset, so results are domain-limited.

Hugging Face
AI Video2026

Replaces a selected person in a source video with a character from a reference image via a MiniMax H3 LoRA adapter, aiming to preserve scene, camera, and background. Trained for 1,000 updates; intended for Ref2VA runtimes and ComfyUI. Short (≈4–5s) continuous shots work best; distributed under the MiniMax H3 Community License.

Hugging Face
AI Deploy2026

Evaluates multi-field JSON schemas in parallel to extract boolean or categorical field values from text, producing guaranteed-valid JSON and per-field calibrated confidences. Uses KV-cache broadcasting, sub-vocabulary logit slicing and token-tree disambiguation to cut latency (5.6x–7.0x on M4 Max) versus autoregressive decoding; requires Apple Silicon and MLX.