Provides a dual-channel, channel-separated sample (8.9 hours) and access path to a 1,000‑hour English conversational corpus for commercial and research use. Delivers 48 kHz per-speaker audio, word-level machine transcripts, and per-speaker metadata designed for full‑duplex/turn-taking and ASR/ TTS research.
Adds a plug-and-play linear-attention branch and LoRA adapters to MiniMax-H3 to run text-to-video generation faster than real-time (near-lossless quality tradeoffs). Includes an optimized FP8 inference stack and a community license with regional restrictions.
Provides 4.5 billion TikTok video records with captions, timestamps, music IDs and engagement counts for research; split across 27 zstd-compressed Parquet files (~289 GB) and sampled via TikTok's mobile API; released for research-use only with privacy and ToS caveats.