AIAny
Icon for item

bones-studio/seed

Provides an annotated multimodal human-motion dataset for language-to-action and robotics research, with BVH and MuJoCo files plus recordings targeted at Unitree-G1 and NVIDIA-SOMA platforms. Covers locomotion, gestures, dance and object interaction with English annotations and 100K–1M samples.

Introduction

Why this matters Most embodied AI and language-to-action research stalls for lack of datasets that link natural language to whole-body, platform-specific motion. SEED fills that gap by combining motion-capture trajectories, formatted robot-ready files, and English action annotations so models can be trained end-to-end from text to executable motion.

What Sets It Apart
  • Multi-target outputs: includes BVH and MuJoCo artifacts plus recordings and assets aimed at Unitree-G1 and NVIDIA-SOMA — so what: eases transfer from learned policies to real/sim robot stacks without manual reformatting.
  • Broad motion coverage with annotations: locomotion, gesture, dance and object interaction labeled in English — so what: supports both low-level control tasks and higher-level language-conditioned behavior learning.
  • Practical scale and provenance: tagged as 100K–1M samples, hosted on Hugging Face with thousands of downloads and community likes — so what: large enough for representation learning while reflecting an applied robotics focus.
Who It's For and Tradeoffs

Great fit if you are training or evaluating language-conditioned controllers, motion-generation models, or sim-to-real pipelines for humanoid/legged platforms and need platform-ready assets and annotated trajectories. Look elsewhere if you need a dataset with a permissive, well-known open license (SEED’s license is unspecified/other) or if your work targets purely vision-only video tasks without motion control.

Where It Fits

SEED sits between pure motion-capture corpora (which may lack robot-format outputs) and robotics benchmarks (which often lack rich language annotations). Use it when your experiment needs both annotated semantics and robot-executable motion representations.

Information

  • Websitehuggingface.co
  • Organizationsbones-studio
  • Published date2026/03/10

More Items

Hugging Face

Generates humanoid robot motion references that preserve object contact locations/timing by solving windowed trajectory optimizations against contact targets in the object frame. Releases retargeted trajectories for two Unitree robots across 75 objects (≈13.9k robot–motion pairs); CC BY‑NC‑SA 4.0.

Hugging Face

Collection of 1.44M unique Turkish voice‑assistant sentences (≈1,819 hours estimated), normalized for TTS and organized by service domains (appointments, banking, e‑commerce). Designed for training and evaluating TTS and text-generation models; licensed CC BY 4.0 with required attribution.

Hugging Face

Provides a monthly Parquet snapshot of ~5.6 billion public TikTok videos (2014–Oct 2026), including captions, hashtags, sounds, engagement metrics and TikTok Shop links. Designed for large-scale querying (DuckDB/Pandas/Polars); licensed CC BY-NC 4.0 for research and personal use.