Home › Datasets › task › Motion Synthesis

Motion Synthesis datasets

archive 2025-07-28

19 datasets carry the task tag "Motion Synthesis" (the task itself: Motion Synthesis), ordered by the archive's paper count. Page 1 of 1: 19 shown of 19. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Motion Synthesis datasets 1–19 of 19

HumanML3D is a 3D human motion-language dataset that originates from a combination of HumanAct12 and Amass dataset.
201 papers · 2 benchmarks
The KIT Motion-Language is a dataset linking human motion and natural language.
48 papers · 2 benchmarks
HumanAct12 is a new 3D human motion dataset adopted from the polar image and 3D pose dataset PHSPD, with proper temporal cropping and action annotating.
37 papers · 2 benchmarks
InterHuman is a multimodal dataset, named InterHuman.
24 papers · 1 benchmark
ARCTIC (Articulated Objects in Free-form Hand Interaction)
ARCTIC is a dataset of free-form interactions of hands and articulated objects.
23 papers · 0 benchmarks
Motion-X is a large-scale 3D expressive whole-body motion dataset, which comprises 15.6M precise 3D whole-body pose annotations (i.e., SMPL-X) covering 81.1K motion sequences from massive scenes, meanwhile providing corresponding semantic…
22 papers · 1 benchmark
AIST++ is a 3D dance dataset which contains 3D motion reconstructed from real dancers paired with music.
21 papers · 2 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
18 papers · 1 benchmark
Inter-X is a large-scale dataset containing ~11K interaction sequences, more than 8.1M frames and 34K fine-grained human textual descriptions.
14 papers · 1 benchmark
AIOZ-GDANCE comprises 16.7 hours of whole-body motion and music audio of group dancing.
7 papers · 1 benchmark
LaFAN1 (Ubisoft La Forge Animation Dataset)
Ubisoft La Forge Animation Dataset ("LAFAN1") Ubisoft La Forge Animation dataset and accompanying code for the SIGGRAPH 2020 paper Robust Motion In-betweening.
7 papers · 1 benchmark
BRACE (The Breakdancing Competition Dataset for Dance Motion Synthesis)
BRACE is a dataset for audio-conditioned dance motion synthesis challenging common assumptions for this task: - strong music-dance correlation - controlled motion data - simple poses and movements To address these issues: - We focus on…
5 papers · 2 benchmarks
A large-scale and diverse duet interactive dance dataset.
5 papers · 0 benchmarks
CHAIRS is a large-scale motion-captured f-AHOI dataset, consisting of 17.3 hours of versatile interactions between 46 participants and 81 articulated and rigid sittable objects.
3 papers · 0 benchmarks
TMD (Text-Music-Dance)
The Text-Music-Dance (TMD) dataset establishes a pioneering benchmark comprising 2,153 text-music-motion pairs.
2 papers · 1 benchmark
Trinity Gesture Dataset includes 23 takes, totalling 244 minutes of motion capture and audio of a male native English speaker producing spontaneous speech on different topics.
2 papers · 2 benchmarks
Data used for the paper SparsePoser: Real-time Full-body Motion Reconstruction from Sparse Data It contains over 1GB of high-quality motion capture data recorded with an Xsens Awinda system while using a variety of VR applications in Meta…
2 papers · 0 benchmarks
BOTH57M is a body-hand dataset with body-level text prompts and finger-level text prompts.
1 paper · 0 benchmarks
Diverse guitar-playing motions about 1 hour long, including: • 12 major scales, • chromatic scales, • diverse chords, • arpeggios, • strumming and picking, • bends, • sliding, • vibrato, • palm mute, • natural harmonics, • artificial…
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.