Browse State-of-the-Art › Motion Synthesis
Motion Synthesis
126 papers with code · 13 benchmarks · 19 datasets archive 2025-07-28
Creating a video where people in the images move (such as blinking or smiling) requires specialized AI technology like Deepfake or Motion Synthesis, which I cannot do directly here.
However, if you want a video made from the images you provided, I can create an animated slideshow with cute effects and royalty-free background music. Would you be interested in this?
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
13 leaderboard tables shown for this task, 13 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 13 until expanded.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
19 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
3 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 126 papers with code (282 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
6 May 2017 8 repositories listed Syntology ran 3 of 5 samples · 2 unverified · 3 pointer-only (licence)Human motion modelling is a classical problem at the intersection of graphics and computer vision, with applications spanning human-computer interaction, motion synthesis, and motion prediction for virtual and augmented…
-
8 Apr 2018 6 repositories listedWe further explore a number of methods for integrating multiple clips into the learning process to develop multi-skilled agents capable of performing a rich repertoire of diverse skills.
-
15 Nov 2021 3 repositories listedIn this work, we follow the trend of rendering the NIMAT effect by introducing a modification on the blur synthesis procedure in portrait mode.
-
16 May 2019 3 repositories listed Syntology ran 1 of 2 samples · 1 unverified · 1 pointer-only (licence)Data-driven modelling and synthesis of motion is an active research area with applications that include animation, games, and social robotics.
-
27 Nov 2017 3 repositories listedOur model, which we call HP-GAN, learns a probability density function of future human poses conditioned on previous poses.
-
11 Mar 2025 2 repositories listedDespite recent advancements in learning-based motion in-betweening, a key limitation has been overlooked: the requirement for character-specific datasets.
-
8 Jun 2024 2 repositories listed Syntology ran 6 of 15 samples · 9 unverified · 15 pointer-only (licence)Motion-based controllable video generation offers the potential for creating captivating visual content.
-
2 Mar 2023 2 repositories listed Syntology ran 1 of 4 samples · 3 unverified · 2 pointer-only (licence)We evaluate the composition methods using an off-the-shelf motion diffusion model, and further compare the results to dedicated models trained for these specific tasks.
-
31 Aug 2022 2 repositories listedInstead of a deterministic language-motion mapping, MotionDiffuse generates motions through a series of denoising steps in which variations are injected.
-
16 Apr 2021 2 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)To improve upon existing models, we propose a generic audio-driven facial animation approach that achieves highly realistic motion synthesis results for the entire face.
-
12 Apr 2021 2 repositories listed Syntology ran 0 of 5 samples · 5 unverifiedBy sampling from this latent space and querying a certain duration through a series of positional encodings, we synthesize variable-length motion sequences conditioned on a categorical action.
-
5 Nov 2019 2 repositories listedIn the analysis phase, we decompose a dance into a series of basic dance units, through which the model learns how to move.
-
3 Jul 2025 1 repository listedAlong with the explosion of large language models, improvements in speech synthesis, advancements in hardware, and the evolution of computer graphics, the current bottleneck in creating digital humans lies in generating…
-
29 Jun 2025 1 repository listedParametric human body models play a crucial role in computer graphics and vision, enabling applications ranging from human motion analysis to understanding human-environment interactions.
-
23 Jun 2025 1 repository listedSubsequently, in the second stage, two generative masked transformers learn to map music signals to these dance tokens: the first producing high-level semantic tokens, and the second, conditioned on music and these…
-
27 Apr 2025 1 repository listedThis survey is intended as a resource for researchers and developers entering the field of generative AI animation or adjacent fields.
-
18 Mar 2025 1 repository listedIn this work, we present a model-free motion imitation framework (KINESIS) to advance the understanding of muscle-based motor control.
-
10 Mar 2025 1 repository listedConditional motion generation has been extensively studied in computer vision, yet two critical challenges remain.
-
30 Jan 2025 1 repository listedRapid progress in text-to-motion generation has been largely driven by diffusion models.
-
25 Jan 2025 1 repository listedAmong the various conditions of motion generation, text can describe motion details elaborately and is easy to acquire, making text-to-motion(T2M) generation important.
-
26 Nov 2024 1 repository listedThis paper introduces MotionLLaMA, a unified framework for motion synthesis and comprehension, along with a novel full-body motion tokenizer called the HoMi Tokenizer.
-
31 Oct 2024 1 repository listed Syntology ran 0 of 11 samples · 11 unverifiedThis issue stems from the internal biases in text encoding, which overlooks motions, and inadequate conditioning mechanisms in T2V generation models.
-
14 Oct 2024 1 repository listed Syntology ran 4 of 6 samples · 2 unverifiedRecent advancements in human motion synthesis have focused on specific types of motions, such as human-scene interaction, locomotion or human-human interaction, however, there is a lack of a unified system capable of…
-
13 Oct 2024 1 repository listedIn this work, we introduce InterMask, a novel framework for generating human interactions using collaborative masked modeling in discrete space.
-
17 Sep 2024 1 repository listedAutoregressive models excel in modeling sequential dependencies by enforcing causal constraints, yet they struggle to capture complex bidirectional patterns due to their unidirectional nature.
-
14 Sep 2024 1 repository listedIn the domain of photorealistic avatar generation, the fidelity of audio-driven lip motion synthesis is essential for realistic virtual interactions.
-
23 Aug 2024 1 repository listedSpeech-driven 3D motion synthesis seeks to create lifelike animations based on human speech, with potential uses in virtual reality, gaming, and the film production.
-
30 Jul 2024 1 repository listedHowever, employing a unified model to achieve various generation tasks with different condition modalities presents two main challenges: motion distribution drifts across different tasks (e.
-
16 Jul 2024 1 repository listedThe target duration of a synthesized human motion is a critical attribute that requires modeling control over the motion dynamics and style.
-
10 Jun 2024 1 repository listedGiven the remarkable results of motion synthesis with diffusion models, a natural question arises: how can we effectively leverage these models for motion editing?
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections