Browse State-of-the-Art › 3D Face Animation
3D Face Animation
25 papers with code · 3 benchmarks · 6 datasets archive 2025-07-28
Image: Cudeiro et al
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
3 leaderboard tables shown for this task, 3 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| BEAT2 (5 rows) | MambaTalk | MambaTalk: Efficient Holistic Gesture Synthesis with Selective... | code | Syntology ran 1 of 1 samples · 0 unverified | Compare |
| Biwi 3D Audiovisual Corpus of Affective Communication - B3D(AC)^2 (5 rows) | SelfTalk | SelfTalk: A Self-Supervised Commutative Training Diagram to... | code | Syntology ran 3 of 3 samples · 0 unverified | Compare |
| VOCASET (2 rows) | FaceFormer | FaceFormer: Speech-Driven 3D Facial Animation with Transformers | code | Syntology ran 4 of 4 samples · 0 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
6 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
25 shown of 25 papers with code (34 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
1 Nov 2017 9 repositories listedFLAME is low-dimensional but more expressive than the FaceWarehouse model and the Basel Face Model.
-
8 Dec 2022 3 repositories listedThis work addresses the problem of generating 3D holistic body motions from human speech.
-
20 Mar 2023 2 repositories listed Syntology ran 4 of 7 samples · 3 unverified · 7 pointer-only (licence)Specifically, we introduce the emotion disentangling encoder (EDE) to disentangle the emotion and content in the speech by cross-reconstructed speech signals with different emotion labels.
-
16 Apr 2021 2 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)To improve upon existing models, we propose a generic audio-driven facial animation approach that achieves highly realistic motion synthesis results for the entire face.
-
7 Dec 2020 2 repositories listedSome methods produce faces that cannot be realistically animated because they do not model how wrinkles vary with expression.
-
23 Mar 2025 1 repository listedWe further propose a personalizer enhancer during distillation to enhance the influence of embeddings on facial animation.
-
18 Mar 2025 1 repository listed Syntology ran 0 of 3 samples · 3 unverifiedWe begin by defining key capabilities for Physical AI reasoning, with a focus on physical common sense and embodied reasoning.
-
17 Jul 2024 1 repository listedOur approach can generate facial expressions with multiple emotions, and has the ability to generate random yet natural blinks and eye movements, while maintaining accurate lip synchronization.
-
22 Mar 2024 1 repository listedTo this end, we propose a method that can produce a highly stylized 3D face model with desired topology.
-
14 Mar 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedGesture synthesis is a vital realm of human-computer interaction, with wide-ranging applications across various fields like film, robotics, and virtual reality.
-
31 Dec 2023 1 repository listed Syntology ran 4 of 4 samples · 0 unverified · 4 pointer-only (licence)We propose EMAGE, a framework to generate full-body human gestures from audio and masked gestures, encompassing facial, local body, hands, and global movements.
-
13 Dec 2023 1 repository listedWe propose a new latent diffusion model for this task, operating in the expression space of neural parametric head models, to synthesize audio-driven realistic head sequences.
-
20 Sep 2023 1 repository listedIn addition, majority of the approaches focus on 3D vertex based datasets and methods that are compatible with existing facial animation pipelines with rigged characters is scarce.
-
10 Aug 2023 1 repository listedThis paper emphasizes the importance of considering both the composite and regional natures of facial movements in speech-driven 3D face animation.
-
19 Jun 2023 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)To enhance the visual accuracy of generated lip movement while reducing the dependence on labeled data, we propose a novel framework SelfTalk, by involving self-supervision in a cross-modals network system to learn 3D…
-
2 Jun 2023 1 repository listedThis paper presents a novel approach for generating 3D talking heads from raw audio inputs.
-
17 Mar 2023 1 repository listedUpon MMFace4D, we construct a non-autoregressive framework for audio-driven 3D face animation.
-
9 Mar 2023 1 repository listedThis paper presents FaceXHuBERT, a text-less speech-driven 3D facial animation generation method that allows to capture personalized and subtle cues in speech (e.
-
6 Jan 2023 1 repository listed Syntology ran 2 of 13 samples · 11 unverifiedIn this paper, we propose to cast speech-driven facial animation as a code query task in a finite proxy space of the learned codebook, which effectively promotes the vividness of the generated motions by reducing the…
-
12 Sep 2022 1 repository listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)In contrast to the traditional avatar creation pipeline which is a costly process, contemporary generative approaches directly learn the data distribution from photographs.
-
10 Dec 2021 1 repository listed Syntology ran 4 of 4 samples · 0 unverifiedSpeech-driven 3D facial animation is challenging due to the complex geometry of human faces and the limited availability of 3D audio-visual data.
-
18 Aug 2021 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)In this paper, we propose a talking face generation method that takes an audio signal as input and a short target video clip as reference, and synthesizes a photo-realistic video of the target face with natural lip…
-
24 Feb 2020 1 repository listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)In this paper, we address this problem by proposing a deep neural network model that takes an audio signal A of a source person and a very short video V of a target person as input, and outputs a synthesized…
-
8 May 2019 1 repository listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)To address this, we introduce a unique 4D face dataset with about 29 minutes of 4D scans captured at 60 fps and synchronized audio from 12 speakers.
-
24 Jun 2014 1 repository listedThe resulting statistical analysis is applied to automatically generate realistic facial animations and to recognize dynamic facial expressions.
Syntology lines on 12 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections