Browse State-of-the-Art › Emotion Recognition
Emotion Recognition
614 papers with code · 9 benchmarks · 52 datasets archive 2025-07-28
Emotion Recognition is an important area of research to enable effective human-computer interaction. Human emotions can be detected using speech signal, facial expressions, body language, and electroencephalography (EEG). Source: Using Deep Autoencoders for Facial Expression Recognition
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
10 leaderboard tables shown for this task, 9 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
52 datasets whose archive record lists this task, ordered by the archive's paper count. 30 shown of 52 until expanded.
Subtasks archive 2025-07-28
12 subtasks in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 614 papers with code (2,041 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
5 Oct 2018 8 repositories listedWe propose several strong multimodal baselines and show the importance of contextual and multimodal information for emotion recognition in conversations.
-
18 Jul 2019 5 repositories listedFinally, investigations on the neuronal activities reveal important brain regions and inter-channel relations for EEG-based emotion recognition.
-
12 Apr 2019 5 repositories listedIn this work, we adopt a feature-engineering based approach to tackle the task of speech emotion recognition.
-
23 Nov 2018 5 repositories listed Syntology ran 0 of 2 samples · 2 unverifiedHumans convey their intentions through the usage of both verbal and nonverbal behaviors during face-to-face communication.
-
16 Dec 2020 4 repositories listedSpecifically, we first modify the recurrence mechanism of XLNet from segment-level to utterance-level in order to better model the conversational data.
-
4 Jun 2019 4 repositories listedEmotion cause extraction (ECE), the task aimed at extracting the potential causes behind certain emotions in text, has gained much attention in recent years due to its wide applications.
-
10 Oct 2018 4 repositories listedSpeech emotion recognition is a challenging task, and extensive reliance has been placed on models that use audio features in building well-performing classifiers.
-
4 Jul 2024 3 repositories listed Syntology ran 10 of 10 samples · 0 unverifiedThis report introduces FunAudioLLM, a model family designed to enhance natural voice interactions between humans and large language models (LLMs).
-
22 Jun 2023 3 repositories listedSpeech Emotion Recognition (SER) typically relies on utterance-level solutions.
-
18 Apr 2023 3 repositories listedThe first Multimodal Emotion Recognition Challenge (MER 2023) was successfully held at ACM Multimedia.
-
19 Oct 2021 3 repositories listed Syntology ran 11 of 16 samples · 5 unverifiedHowever, pure Transformer models tend to require more training data compared to CNNs, and the success of the AST relies on supervised pretraining that requires a large amount of labeled data and a complex training…
-
17 Sep 2021 3 repositories listedThe objective of the sheet is to facilitate and encourage more thoughtfulness on why to automate, how to automate, and how to judge success well before the building of AER systems.
-
2 Jul 2021 3 repositories listedI will also present a template for ethics sheets with 50 ethical considerations, using the task of emotion recognition as a running example.
-
5 Aug 2020 3 repositories listedWe propose a deep graph approach to address the task of speech emotion recognition.
-
30 Mar 2020 3 repositories listedIn this paper we present EMOTIC, a dataset of images of people in a diverse set of natural situations, annotated with their apparent emotion.
-
10 Feb 2020 3 repositories listed Syntology ran 2 of 7 samples · 5 unverifiedWe use the soft labels and the ground truth to train the student model.
-
17 Apr 2019 3 repositories listedTherefore, in this paper, based on audio and text, we consider the task of multimodal sentiment analysis and propose a novel fusion strategy including both multi-feature fusion and multi-modality fusion to improve the…
-
31 May 2018 3 repositories listedPrevious research in this field has exploited the expressiveness of tensors for multimodal representation.
-
17 Sep 2015 3 repositories listedThe proposed architecture achieves 99.
-
20 Dec 2014 3 repositories listedOn MNIST handwritten digits, we show that our model is robust to label corruption.
-
6 Jul 2025 2 repositories listedTo address this gap, we present and release a large-scale rPPG dataset collected under dynamic lighting conditions at night, named DLCN.
-
23 May 2025 2 repositories listed Syntology ran 1 of 4 samples · 3 unverifiedDespite these advancements, CosyVoice 2 exhibits limitations in language coverage, domain diversity, data volume, text formats, and post-training techniques.
-
10 Apr 2025 2 repositories listed Syntology ran 1 of 9 samples · 8 unverifiedMost existing emotion analysis emphasizes which emotion arises (e.
-
28 Mar 2025 2 repositories listedIn the second stage, it learns CLAP features using the audio features learned from the LLM-based embeddings.
-
30 Dec 2024 2 repositories listedFace recognition has witnessed remarkable advancements in recent years, thanks to the development of deep learning techniques.
-
19 Dec 2024 2 repositories listedIn addition, we propose a new prototypical diversity amplification loss to strengthen the model's capacity by amplifying the differences between different prototypes.
-
11 Dec 2024 2 repositories listedWhile Multimodal Large Language Models (MLLMs) demonstrate robust general capabilities, they face considerable challenges in the field of affective computing, particularly in detecting subtle facial expressions and…
-
13 Oct 2024 2 repositories listedEEG-based emotion recognition (EER) has gained significant attention due to its potential for understanding and analyzing human emotions.
-
23 Sep 2024 2 repositories listedThe complex nature of musical emotion introduces inherent bias in both recognition and generation, particularly when relying on a single audio encoder, emotion classifier, or evaluation metric.
-
17 Jul 2024 2 repositories listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)Our results indicate that multimodal textualization provides lower accuracy than feature-based models on C-EXPR-DB, where text transcripts are captured in the wild.
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections