Home › Datasets › task › Emotion Recognition
Emotion Recognition datasets
archive 2025-07-28
52 datasets carry the task tag "Emotion Recognition" (the task itself: Emotion Recognition), ordered by the archive's paper count. Page 1 of 2: 48 shown of 52. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Emotion Recognition datasets 1–48 of 52
The Hateful Memes data set is a multimodal dataset for hateful meme detection (image + text) that contains 10,000+ new multimodal examples created by Facebook AI.
177 papers · 3 benchmarks
FER2013 (Facial Expression Recognition 2013 Dataset)
Fer2013 contains approximately 30,000 facial RGB images of different expressions with size restricted to 48×48, and the main labels of it can be divided into 7 types: 0=Angry, 1=Disgust, 2=Fear, 3=Happy, 4=Sad, 5=Surprise, 6=Neutral.
168 papers · 5 benchmarks
Aff-Wild2 is a large-scale in-the-wild database and an extension of the Aff-Wild dataset for affect recognition.
142 papers · 2 benchmarks
Aff-Wild is a large-scale in-the-wild dataset for valence-arousal estimation from videos with a variety of head poses, illumination conditions and occlusions.
125 papers · 0 benchmarks
CARER (Contextualized Affect Representations for Emotion Recognition)
CARER is an emotion dataset collected through noisy labels, annotated via distant supervision as in (Go et al., 2009).
123 papers · 2 benchmarks
SEED (SJTU Emotion EEG Dataset)
The SEED dataset contains subjects' EEG signals when they were watching films clips.
119 papers · 4 benchmarks
MSP-IMPROV (MSP-IMPROV: An Acted Corpus of Dyadic Interactions to Study Emotion Perception)
We present the MSP-IMPROV corpus, a multimodal emotional database, where the goal is to have control over lexical content and emotion while also promoting naturalness in the recordings.
70 papers · 1 benchmark
ISEAR (International Survey on Emotion Antecedents and Reactions)
Over a period of many years during the 1990s, a large group of psychologists all over the world collected data in the ISEAR project, directed by Klaus R.
55 papers · 0 benchmarks
A new challenge set for multimodal classification, focusing on detecting hate speech in multimodal memes.
48 papers · 0 benchmarks
EmotionLines contains a total of 29245 labeled utterances from 2000 dialogues.
44 papers · 1 benchmark
ExpW (Expression in-the-Wild)
The Expression in-the-Wild (ExpW) dataset is for facial expression recognition and contains 91,793 faces manually labeled with expressions.
41 papers · 1 benchmark
The EMOTIC dataset, named after EMOTions In Context, is a database of images with people in real environments, annotated with their apparent emotions.
38 papers · 2 benchmarks
EmoBank is a corpus of 10k English sentences balancing multiple genres, annotated with dimensional emotion metadata in the Valence-Arousal-Dominance (VAD) representation format.
32 papers · 0 benchmarks
MAFW is a large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.
31 papers · 2 benchmarks
RAVDESS (Ryerson Audio-Visual Database of Emotional Speech and Song)
The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS) contains 7,356 files (total size: 24.8 GB).
27 papers · 6 benchmarks
EMOPIA (A Multi-Modal Pop Piano Dataset For Emotion Recognition and Emotion-based Music Generation)
EMOPIA (pronounced ‘yee-mò-pi-uh’) dataset is a shared multi-modal (audio and MIDI) database focusing on perceived emotion in pop piano music, to facilitate research on various tasks related to music emotion.
23 papers · 0 benchmarks
CPED (Chinese Personalized and Emotional Dialogue)
We construct a dataset named CPED from 40 Chinese TV shows.
15 papers · 3 benchmarks
The DEAP dataset consists of two parts: - The ratings from an online self-assessment where 120 one-minute extracts of music videos were each rated by 14-16 volunteers based on arousal, valence and dominance.
11 papers · 1 benchmark
The Video2GIF dataset contains over 100,000 pairs of GIFs and their source videos.
11 papers · 0 benchmarks
EmoWOZ is the first large-scale open-source dataset for emotion recognition in task-oriented dialogues.
10 papers · 2 benchmarks
MUStARD++ is a multimodal sarcasm detection dataset (MUStARD) pre-annotated with 9 emotions.
10 papers · 1 benchmark
A multimodal dataset with comprehensive annotations of continuous emotions during naturalistic conversations.
9 papers · 0 benchmarks
MSP-Podcast (A large naturalistic speech emotional dataset)
The MSP-Podcast corpus contains speech segments from podcast recordings which are perceptually annotated using crowdsourcing.
9 papers · 4 benchmarks
SEND (Stanford Emotional Narratives Dataset)
SEND (Stanford Emotional Narratives Dataset) is a set of rich, multimodal videos of self-paced, unscripted emotional narratives, annotated for emotional valence over time.
8 papers · 0 benchmarks
ArmanEmo is a human-labeled emotion dataset of more than 7000 Persian sentences labeled for seven categories.
6 papers · 1 benchmark
Ulm-TSST (Ulm-Trier Social Stress Dataset)
Ulm-TSST is a dataset continuous emotion (valence and arousal) prediction and physiological-emotion' prediction.
6 papers · 0 benchmarks
Emotional Dialogue Acts data contains dialogue act labels for existing emotion multi-modal conversational datasets.
3 papers · 0 benchmarks
FindingEmo is an image dataset containing annotations for 25k images, specifically tailored to Emotion Recognition.
3 papers · 0 benchmarks
HUME-VB (The Hume Vocal Bursts Dataset)
The Hume Vocal Burst Database (H-VB) includes all train, validation, and test recordings and corresponding emotion ratings for the train and validation recordings.
3 papers · 7 benchmarks
A large-scale dataset of memes with captions and class labels.
3 papers · 0 benchmarks
A database of more than 2000 minutes of audio-visual data of 398 people coming from six cultures, 50% female, and uniformly spanning the age range of 18 to 65 years old.
3 papers · 0 benchmarks
Dusha (Dusha Crowd, Dusha Podcast)
Dusha is a dataset for speech emotion recognition (SER) tasks.
2 papers · 2 benchmarks
EmoPars is a dataset of 30,000 Persian Tweets labeled with Ekman’s six basic emotions (Anger, Fear, Happiness, Sadness, Hatred, and Wonder).
2 papers · 0 benchmarks
1000 songs has been selected from Free Music Archive (FMA).
2 papers · 1 benchmark
The E-MASAC Dataset is a collection of code-mixed conversations sourced from an Indian TV series, focusing on Hindi-English interactions.
2 papers · 1 benchmark
The Sentimental LIAR dataset is a modified and further extended version of the LIAR extension introduced by Kirilin et al.
2 papers · 0 benchmarks
Overview nEMO is a simulated dataset of emotional speech in the Polish language.
2 papers · 0 benchmarks
BERSt (Basic Emotion Random phrase Shouts)
BERSt Dataset We release the BERSt Dataset for various speech recognition tasks including Automatic Speech Recognition (ASR) and Speech Emotion Recogniton (SER) Overview 4526 single phrase recordings (~3.75h) 98 professional actors 19…
1 paper · 1 benchmark
BanglaEmotion (BanglaEmotion: A Benchmark Dataset for Bangla Textual Emotion Analysis)
BanglaEmotion is a manually annotated Bangla Emotion corpus, which incorporates the diversity of fine-grained emotion expressions in social-media text.
1 paper · 0 benchmarks
EmoFilm (Emotional speech from Films)
EmoFilm is a multilingual emotional speech corpus comprising 1115 audio instances produced in English, Italian, and Spanish languages.
1 paper · 0 benchmarks
A corpus designed in analogy to the well-established English ISEAR emotion dataset.
1 paper · 0 benchmarks
MFA (Many Faces of Anger)
The MFA (Many Faces of Anger) dataset includes 200 in-the-wild videos from North American and Persian cultures with fine-grained labels of: 'annoyed', 'anger', 'disgust', 'hatred' and 'furious' and 13 related emojis.
1 paper · 1 benchmark
Reader Emotion News 20k Dataset
1 paper · 0 benchmarks
VESUS (Varied Emotion in Syntactically Uniform Speech)
The Varied Emotion in Syntactically Uniform Speech (VESUS) repository is a lexically controlled database collected by the NSA lab.
1 paper · 0 benchmarks
VNEMOS (Vietnamese Speech Emotion Dataset)
This research introduces the dataset that we created to test voice emotional recognition models with Vietnamese data.
1 paper · 0 benchmarks
Existing databases usually record posed or induced human behavior in individual or dyadic settings with biased annotations in which a basic emotional class label or a Valence–Arousal pair value represents the emotional states.
1 paper · 0 benchmarks
YM2413-MDB is an 80s FM video game music dataset with multi-label emotion annotations.
1 paper · 0 benchmarks
A modification on the ShEMO dataset with help of an Automatic Speech Recognition (ASR) system.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.