Home › Datasets › task › Sentiment Analysis
Sentiment Analysis datasets
archive 2025-07-28
104 datasets carry the task tag "Sentiment Analysis" (the task itself: Sentiment Analysis), ordered by the archive's paper count. Page 2 of 3: 48 shown of 104. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Sentiment Analysis datasets 49–96 of 104
UIT-VSMEC (Vietnamese Social Media Emotion Corpus)
Emotion recognition is a higher approach or special case of sentiment analysis.
4 papers · 0 benchmarks
The WikiSem500 dataset contains around 500 per-language cluster groups for English, Spanish, German, Chinese, and Japanese (a total of 13,314 test cases).
4 papers · 0 benchmarks
XED is a multilingual fine-grained emotion dataset.
4 papers · 0 benchmarks
Pars-ABSA is a manually annotated Persian dataset, Pars-ABSA, which is verified by 3 native Persian speakers.
3 papers · 0 benchmarks
RETWEET is a dataset of tweets and overall predominant sentiment of their replies.
3 papers · 2 benchmarks
We develop a primary dataset based on our task of suicide or depression classification.
3 papers · 0 benchmarks
A first-of-its-kind large dataset of sarcastic/non-sarcastic tweets with high-quality labels and extra features: (1) sarcasm perspective labels (2) new contextual features.
3 papers · 0 benchmarks
SPOT (Sentiment Polarity Annotations Dataset)
The SPOT dataset contains 197 reviews originating from the Yelp'13 and IMDB collections ([1][2]), annotated with segment-level polarity labels (positive/neutral/negative).
3 papers · 0 benchmarks
TCAB (Text Classification Attack Benchmark)
Text Classification Attack Benchmark (TCAB) is a dataset for analyzing, understanding, detecting, and labeling adversarial attacks against text classifiers.
3 papers · 0 benchmarks
Youtbean is a dataset created from closed captions of YouTube product review videos.
3 papers · 0 benchmarks
Antonio Gulli’s corpus of news articles is a collection of more than 1 million news articles.
2 papers · 0 benchmarks
AWARE (AWARE: Aspect-Based Sentiment Analysis Dataset of Apps Reviews for Requirements Elicitation)
The peer-reviewed paper of AWARE dataset is published in ASEW 2021, and can be accessed through: http://doi.org/10.1109/ASEW52652.2021.00049.
2 papers · 3 benchmarks
Sentiment detection remains a pivotal task in natural language processing, yet its development in Arabic lags due to a scarcity of training materials compared to English.
2 papers · 0 benchmarks
DBRD (Dutch Book Reviews Dataset)
The DBRD (pronounced dee-bird) dataset contains over 110k book reviews along with associated binary sentiment polarity labels.
2 papers · 1 benchmark
FinnSentiment introduces a 27,000 sentence dataset (in Finnish) annotated independently with sentiment polarity by three native annotators.
2 papers · 0 benchmarks
LEPISZCZE is an open-source comprehensive benchmark for Polish NLP and a continuous-submission leaderboard, concentrating public Polish datasets (existing and new) in specific tasks.
2 papers · 0 benchmarks
LatamXIX (19th Century Latin American Spanish Newspaper Corpus with LLM OCR Correction)
A novel dataset of 19th-century Latin American press texts, which addresses the lack of specialized corpora for historical and linguistic analysis in this region.
2 papers · 0 benchmarks
ReactionGIF is an affective dataset of 30K tweets which can be used for tasks like induced sentiment prediction and multilabel classification of induced emotions.
2 papers · 0 benchmarks
ReviewQA is a question-answering dataset based on hotel reviews.
2 papers · 0 benchmarks
https://github.com/dialogue-evaluation/RuSentNE-evaluation
2 papers · 0 benchmarks
This corpus was constructed by collecting 10,008 reviews from various domains, including sports, food, software, politics, and entertainment.
2 papers · 1 benchmark
In AISIA-VN-Review-S and AISIA-VN-Review-F datasets, we first collect 450K customer reviewing comments from various e–commerce websites.
1 paper · 0 benchmarks
The Amazon Polarity dataset is a set of reviews from Amazon.
1 paper · 1 benchmark
Sentiment analysis is pivotal in Natural Language Processing for understanding opinions and emotions in text.
1 paper · 0 benchmarks
A Bambara dialectal dataset dedicated for Sentiment Analysis, available freely for Natural Language Processing research purposes
1 paper · 0 benchmarks
BanglaBook (Large-scale Bangla Dataset for Sentiment Analysis from Book Reviews)
This repository contains the code, data, and models of the paper titled "BᴀɴɢʟᴀBᴏᴏᴋ: A Large-scale Bangla Dataset for Sentiment Analysis from Book Reviews" published in the Findings of the Association for Computational Linguistics: ACL…
1 paper · 1 benchmark
BanglaEmotion (BanglaEmotion: A Benchmark Dataset for Bangla Textual Emotion Analysis)
BanglaEmotion is a manually annotated Bangla Emotion corpus, which incorporates the diversity of fine-grained emotion expressions in social-media text.
1 paper · 0 benchmarks
The Covid19-CountryImage dataset is a Twitter dataset which contains COVID-19-related tweets.
1 paper · 0 benchmarks
Capriccio is a sentiment classification dataset on tweets that simulates data drift.
1 paper · 0 benchmarks
A dataset of games played in the card game "Cards Against Humanity" (CAH), by human players, derived from the online CAH labs.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Fallout New Vegas Dialog is a multilingual sentiment annotated dialog dataset from Fallout New Vegas.
1 paper · 0 benchmarks
Financial Language Understanding Evaluation is an open-source comprehensive suite of benchmarks for the financial domain.
1 paper · 0 benchmarks
This dataset contains news headlines relevant to key forex pairs: AUDUSD, EURCHF, EURUSD, GBPUSD, and USDJPY.
1 paper · 0 benchmarks
HARD (Hotel Arabic-Reviews Dataset)
The Hotel Arabic-Reviews Dataset (HARD) contains 93700 hotel reviews in Arabic language.
1 paper · 1 benchmark
L3Cube-MahaCorpus is a Marathi monolingual data set scraped from different internet sources.
1 paper · 0 benchmarks
LSICC (Large Scale Informal Chinese Corpus)
Large Scale Informal Chinese Corpus (LSICC) is a large-scale corpus of informal Chinese.
1 paper · 0 benchmarks
MalayalamMixSentiment is a Sentiment Analysis Dataset for Code-Mixed Malayalam-English.
1 paper · 0 benchmarks
Modern Hebrew Sentiment Dataset is a sentiment analysis benchmark for Hebrew, based on 12K social media comments, and provide two instances of these data: in token-based and morpheme-based settings.
1 paper · 0 benchmarks
MultiSenti presents a labeled dataset called MultiSenti for sentiment classification of code-switched informal short text, (2) explore the feasibility of adapting resources from a resource-rich language for an informal one, and (3) propose…
1 paper · 0 benchmarks
NAIST COVID is a multilingual dataset of social media posts related to COVID-19, consisting of microblogs in English and Japanese from Twitter and those in Chinese from Weibo.
1 paper · 0 benchmarks
NoReCfine is a dataset for fine-grained sentiment analysis in Norwegian, annotated with respect to polar expressions, targets and holders of opinion.
1 paper · 0 benchmarks
ROAST (Review level Opinion Aspect Sentiment Target Joint Detection for ABSA)
This repository has a review-level multidomain multilingual dataset for Aspect-based Sentiment Analysis(ABSA) for the paper ROAST: Review-level Opinion Aspect Sentiment Target Joint Detection.
1 paper · 0 benchmarks
https://arxiv.org/abs/2503.15222
1 paper · 0 benchmarks
https://github.com/dialogue-evaluation/RuOpinionNE-2024
1 paper · 0 benchmarks
SAIL 2017 (Sentiment Analysis for Indian Languages)
India is a linguistic area with one of the longest histories of contact, influence, use, teaching and learning of English-in-diaspora in the world (Kachru and Nelson, 2006).
1 paper · 1 benchmark
The Sequence labellIng evaLuatIon benChmark fOr spoken laNguagE (SILICONE) benchmark is a collection of resources for training, evaluating, and analyzing natural language understanding systems specifically designed for spoken language.
1 paper · 1 benchmark
SVLD (Social Vision and Language Dataset)
The social vision and language dataset is a large-scale multimodal dataset designed for research into social contextual learning.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.