Home › Datasets › task › Recommendation Systems
Recommendation Systems datasets
archive 2025-07-28
56 datasets carry the task tag "Recommendation Systems" (the task itself: Recommendation Systems), ordered by the archive's paper count. Page 1 of 2: 48 shown of 56. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Recommendation Systems datasets 1–48 of 56
The MovieLens datasets, first released in 1998, describe people’s expressed preferences for movies.
1,246 papers · 17 benchmarks
Netflix Prize consists of about 100,000,000 ratings for 17,770 movies given by 480,189 users.
370 papers · 1 benchmark
Gowalla is a location-based social networking website where users share their locations by checking-in.
203 papers · 4 benchmarks
The Yelp2018 dataset is adopted from the 2018 edition of the yelp challenge.
125 papers · 2 benchmarks
ReDial (Recommendation Dialogues) is an annotated dataset of dialogues, where users recommend movies to each other.
105 papers · 2 benchmarks
The Yelp Dataset is a valuable resource for academic research, teaching, and learning.
86 papers · 15 benchmarks
Douban (Douban Conversation Corpus)
We release Douban Conversation Corpus, comprising a training data set, a development set and a test set for retrieval based chatbot.
81 papers · 4 benchmarks
This dataset contains 21,889 outfits from polyvore.com, in which 17,316 are for training, 1,497 for validation and 3,076 for testing.
62 papers · 3 benchmarks
The Epinions dataset is built form a who-trust-whom online social network of a general consumer review site Epinions.com.
54 papers · 2 benchmarks
KuaiRec is a real-world dataset collected from the recommendation logs of the video-sharing mobile app Kuaishou.
51 papers · 0 benchmarks
The Foursquare dataset consists of check-in data for different cities.
49 papers · 0 benchmarks
KuaiRand is an unbiased sequential recommendation dataset collected from the recommendation logs of the video-sharing mobile app, Kuaishou (快手).
42 papers · 1 benchmark
This dataset contains product reviews and metadata from Amazon, including 142.8 million reviews spanning May 1996 - July 2014.
41 papers · 5 benchmarks
The Memetracker corpus contains articles from mainstream media and blogs from August 1 to October 31, 2008 with about 1 million documents per day.
39 papers · 1 benchmark
GVGAI (General Video Game AI)
The General Video Game AI (GVGAI) framework is widely used in research which features a corpus of over 100 single-player games and 60 two-player games.
36 papers · 0 benchmarks
Amazon Review is a dataset to tackle the task of identifying whether the sentiment of a product review is positive or negative.
35 papers · 1 benchmark
The Ciao dataset contains rating information of users given to items, and also contain item category information.
34 papers · 1 benchmark
The Pinterest dataset contains more than 1 million images associated to Pinterest users’ who have “pinned” them.
33 papers · 1 benchmark
A human-to-human Chinese dialog dataset (about 10k dialogs, 156k utterances), which contains multiple sequential dialogs for every pair of a recommendation seeker (user) and a recommender (bot).
28 papers · 0 benchmarks
TG-ReDial is a a topic-guided conversational recommendation dataset for research on conversational/interactive recommender systems.
27 papers · 0 benchmarks
This dataset includes reviews (ratings, text, helpfulness votes), product metadata (descriptions, category information, price, brand, and image features), and links (also viewed/also bought graphs).
18 papers · 3 benchmarks
The MMD (MultiModal Dialogs) dataset is a dataset for multimodal domain-aware conversations.
18 papers · 0 benchmarks
BeerAdvocate is a dataset that consists of beer reviews from beeradvocate.
15 papers · 1 benchmark
A dataset containing 404,683 shop photos collected from 25 different online retailers and 20,357 street photos, providing a total of 39,479 clothing item matches between street and shop photos.
15 papers · 1 benchmark
This datasets is a subset of the Amazon reviews dataset which contain Men related products
14 papers · 2 benchmarks
Amazon-Sports is a sub-category of the Amazon dataset, which contains a series of product reviews crawled from Amazon.com.
14 papers · 1 benchmark
MSD (Million Song Dataset)
The Million Song Dataset is a freely-available collection of audio features and metadata for a million contemporary popular music tracks.
11 papers · 2 benchmarks
This datasets is a subset of the Amazon reviews dataset which contain Fashion related products
8 papers · 1 benchmark
The WeChat dataset for fake news detection contains more than 20k news labelled as fake news or not.
7 papers · 1 benchmark
CITE is a crowd-sourced resource for multimodal discourse: this resource characterises inferences in image-text contexts in the domain of cooking recipes in the form of coherence relations.
6 papers · 1 benchmark
A set of approximately 100K podcast episodes comprised of raw audio files along with accompanying ASR transcripts.
6 papers · 0 benchmarks
Coached Conversational Preference Elicitation is a dataset consisting of 502 English dialogs with 12,000 annotated utterances between a user and an assistant discussing movie preferences in natural language.
5 papers · 0 benchmarks
The Epinions dataset is trust network dataset.
5 papers · 1 benchmark
KG20C (A scholarly knowledge graph benchmark dataset)
KG20C is a Knowledge Graph about high quality papers from 20 top computer science Conferences.
5 papers · 1 benchmark
Amazon Fine Foods is a dataset that consists of reviews of fine foods from amazon.
4 papers · 0 benchmarks
Publicly available dataset in the hotel domain (50M versus 0.9M) and additionally, the largest recommendation dataset in a single domain and with textual reviews (50M versus 22M).
3 papers · 0 benchmarks
We ran 21 recommender systems on three datasets (BeerAdvocate, LibraryThing and MovieLens 1M).
3 papers · 0 benchmarks
Dataset of restaurant reviews from TripAdvisor that includes images and texts uploaded in reviews by users.
3 papers · 0 benchmarks
xMIND (A Multilingual Dataset for Cross-lingual News Recommendation)
xMIND is an open, large-scale multilingual news dataset for multi- and cross-lingual news recommendation.
3 papers · 0 benchmarks
Delicious : This data set contains tagged web pages retrieved from the website delicious.com.
2 papers · 1 benchmark
E-ReDial (Explainable Recommendation Dialogues)
E-ReDial is a conversational recommender system dataset with high-quality explanations.
2 papers · 0 benchmarks
A dataset consisting of recipient 46 users and, 26180 tweets.
2 papers · 0 benchmarks
Uses a platform with 77 candies and sweets to rank.
2 papers · 0 benchmarks
Wikidata-14M is a recommender system dataset for recommending items to Wikidata editors.
2 papers · 0 benchmarks
Wyze Rule Recommendation Dataset.
2 papers · 0 benchmarks
CAL10K (Computer Audition Lab 10000)
The CAL10K dataset (introduced as Swat10k) contains 10,870 songs that are weakly-labelled using a tag vocabulary of 475 acoustic tags and 153 genre tags.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.