Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 51 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 2401–2448 of 12,172
ACRE (Abstract Causal REasoning)
Abstract Causal REasoning (ACRE) is a dataset for the systematic evaluation of current vision systems in causal induction, i.e., identifying unobservable mechanisms that lead to the observable relations among variables.
15 papers · 0 benchmarks
ACSPublicCoverage: predict whether an individual is covered by public health insurance, after filtering the ACS PUMS data sample to only include individuals under the age of 65, and those with an income of less than $30,000.
15 papers · 0 benchmarks
10 classes with 50, 000 training and 5, 000 testing images.
15 papers · 1 benchmark
ArSarcasm-v2 is an extension of the original ArSarcasm dataset published along with the paper From Arabic Sentiment Analysis to Sarcasm Detection: The ArSarcasm Dataset.
15 papers · 0 benchmarks
BIRD (BIg Bench for LaRge-scale Database Grounded Text-to-SQL Evaluation) represents a pioneering, cross-domain dataset that examines the impact of extensive database contents on text-to-SQL parsing.
15 papers · 1 benchmark
Contains 800 sequences at various spatial resolutions from 270p to 2160p and has been evaluated on ten existing network architectures for four different coding tools.
15 papers · 0 benchmarks
BeerAdvocate is a dataset that consists of beer reviews from beeradvocate.
15 papers · 1 benchmark
Breaking Bad is a large-scale dataset of fractured objects.
15 papers · 0 benchmarks
BugSwarm is a dataset of reproducible faults and fixes to perform experimental evaluation of approaches to software quality.
15 papers · 0 benchmarks
CASIA V2 is a dataset for forgery classification.
15 papers · 0 benchmarks
CCVID (Clothes-Changing Video person re-ID)
Clothes-Changing Video person re-ID (CCVID) is a dataset constructed from the raw data of a gait recognition dataset, i.e.
15 papers · 1 benchmark
CED (Color Event Camera Dataset)
Contains 50 minutes of footage with both color frames and events.
15 papers · 0 benchmarks
CMU DoG (CMU Document Grounded Conversations Dataset)
This is a document grounded dataset for text conversations.
15 papers · 0 benchmarks
CPED (Chinese Personalized and Emotional Dialogue)
We construct a dataset named CPED from 40 Chinese TV shows.
15 papers · 3 benchmarks
Data was collected for normal bearings, single-point drive end and fan end defects.
15 papers · 1 benchmark
CelebV-Text comprises 70,000 in-the-wild face video clips with diverse visual content, each paired with 20 texts generated using the proposed semi-automatic text generation strategy.
15 papers · 0 benchmarks
Node classification on Citeseer with the fixed 48%/32%/20% splits provided by Geom-GCN.
15 papers · 1 benchmark
CitySim Dataset (A Drone-Based Vehicle Trajectory Dataset for Safety Oriented Research and Digital Twins)
The development of safety-oriented research ideas and applications requires fine-grained vehicle trajectory data that not only has high accuracy but also captures a substantial number of critical safety events.
15 papers · 0 benchmarks
A large-scale video dataset, featuring clips from movies with detailed captions.
15 papers · 1 benchmark
Continual World is a benchmark consisting of realistic and meaningfully diverse robotic tasks built on top of Meta-World as a testbed.
15 papers · 0 benchmarks
Node classification on Cora with the fixed 48%/32%/20% splits provided by Geom-GCN.
15 papers · 1 benchmark
Countix is a real world dataset of repetition videos collected in the wild (i.e.YouTube) covering a wide range of semantic settings with significant challenges such as camera and object motion, diverse set of periods and counts, and…
15 papers · 1 benchmark
The beginnings of a question answering dataset specifically designed for COVID-19, built by hand from knowledge gathered from Kaggle's COVID-19 Open Research Dataset Challenge.
15 papers · 0 benchmarks
CrossNER is a cross-domain NER (Named Entity Recognition) dataset, a fully-labeled collection of NER data spanning over five diverse domains (Politics, Natural Science, Music, Literature, and Artificial Intelligence) with specialized…
15 papers · 1 benchmark
CustomHumans is recorded by a multi-view photogrammetry system equipped with 53 RGB (12 Megapixels) and 53 (4 Megapixels) IR cameras.
15 papers · 1 benchmark
Dataset consisting of IMU measurements and corresponding SMPL poses.
15 papers · 0 benchmarks
The database consists of 150 annotated pages of three different medieval manuscripts with challenging layouts.
15 papers · 2 benchmarks
The DUC2004 dataset is a dataset for document summarization.
15 papers · 4 benchmarks
A novel benchmark dataset that includes a manually annotated point cloud for over 260 million laser scanning points into 100'000 (approx.) assets from Dublin LiDAR point cloud [12] in 2015.
15 papers · 0 benchmarks
This data set contains electricity consumption of 370 points/clients.
15 papers · 7 benchmarks
Everybody Dance Now is a dataset of videos that can be used for training and motion transfer.
15 papers · 0 benchmarks
A dataset containing 404,683 shop photos collected from 25 different online retailers and 20,357 street photos, providing a total of 39,479 clothing item matches between street and shop photos.
15 papers · 1 benchmark
A dataset of over 24,000 images exhibiting the broadest range of exposure values to date with a corresponding properly exposed image.
15 papers · 1 benchmark
FaVIQ (Fact Verification from Information-seeking Questions)
FaVIQ (Fact Verification from Information-seeking Questions) is a challenging and realistic fact verification dataset that reflects confusions raised by real users.
15 papers · 0 benchmarks
Fakeddit is a novel multimodal dataset for fake news detection consisting of over 1 million samples from multiple categories of fake news.
15 papers · 0 benchmarks
First-Person Hand Action Benchmark is a collection of RGB-D video sequences comprised of more than 100K frames of 45 daily hand action categories, involving 26 different objects in several hand configurations.
15 papers · 2 benchmarks
The gtzan8 audio dataset contains 1000 tracks of 30 second length.
15 papers · 4 benchmarks
The GoodsAD dataset contains 6124 images with 6 categories of common supermarket goods.
15 papers · 1 benchmark
HAKE is built upon existing activity datasets and provides human body part level atomic action labels (Part States).
15 papers · 0 benchmarks
HaGRID (HaGRID - HAnd Gesture Recognition Image Dataset)
We introduce a large image dataset HaGRID (HAnd Gesture Recognition Image Dataset) for hand gesture recognition (HGR) systems.
15 papers · 0 benchmarks
Extension test cases of HumanEval, as well as generated code.
15 papers · 1 benchmark
Humicroedit is a humorous headline dataset.
15 papers · 0 benchmarks
The ISIC 2017 dataset was published by the International Skin Imaging Collaboration (ISIC) as a large-scale dataset of dermoscopy images.
15 papers · 0 benchmarks
ISTD+ consists of shadow images, shadow-free images, and shadow masks, with 1,330 training images and 540 testing images from 135 unique background scenes.
15 papers · 1 benchmark
JEEBench is a considerably more challenging benchmark dataset for evaluating the problem solving abilities of LLMs.
15 papers · 0 benchmarks
The dataset is constructed from images of defective production items that were provided and annotated by Kolektor Group d.o.o..
15 papers · 1 benchmark
KolektorSDD2 is a surface-defect detection dataset with over 3000 images containing several types of defects, obtained while addressing a real-world industrial problem.
15 papers · 2 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.