Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 33 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1537–1584 of 12,172
V2X-Sim, short for vehicle-to-everything simulation, is the a synthetic collaborative perception dataset in autonomous driving developed by AI4CE Lab at NYU and MediaBrain Group at SJTU to facilitate collaborative perception between…
32 papers · 1 benchmark
A large-scale VIdeo Panoptic Segmentation dataset
32 papers · 1 benchmark
Worldtree is a corpus of explanation graphs, explanatory role ratings, and associated tablestore.
32 papers · 0 benchmarks
The Extreme Summarization (XSum) dataset is a dataset for evaluation of abstractive single-document summarization systems.
32 papers · 5 benchmarks
The Actor-Action Dataset (A2D) by Xu et al.
31 papers · 1 benchmark
We introduce ACDC, the Adverse Conditions Dataset with Correspondences for training and testing semantic segmentation methods on adverse visual conditions.
31 papers · 5 benchmarks
The AGIQA-3K is a fine-grained AI-generated image (AGI) subjective quality assessment database.
31 papers · 0 benchmarks
ASTD (Arabic Sentiment Tweets Dataset)
Arabic Sentiment Tweets Dataset (ASTD) is an Arabic social sentiment analysis dataset gathered from Twitter.
31 papers · 1 benchmark
ArtEmis is a large-scale dataset aimed at providing a detailed understanding of the interplay between visual content, its emotional effect, and explanations for the latter in language.
31 papers · 0 benchmarks
CAVE (Multispectral imaging using multiplexed illumination.)
Multispectral imaging using multiplexed illumination.
31 papers · 1 benchmark
Source: CHANGE DETECTION IN REMOTE SENSING IMAGES USING CONDITIONAL ADVERSARIAL NETWORKS
31 papers · 2 benchmarks
ContractNLI is a dataset for document-level natural language inference (NLI) on contracts whose goal is to automate/support a time-consuming procedure of contract review.
31 papers · 0 benchmarks
DCASE 2016 is a dataset for sound event detection.
31 papers · 0 benchmarks
The DQN Replay Dataset was collected as follows: We first train a [DQN][naturedqn] agent, on all 60 [Atari 2600 games][ale] with [sticky actions][stochasticale] enabled for 200 million frames (standard protocol) and save all of the…
31 papers · 0 benchmarks
A novel benchmark and dataset for the evaluation of image-based garment reconstruction systems.
31 papers · 0 benchmarks
Consists of 190K posts from five different categories of Reddit communities.
31 papers · 0 benchmarks
EGAD (Evolved Grasping Analysis Dataset)
The Evolved Grasping Analysis Dataset (EGAD) comprises over 2000 generated objects aimed at training and evaluating robotic visual grasp detection algorithms.
31 papers · 0 benchmarks
ExPose (EXpressive POse and Shape rEgression)
Curates a dataset of SMPL-X fits on in-the-wild images.
31 papers · 0 benchmarks
The friedman1 data set is commonly used to test semi-supervised regression methods.
31 papers · 0 benchmarks
GuitarSet is a dataset of high-quality guitar recordings and rich annotations.
31 papers · 2 benchmarks
IGLUE (Image-Grounded Language Understanding Evaluation)
The Image-Grounded Language Understanding Evaluation (IGLUE) benchmark brings together—by both aggregating pre-existing datasets and creating new ones—visual question answering, cross-modal retrieval, grounded reasoning, and grounded…
31 papers · 0 benchmarks
The Image Paragraph Captioning dataset allows researchers to benchmark their progress in generating paragraphs that tell a story about an image.
31 papers · 1 benchmark
The IMAGE-CHAT dataset is a large collection of (image, style trait for speaker A, style trait for speaker B, dialogue between A & B) tuples that we collected using crowd-workers, Each dialogue consists of consecutive turns by speaker A…
31 papers · 2 benchmarks
KVQA (Knowledge-aware VQA)
It contains manually verified 183K question-answer pairs about more than 18K persons and 24K images.
31 papers · 0 benchmarks
LC25000 (Lung And Colon Histopathological Image Dataset)
The LC25000 dataset contains 25,000 color images with 5 classes of 5,000 images each.
31 papers · 0 benchmarks
MAFW is a large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.
31 papers · 2 benchmarks
Under Institutional Review Board (IRB) supervision, 50 abdomen CT scans of were randomly selected from a combination of an ongoing colorectal cancer chemotherapy trial, and a retrospective ventral hernia study.
31 papers · 3 benchmarks
The MIT-BIH Arrhythmia Database contains 48 half-hour excerpts of two-channel ambulatory ECG recordings, obtained from 47 subjects studied by the BIH Arrhythmia Laboratory between 1975 and 1979.
31 papers · 5 benchmarks
Multimodal C4 (MMC4) is an augmentation of the popular text-only c4 corpus with images interleaved.
31 papers · 0 benchmarks
The MQ2008 dataset is a dataset for Learning to Rank.
31 papers · 0 benchmarks
The MSLR-WEB30K dataset consists of 30,000 search queries over the documents from search results.
31 papers · 1 benchmark
MSRC-12 (MSRC-12 Kinect Gesture Dataset)
The Microsoft Research Cambridge-12 Kinect gesture data set consists of sequences of human movements, represented as body-part locations, and the associated gesture to be recognized by the system.
31 papers · 2 benchmarks
OIE2016 is the first large-scale OpenIE benchmark.
31 papers · 1 benchmark
OpenEDS (Open Eye Dataset) is a large scale data set of eye-images captured using a virtual-reality (VR) head mounted display mounted with two synchronized eyefacing cameras at a frame rate of 200 Hz under controlled illumination.
31 papers · 1 benchmark
OpinionQA is a dataset for evaluating the alignment of LM opinions with those of 60 US demographic groups over topics ranging from abortion to automation.
31 papers · 0 benchmarks
PMC-OA (PubmedCentral OpenAcess)
PMC-OA is a large-scale dataset that contains 1.65M image-text pairs.
31 papers · 0 benchmarks
PhraseCut is a dataset consisting of 77,262 images and 345,486 phrase-region pairs.
31 papers · 1 benchmark
SOC (Salient Objects in Clutter)
SOC (Salient Objects in Clutter) is a dataset for Salient Object Detection (SOD).
31 papers · 1 benchmark
When glancing at a magazine, or browsing the Internet, we are continuously being exposed to photographs.
31 papers · 4 benchmarks
The SemEval-2013 Task 2 dataset contains data for two subtasks: A, an expression-level subtask, and B, a message-level subtask.
31 papers · 0 benchmarks
TIMIT (TIMIT Acoustic-Phonetic Continuous Speech Corpus)
The TIMIT Acoustic-Phonetic Continuous Speech Corpus is a standard dataset used for evaluation of automatic speech recognition systems.
31 papers · 6 benchmarks
A large scale dataset with daily-living activities performed in a natural manner.
31 papers · 1 benchmark
WCEP (Wikipedia Current Events Portal)
The WCEP dataset for multi-document summarization (MDS) consists of short, human-written summaries about news events, obtained from the Wikipedia Current Events Portal (WCEP), each paired with a cluster of news articles associated with an…
31 papers · 1 benchmark
Weather (Max-Planck-Institut Weather Dataset for Long-term Time Series Forecasting)
Weather is recorded every 10 minutes for the 2020 whole year, which contains 21 meteorological indicators, such as air temperature, humidity, etc.
31 papers · 9 benchmarks
AGQA (Action Genome Question Answering)
Action Genome Question Answering (AGQA) is a benchmark for compositional spatio-temporal reasoning.
30 papers · 0 benchmarks
BRACS (BReAst Carcinoma Subtyping)
BReAst Carcinoma Subtyping (BRACS) dataset, a large cohort of annotated Hematoxylin & Eosin (H&E)-stained images to facilitate the characterization of breast lesions.
30 papers · 0 benchmarks
A testbed for commonsense reasoning about entity knowledge, bridging fact-checking about entities with commonsense inferences.
30 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.