Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 22 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1009–1056 of 12,172
Dark Zurich is an image dataset containing a total of 8779 images captured at nighttime, twilight, and daytime, along with the respective GPS coordinates of the camera for each image.
57 papers · 3 benchmarks
The EYEDIAP dataset is a dataset for gaze estimation from remote RGB, and RGB-D (standard vision and depth), cameras.
57 papers · 3 benchmarks
Europarl-ST is a multilingual Spoken Language Translation corpus containing paired audio-text samples for SLT from and into 9 European languages, for a total of 72 different translation directions.
57 papers · 0 benchmarks
Fluent Speech Commands is an open source audio dataset for spoken language understanding (SLU) experiments.
57 papers · 1 benchmark
Lost and Found is a novel lost-cargo image sequence dataset comprising more than two thousand frames with pixelwise annotations of obstacle and free-space and provide a thorough comparison to several stereo-based baseline methods.
57 papers · 1 benchmark
MSLS (Mapillary Street-level Sequences Dataset)
The largest and most diverse dataset for lifelong place recognition from image sequences in urban and suburban settings.
57 papers · 1 benchmark
Open-X-Embodiment robot manipulation dataset, see https://robotics-transformer-x.github.io/
57 papers · 0 benchmarks
PGM (Procedurally Generated Matrices (PGM))
PGM dataset serves as a tool for studying both abstract reasoning and generalisation in models.
57 papers · 0 benchmarks
RELLIS-3D is a multi-modal dataset for off-road robotics.
57 papers · 3 benchmarks
SciDocs evaluation framework consists of a suite of evaluation tasks designed for document-level tasks.
57 papers · 2 benchmarks
The Slashdot dataset is a relational dataset obtained from Slashdot.
57 papers · 2 benchmarks
The Stanford Dogs dataset contains 20,580 images of 120 classes of dogs from around the world, which are divided into 12,000 images for training and 8,580 images for testing.
57 papers · 6 benchmarks
The TotalCapture dataset consists of 5 subjects performing several activities such as walking, acting, a range of motion sequence (ROM) and freestyle motions, which are recorded using 8 calibrated, static HD RGB cameras and 13 IMUs…
57 papers · 2 benchmarks
WHAMR! (WHAM! with synthetic reverberated sources)
WHAMR!
57 papers · 3 benchmarks
The Wireframe dataset consists of 5,462 images (5,000 for training, 462 for test) of indoor and outdoor man-made scenes.
57 papers · 2 benchmarks
Data Set Information: Extraction was done by Barry Becker from the 1994 Census database.
56 papers · 2 benchmarks
BEAT (Body-Expression-Audio-Text)
BEAT has i) 76 hours, high-quality, multi-modal data captured from 30 speakers talking with eight different emotions and in four different languages, ii) 32 millions frame-level emotion and semantic relevance annotations.
56 papers · 1 benchmark
CoNLL++ is a corrected version of the CoNLL03 NER dataset where 5.38% of the test sentences have been fixed.
56 papers · 2 benchmarks
Consists of over one million high-resolution images of varying gaze under extreme head poses.
56 papers · 1 benchmark
From scientific research to commercial applications, eye tracking is an important tool across many domains.
56 papers · 1 benchmark
The Jobs dataset by LaLonde [36] is a widely used benchmark in the causal inference community, where the treatment is job training and the outcomes are income and employment status after training.
56 papers · 1 benchmark
MINC (Materials in Context Database)
MINC is a large-scale, open dataset of materials in the wild.
56 papers · 0 benchmarks
MasakhaNER is a collection of Named Entity Recognition (NER) datasets for 10 different African languages.
56 papers · 1 benchmark
A large real-world event-based dataset for object classification.
56 papers · 2 benchmarks
PA-100K is a recent-proposed large pedestrian attribute dataset, with 100,000 images in total collected from outdoor surveillance cameras.
56 papers · 1 benchmark
RTE (Recognizing Textual Entailment)
The Recognizing Textual Entailment (RTE) datasets come from a series of textual entailment challenges.
56 papers · 2 benchmarks
The Re-TACRED dataset is a significantly improved version of the TACRED dataset for relation extraction.
56 papers · 1 benchmark
This dataset contains images of unusual dangers which can be encountered by a vehicle on the road – animals, rocks, traffic cones and other obstacles.
56 papers · 1 benchmark
UCI Machine Learning Repository is a collection of over 550 datasets.
56 papers · 9 benchmarks
VOT2017 (Visual Object Tracking Challenge)
VOT2017 is a Visual Object Tracking dataset for different tasks that contains 60 short sequences annotated with 6 different attributes.
56 papers · 2 benchmarks
ISEAR (International Survey on Emotion Antecedents and Reactions)
Over a period of many years during the 1990s, a large group of psychologists all over the world collected data in the ISEAR project, directed by Klaus R.
55 papers · 0 benchmarks
MuTual is a retrieval-based dataset for multi-turn dialogue reasoning, which is modified from Chinese high school English listening comprehension test data.
55 papers · 0 benchmarks
OpenDialKG contains utterance from 15K human-to-human role-playing dialogs is manually annotated with ground-truth reference to corresponding entities and paths from a large-scale KG with 1M+ facts.
55 papers · 0 benchmarks
PMC-VQA is a large-scale medical visual question-answering dataset that contains 227k VQA pairs of 149k images that cover various modalities or diseases.
55 papers · 2 benchmarks
PrOntoQA (Proof and Ontology-Generated Question-Answering)
PrOntoQA is a question-answering dataset which generates examples with chains-of-thought that describe the reasoning required to answer the questions correctly.
55 papers · 0 benchmarks
QUASAR-T (QUestion Answering by Search And Reading – Trivia)
QUASAR-T is a large-scale dataset aimed at evaluating systems designed to comprehend a natural language query and extract its answer from a large corpus of text.
55 papers · 1 benchmark
Quora Question Pairs (QQP) dataset consists of over 400,000 question pairs, and each question pair is annotated with a binary value indicating whether the two questions are paraphrase of each other.
55 papers · 8 benchmarks
The REVERB (REverberant Voice Enhancement and Recognition Benchmark) challenge is a benchmark for evaluation of automatic speech recognition techniques.
55 papers · 1 benchmark
The Shifts Dataset is a dataset for evaluation of uncertainty estimates and robustness to distributional shift.
55 papers · 1 benchmark
We present a benchmark for image-based 3D reconstruction.
55 papers · 2 benchmarks
Tiny ImageNet-C is an open-source data set comprising algorithmically generated corruptions applied to the Tiny ImageNet (ImageNet-200) test set comprising 200 classes following the concept of ImageNet-C.
55 papers · 0 benchmarks
VLN-CE (Vision-and-Language Navigation in Continuous Environments)
Vision and Language Navigation in Continuous Environments (VLN-CE) is an instruction-guided navigation task with crowdsourced instructions, realistic environments, and unconstrained agent navigation.
55 papers · 1 benchmark
Fisheye cameras are commonly employed for obtaining a large field of view in surveillance, augmented reality and in particular automotive applications.
55 papers · 1 benchmark
XTREME (Cross-Lingual Transfer Evaluation of Multilingual Encoders)
The Cross-lingual TRansfer Evaluation of Multilingual Encoders (XTREME) benchmark was introduced to encourage more research on multilingual transfer learning,.
55 papers · 2 benchmarks
Emotion-cause pair extraction (ECPE) aims to extract the potential pairs of emotions and corresponding causes in a document.
55 papers · 0 benchmarks
ASSET is a new dataset for assessing sentence simplification in English.
54 papers · 1 benchmark
The Bamboogle dataset is a collection of questions that was constructed to investigate the ability of language models to perform compositional reasoning tasks.
54 papers · 1 benchmark
C3 is a free-form multiple-Choice Chinese machine reading Comprehension dataset.
54 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.