Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 87 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 4129–4176 of 12,172
XImageNet-12 (XIMAGENET-12: An Explainable AI Benchmark Dataset for Model Robustness Evaluation)
Enlarge the dataset to understand how image background effect the Computer Vision ML model.
6 papers · 1 benchmark
XL-BEL is a benchmark for cross-lingual biomedical entity linking (XL-BEL).
6 papers · 0 benchmarks
XQA is a data which consists of a total amount of 90k question-answer pairs in nine languages for cross-lingual open-domain question answering.
6 papers · 0 benchmarks
A new dataset with significant occlusions related to object manipulation.
6 papers · 0 benchmarks
Dataset for large-scale yoga pose recognition with 82 classes.
6 papers · 0 benchmarks
YouTube Driving Dataset contains a massive amount of real-world driving frames with various conditions, from different weather, different regions, to diverse scene types
6 papers · 1 benchmark
Video object segmentation has been studied extensively in the past decade due to its importance in understanding video spatial-temporal structures as well as its value in industrial applications.
6 papers · 1 benchmark
ZEGGS dataset contains 67 sequences of monologues performed by a female actor speaking in English and covers 19 different motion styles.
6 papers · 0 benchmarks
The ZS-F-VQA dataset is a new split of the F-VQA dataset for zero-shot problem.
6 papers · 1 benchmark
Benchmark dataset for abstracts and titles of 100,000 ArXiv scientific papers.
6 papers · 1 benchmark
This benchmark includes 11 image classification datasets that were used to evaluate the transferability of metrics.
6 papers · 1 benchmark
The exiD dataset introduces a groundbreaking collection of naturalistic road user trajectories at highway entries and exits in Germany, meticulously captured with drones to navigate past the limitations of conventional traffic data…
6 papers · 0 benchmarks
iQIYI-VID dataset, which comprises video clips from iQIYI variety shows, films, and television dramas.
6 papers · 0 benchmarks
legalNER is a corpus of 46545 annotated legal named entities mapped to 14 legal entity types.
6 papers · 0 benchmarks
A multimodal database for eye blink detection and attention level estimation.
6 papers · 0 benchmarks
The sStoryCloze refers to the Spoken StoryCloze benchmark, which is a spoken version of the StoryCloze dataset.
6 papers · 0 benchmarks
safe-control-gym is an open-source benchmark suite that extends OpenAI's Gym API with (i) the ability to specify (and query) symbolic models and constraints and (ii) introduce simulated disturbances in the control inputs, measurements, and…
6 papers · 0 benchmarks
Therapeutics Data Commons is an open-science initiative with AI/ML-ready datasets and AI/ML tasks for therapeutics, spanning the discovery and development of safe and effective medicines.
6 papers · 1 benchmark
1QIsaa data collection (1QIsaa data collection (binarized images, feature files, and plotting scripts) for writer identification test)
This data set is collected for the ERC project: The Hands that Wrote the Bible: Digital Palaeography and Scribal Culture of the Dead Sea Scrolls PI: Mladen Popović Grant agreement ID: 640497 Project website:…
5 papers · 0 benchmarks
20-MAD (20-MAD: Mozilla Apache Dataset)
20-MAD, a dataset linking the commit and issue data of Mozilla and Apache projects.
5 papers · 0 benchmarks
25kTrees (Individual Tree Crown Annotations)
Manual crown delineation of individual trees in two countries: Denmark and Finland.
5 papers · 0 benchmarks
2DeteCT (2DeteCT - A large 2D expandable, trainable, experimental Computed Tomography dataset for machine learning)
Maximilian B.
5 papers · 0 benchmarks
A Large Dataset of Object Scans is a dataset of more than ten thousand 3D scans of real objects.
5 papers · 0 benchmarks
AART (AI-Assisted Red-Teaming)
AART serves as an automated alternative to the current manual red-teaming efforts.
5 papers · 0 benchmarks
The ADL Piano MIDI is a dataset of 11,086 piano pieces from different genres.
5 papers · 0 benchmarks
ADVErsarial Table perturbAtion (ADVETA) is a robustness evaluation benchmark featuring natural and realistic ATPs.
5 papers · 0 benchmarks
The 2017 PhysioNet/CinC Challenge aims to encourage the development of algorithms to classify, from a single short ECG lead recording (between 30 s and 60 s in length), whether the recording shows normal sinus rhythm, atrial fibrillation…
5 papers · 0 benchmarks
AG-ReID (Aerial-Ground Person Re-identification)
Person re-ID matches persons across multiple non-overlapping cameras.
5 papers · 1 benchmark
AHP (Amodal Human Perception)
The AHP dataset consists of 56,599 images in total which are collected from several large-scale instance segmentation and detection datasets, including COCO, VOC (w/ SBD), LIP, Objects365 and OpenImages.
5 papers · 0 benchmarks
AI2D-RST is a multimodal corpus of 1000 English-language diagrams that represent topics in primary school natural sciences, such as food webs, life cycles, moon phases and human physiology.
5 papers · 0 benchmarks
AMALGUM (A Machine Annotated Lookalike of GUM)
AMALGUM is a machine annotated multilayer corpus following the same design and annotation layers as GUM, but substantially larger (around 4M tokens).
5 papers · 0 benchmarks
ARID (Autonomous Robot Indoor Dataset)
ARID is a large-scale, multi-view object dataset collected with an RGB-D camera mounted on a mobile robot.
5 papers · 0 benchmarks
Dataset to address the problem of detecting people Looking At Each Other (LAEO) in video sequences.
5 papers · 0 benchmarks
Predicting the age of abalone from physical measurements.
5 papers · 1 benchmark
Acappella comprises around 46 hours of a cappella solo singing videos sourced from YouTbe, sampled across different singers and languages.
5 papers · 0 benchmarks
AccentDB is a database that contains samples of 4 Indian-English accents, and a compilation of samples from 4 native-English, and a metropolitan Indian-English accent.
5 papers · 0 benchmarks
ActorShift is a dataset where the domain shift comes from the change in actor species: we use humans in the source domain and animals in the target domain.
5 papers · 0 benchmarks
AdaptiX (AdaptiX – A Transitional XR Framework for Development and Evaluation of Shared Control Applications in Assistive Robotics)
GitHub repository for "AdaptiX – A Transitional XR Framework for Development and Evaluation of Shared Control Applications in Assistive Robotics", which is used in several shared control applications
5 papers · 0 benchmarks
Adressa (SmartMedia Adressa News Dataset)
The Adressa Dataset is a news dataset that includes news articles (in Norwegian) in connection with anonymized users.
5 papers · 0 benchmarks
[1]: https://www.projectaria.com/datasets/ase/ "" [2]: https://facebookresearch.github.io/projectariatools/docs/opendatasets/ariasyntheticenvironmentsdataset "" [3]: https://www.projectaria.com/research/ "" Aria Synthetic Environments is a…
5 papers · 2 benchmarks
Artie Bias Corpus is an open dataset for detecting demographic bias in speech applications.
5 papers · 0 benchmarks
AstroVision is a large-scale dataset comprised of 115,970 densely annotated, real images of 16 different small bodies from both legacy and ongoing deep space missions to facilitate the study of deep learning for autonomous navigation in…
5 papers · 0 benchmarks
Large vision-language models (LVLMs) are prone to hallucinations, where certain contextual cues in an image can trigger the language module to produce overconfident and incorrect reasoning about abnormal or hypothetical objects.
5 papers · 1 benchmark
Avalon is a benchmark for generalization in Reinforcement Learning (RL).
5 papers · 0 benchmarks
BB-norm-habitat (Bacteria Biotope - entity normalization - bacterial habitat)
In the BB-norm modality of this task, participant systems had to normalize textual entity mentions according to the OntoBiotope ontology for habitats.
5 papers · 0 benchmarks
In the BB-norm modality of this task, participant systems had to normalize textual entity mentions according to the OntoBiotope ontology for phenotypes.
5 papers · 0 benchmarks
The BEHAVIOR-1K dataset is a comprehensive simulation benchmark for human-centered robotics¹.
5 papers · 0 benchmarks
BEOID (Bristol Egocentric Object Interactions Dataset)
The BEOID dataset includes object interactions ranging from preparing a coffee to operating a weight lifting machine and opening a door.
5 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.