Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 72 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 3409–3456 of 12,172
FERG (Facial Expression Research Group Database)
FERG is a database of cartoon characters with annotated facial expressions containing 55,769 annotated face images of six characters.
8 papers · 1 benchmark
FVI (Free-form Video Inpainting)
The Free-Form Video Inpainting dataset is a dataset used for training and evaluation video inpainting models.
8 papers · 0 benchmarks
FeTS2022 (Federated Tumor Segmentation Challenge 2022)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
8 papers · 0 benchmarks
A dataset for fine-grained entity typing of knowledge graph entities built from Freebase.
8 papers · 0 benchmarks
FloorPlanCAD is a large-scale real-world CAD drawing dataset containing over 15,000 floor plans, ranging from residential to commercial buildings.
8 papers · 0 benchmarks
GLUE-X is a benchmark dataset used to evaluate the out-of-distribution (OOD) robustness of Natural Language Understanding (NLU) models.
8 papers · 0 benchmarks
GPA (Geometric Pose Affordance)
multi-view imagery of people interacting with a variety of rich 3D environments
8 papers · 2 benchmarks
A GQA-based dataset with 1,040,830 multi-modal explanations of visual reasoning processes.
8 papers · 1 benchmark
GTA (A Benchmark for General Tool Agents)
A benchmark to evaluate the tool-use capabilities of LLM-based agents in real-world scenarios.
8 papers · 0 benchmarks
GVFC (Gun Violence Frame Corpus)
This is a new dataset of news headlines and their frames related to the issue of gun violence in the United States.
8 papers · 0 benchmarks
GermanQuAD is a Question Answering (QA) dataset of 13,722 extractive question/answer pairs in German.
8 papers · 1 benchmark
Ghera is a repository of Android app vulnerabilities.
8 papers · 0 benchmarks
GigaST is a large-scale pseudo speech translation (ST) corpus.
8 papers · 0 benchmarks
H3WB (Human 3.6M 3D WholeBody)
Human3.6M 3D WholeBody (H3WB) is a large scale dataset with 133 whole-body keypoint annotations on 100K images, made possible by a new multi-view pipeline.
8 papers · 3 benchmarks
HANDAL (HANDAL: A Dataset of Real-World Manipulable Object Categories with Pose Annotations, Affordances, and Reconstructions)
We present the HANDAL dataset for category-level object pose estimation and affordance prediction.
8 papers · 0 benchmarks
HANNA (HANNA, a large annotated dataset of Human-ANnotated NArratives for ASG evaluation.)
HANNA, a large annotated dataset of Human-ANnotated NArratives for Automatic Story Generation (ASG) evaluation, has been designed for the benchmarking of automatic metrics for ASG.
8 papers · 0 benchmarks
The HInt dataset is frequently used as a generalizability benchmark for 3D Hand Reconstruction.
8 papers · 1 benchmark
The official HOList benchmark for automated theorem proving consists of all theorem statements in the core, complex, and flyspeck corpora.
8 papers · 1 benchmark
HS-SOD (HyperSpectral Salient Object Detection Dataset)
HS-SOD is a hyperspectral salient object detection dataset with a collection of 60 hyperspectral images with their respective ground-truth binary images and representative rendered colour images (sRGB).
8 papers · 0 benchmarks
HumAID (Human-Annotated Disaster Incidents Data)
Social networks are widely used for information consumption and dissemination, especially during time-critical events such as natural disasters.
8 papers · 0 benchmarks
The IMUPoser Dataset is a dataset for estimating body pose using IMUs already in devices that many users own -- namely smartphones, smartwatches, and earbuds.
8 papers · 0 benchmarks
Contains 446,684 images annotated by humans that cover 43 incidents across a variety of scenes.
8 papers · 0 benchmarks
JSRT (Japanese Society of Radiological Technology Database)
The standard digital image database with and without chest lung nodules (JSRT database) was created(1) by the Japanese Society of Radiological Technology (JSRT) in cooperation with the Japanese Radiological Society (JRS) in 1998.
8 papers · 0 benchmarks
JSUT Corpus is a free large-scale speech corpus that can be shared between academic institutions and commercial companies has an important role.
8 papers · 0 benchmarks
The dataset contains transactions made by credit cards in September 2013 by European cardholders.
8 papers · 2 benchmarks
Kinetics-100 is a dataset split created from the Kinetics dataset to evaluate the performance of few-shot action recognition models.
8 papers · 1 benchmark
L3CubeMahaSent is a large publicly available Marathi Sentiment Analysis dataset.
8 papers · 0 benchmarks
LLCM (Low-Light Cross-Modality Person Dataset)
LLCM (Low-Light Cross-Modality) dataset is constructed to facilitate the study of low-light cross-modality person Re-ID task.
8 papers · 0 benchmarks
LPW (Labeled Pedestrian in the Wild)
Labeled Pedestrian in the Wild (LPW) is a pedestrian detection dataset that contains 2,731 pedestrians in three different scenes where each annotated identity is captured by from 2 to 4 cameras.
8 papers · 0 benchmarks
LReID is a benchmark for lifelong person reidentification.
8 papers · 0 benchmarks
LSOIE (Large-Scale dataset for Supervised Open Information Extraction)
LSOIE is a large-scale OpenIE data converted from QA-SRL 2.0 in two domains, i.e., Wikipedia and Science.
8 papers · 2 benchmarks
LUN is used for unreliable news source classification, this dataset includes 17,250 articles from satire, propaganda, and hoaxe.
8 papers · 1 benchmark
Logic2Text is a large-scale dataset with 10,753 descriptions involving common logic types paired with the underlying logical forms.
8 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
8 papers · 1 benchmark
LongForm dataset is created by leveraging English corpus examples with augmented instructions.
8 papers · 0 benchmarks
A Benchmark to evaluate complex instruction following.
8 papers · 0 benchmarks
MISAW (MIcro-Surgical Anastomose Workflow recognition on training sessions)
The MISAW data set is composed of 27 sequences of micro-surgical anastomosis on artificial blood vessels performed by 3 surgeons and 3 engineering students.
8 papers · 1 benchmark
MMDU (Multi-Turn Multi-Image Dialog Understanding)
MMDU, a comprehensive benchmark, and MMDU-45k, a large-scale instruction tuning dataset, are designed to evaluate and improve LVLMs' abilities in multi-turn and multi-image conversations.
8 papers · 0 benchmarks
MO-Gymnasium is an open source Python library for developing and comparing multi-objective reinforcement learning algorithms by providing a standard API to communicate between learning algorithms and environments, as well as a standard set…
8 papers · 0 benchmarks
MRDA (ICSI Meeting Recorder Dialog Act Corpus)
The MRDA corpus consists of about 75 hours of speech from 75 naturally-occurring meetings among 53 speakers.
8 papers · 1 benchmark
Serving as an out-of-domain xMR test dataset, MSVAMP allows for a more exhaustive and comprehensive evaluation of the model’s multilingual mathematical capabilities.
8 papers · 0 benchmarks
Based on the MVSEC dataset, we select some image-event pairs to evaluate the segmentation performance, namely MVSEC-SEG, which only serves as a test set.
8 papers · 1 benchmark
ManyTypes4Py is a large Python dataset for machine learning (ML)-based type inference.
8 papers · 0 benchmarks
MathMLben is a benchmark to the evaluate tools for mathematical format conversion (LaTeX ↔ MathML ↔ CAS).
8 papers · 0 benchmarks
MeGlass is an eyeglass dataset originally designed for eyeglass face recognition evaluation.
8 papers · 0 benchmarks
Meta-Album (Multi-domain Meta-Dataset for Few-Shot Image Classification)
Meta Album is a meta-dataset created for few-shot learning, meta-learning, continual learning and so on.
8 papers · 0 benchmarks
A large new multilingual dataset for multilingual entity linking.
8 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.