Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 67 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 3169–3216 of 12,172
BLVD is a large scale 5D semantics dataset collected by the Visual Cognitive Computing and Intelligent Vehicles Lab.
9 papers · 0 benchmarks
Bactrian-X is a comprehensive multilingual parallel dataset of 3.4 million instruction-response pairs across 52 languages.
9 papers · 0 benchmarks
BiToD is a bilingual multi-domain dataset for end-to-end task-oriented dialogue modeling.
9 papers · 0 benchmarks
The Blackbird unmanned aerial vehicle (UAV) dataset is a large-scale, aggressive indoor flight dataset collected using a custom-built quadrotor platform for use in evaluation of agile perception.
9 papers · 0 benchmarks
Bonn RGB-D Dynamic is a dataset for RGB-D SLAM, containing highly dynamic sequences.
9 papers · 1 benchmark
The Japanese-English business conversation corpus, namely Business Scene Dialogue corpus, was constructed in 3 steps: 1.
9 papers · 2 benchmarks
LRW-1000 has been renamed as CAS-VSR-W1k.
9 papers · 1 benchmark
The CC3D dataset [1] of 3D CAD models was collected from a free online service for sharing CAD designs [2].
9 papers · 1 benchmark
CCGbank is a translation of the Penn Treebank into a corpus of Combinatory Categorial Grammar derivations.
9 papers · 1 benchmark
CHAOS (CHAOS - Combined (CT-MR) Healthy Abdominal Organ Segmentation)
CHAOS challenge aims the segmentation of abdominal organs (liver, kidneys and spleen) from CT and MRI data.
9 papers · 0 benchmarks
his is an academic intrusion detection dataset.
9 papers · 0 benchmarks
CLIRMatrix is a large collection of bilingual and multilingual datasets for Cross-Lingual Information Retrieval.
9 papers · 0 benchmarks
CODEBRIM (COncrete DEfect BRidge IMage Dataset)
Dataset for multi-target classification of five commonly appearing concrete defects.
9 papers · 0 benchmarks
In this work, we propose a general dataset for Color-Event camera based Single Object Tracking, termed COESOT.
9 papers · 1 benchmark
COVERAGE (Copy-Move Forgery Database with Similar but Genuine Objects)
COVERAGE contains copymove forged (CMFD) images and their originals with similar but genuine objects (SGOs).
9 papers · 2 benchmarks
CQASUMM is a dataset for CQA (Community Question Answering) summarization, constructed from the 4.4 million Yahoo!
9 papers · 0 benchmarks
CRD3 (Critical Role Dungeons and Dragons Dataset)
The dataset is collected from 159 Critical Role episodes transcribed to text dialogues, consisting of 398,682 turns.
9 papers · 0 benchmarks
CRUW is a dataset for the radar object detection (ROD) task, which aims to classify and localize the objects in 3D purely from radar's radio frequency (RF) images.
9 papers · 0 benchmarks
CTSpine1K is a large-scale and comprehensive dataset for research in spinal image analysis.
9 papers · 0 benchmarks
CUHK03-C is an evaluation set that consists of algorithmically generated corruptions applied to the CUHK03 test-set.
9 papers · 1 benchmark
CaDIS (Cataract Dataset for Image Segmentation)
CaDIS: a Cataract Dataset for Image Segmentation is a dataset for semantic segmentation created by Digital Surgery Ltd.
9 papers · 1 benchmark
CaSiNo is a dataset of 1030 negotiation dialogues in English.
9 papers · 0 benchmarks
We provide manual annotations of 14 semantic keypoints for 100,000 car instances (sedan, suv, bus, and truck) from 53,000 images captured from 18 moving cameras at Multiple intersections in Pittsburgh, PA.
9 papers · 2 benchmarks
Chest X-ray images for pneumonia detection.
9 papers · 2 benchmarks
ClueWeb22 is the newest iteration of the ClueWeb line of datasets, provides 10 billion web pages affiliated with rich information.
9 papers · 0 benchmarks
CodeQA is a free-form question answering dataset for the purpose of source code comprehension: given a code snippet and a question, a textual answer is required to be generated.
9 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
9 papers · 0 benchmarks
The Collective Activity Dataset contains 5 different collective activities: crossing, walking, waiting, talking, and queueing and 44 short video sequences some of which were recorded by consumer hand-held digital camera with varying view…
9 papers · 1 benchmark
ComQA is a large dataset of real user questions that exhibit different challenging aspects such as compositionality, temporal reasoning, and comparisons.
9 papers · 0 benchmarks
Language Model (LM) agents for cybersecurity that are capable of autonomously identifying vulnerabilities and executing exploits have potential to cause real-world impact.
9 papers · 1 benchmark
DAD (Driver Anomaly Detection)
Contains normal driving videos together with a set of anomalous actions in its training set.
9 papers · 0 benchmarks
Deep Learning Hard (DL-HARD) is an annotated dataset designed to more effectively evaluate neural ranking models on complex topics.
9 papers · 0 benchmarks
SMAC+ defensive armored scenario with sequential episodic buffer
9 papers · 1 benchmark
smac+ defense infantry scenario with parallel episodic buffer
9 papers · 1 benchmark
SMAC+ defensive infantry scenario with sequential episodic buffer
9 papers · 1 benchmark
SMAC+ defensive outnumbered scenario with sequential episodic buffer
9 papers · 1 benchmark
Defects4J is a collection of reproducible bugs and a supporting infrastructure with the goal of advancing software engineering research.
9 papers · 1 benchmark
Contains 4,677 videos with temporal, spatial, and categorical annotations.
9 papers · 0 benchmarks
E-KAR (Benchmark for Explainable Knowledge-intensive Analogical Reasoning)
The ability to recognize analogies is fundamental to human cognition.
9 papers · 0 benchmarks
EXPY-TKY contains the traffic speed information and the corresponding traffic incident information in 10-minute interval for 1843 expressway road links in Tokyo over three months (2021/10∼2021/12).
9 papers · 1 benchmark
The Earning Calls dataset consists of processed earning conference calls data (text and audio).
9 papers · 0 benchmarks
The Ecoli dataset is a dataset for protein localization.
9 papers · 0 benchmarks
EgoCap is a dataest of 100,000 egocentric images of eight people in different clothing, with 75,000 images from six people used for training.
9 papers · 0 benchmarks
EgoHOS (Fine-Grained Egocentric Hand-Object Segmentation Dataset)
EgoHOS is a labeled dataset consisting of 11243 egocentric images with per-pixel segmentation labels of hands and objects being interacted with during a diverse array of daily activities.
9 papers · 0 benchmarks
EgoProceL is a large-scale dataset for procedure learning.
9 papers · 0 benchmarks
The EntitySeg dataset contains 33,227 images with high-quality mask annotations.
9 papers · 0 benchmarks
FAIR-Play is a video-audio dataset consisting of 1,871 video clips and their corresponding binaural audio clips recording in a music room.
9 papers · 0 benchmarks
Natural Language Inference (NLI), also called Textual Entailment, is an important task in NLP with the goal of determining the inference relationship between a premise p and a hypothesis h.
9 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.