Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 25 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1153–1200 of 12,172
Avenue Dataset contains 16 training and 21 testing video clips.
48 papers · 2 benchmarks
In Clipart1k, the target domain classes to be detected are the same as those in the source domain.
48 papers · 3 benchmarks
The DNS Challenge at INTERSPEECH 2020 intended to promote collaborative research in single-channel Speech Enhancement aimed to maximize the perceptual quality and intelligibility of the enhanced speech.
48 papers · 3 benchmarks
A new challenge set for multimodal classification, focusing on detecting hate speech in multimodal memes.
48 papers · 0 benchmarks
JHU-CROWD++ is A large-scale unconstrained crowd counting dataset with 4,372 images and 1.51 million annotations.
48 papers · 1 benchmark
The KIT Motion-Language is a dataset linking human motion and natural language.
48 papers · 2 benchmarks
MedMentions is a new manually annotated resource for the recognition of biomedical concepts.
48 papers · 1 benchmark
The Objectron dataset is a collection of short, object-centric video clips, which are accompanied by AR session metadata that includes camera poses, sparse point-clouds and characterization of the planar surfaces in the surrounding…
48 papers · 0 benchmarks
SRD (Shadow Removal Dataset)
SRD is a dataset for shadow removal that contains 3088 shadow and shadow-free image pairs.
48 papers · 1 benchmark
TQA (Textbook Question Answering)
The TextbookQuestionAnswering (TQA) dataset is drawn from middle school science curricula.
48 papers · 1 benchmark
VOICES (Voices Obscured In Complex Environmental Settings)
The VOICES corpus is a dataset to promote speech and signal processing research of speech recorded by far-field microphones in noisy room conditions.
48 papers · 0 benchmarks
BDD-X (Berkeley Deep Drive-X (eXplanation))
Berkeley Deep Drive-X (eXplanation) is a dataset is composed of over 77 hours of driving within 6,970 videos.
47 papers · 0 benchmarks
BillSum is the first dataset for summarization of US Congressional and California state bills.
47 papers · 2 benchmarks
CityFlow is a city-scale traffic camera dataset consisting of more than 3 hours of synchronized HD videos from 40 cameras across 10 intersections, with the longest distance between two simultaneous cameras being 2.5 km.
47 papers · 1 benchmark
ECHR is an English legal judgment prediction dataset of cases from the European Court of Human Rights (ECHR).
47 papers · 1 benchmark
The Elephant MIL dataset is a benchmark used in multiple instance learning (MIL), which falls under the broader categories of image classification and content-based image retrieval.
47 papers · 1 benchmark
Gait3D is a large-scale 3D representation-based gait recognition dataset.
47 papers · 2 benchmarks
The IMPACT dataset contains 50 human created prompts for each category, 200 in total, to test LLMs general writing ability.
47 papers · 0 benchmarks
MedleyDB, is a dataset of annotated, royalty-free multitrack recordings.
47 papers · 0 benchmarks
MultiCoNER is a large multilingual dataset (11 languages) for Named Entity Recognition.
47 papers · 0 benchmarks
QUASAR (QUestion Answering by Search And Reading)
The Question Answering by Search And Reading (QUASAR) is a large-scale dataset consisting of QUASAR-S and QUASAR-T.
47 papers · 1 benchmark
Reddit TIFU dataset is a newly collected Reddit dataset, where TIFU denotes the name of /r/tifu subbreddit.
47 papers · 1 benchmark
ScreenSpot Evaluation Benchmark ScreenSpot is an evaluation benchmark for GUI grounding, comprising over 1,200 instructions from various environments, including iOS, Android, macOS, Windows, and Web.
47 papers · 1 benchmark
UAV-Human is a large dataset for human behavior understanding with UAVs.
47 papers · 5 benchmarks
UBnormal (University of Bucharest Abnormal Videos)
UBnormal is a new supervised open-set benchmark composed of multiple virtual scenes for video anomaly detection.
47 papers · 3 benchmarks
WildDash is a benchmark evaluation method is presented that uses the meta-information to calculate the robustness of a given algorithm with respect to the individual hazards.
47 papers · 2 benchmarks
- We present a large and diverse abdominal CT organ segmentation dataset, termed AbdomenCT-1K, with more than 1000 (1K) CT scans from 12 medical centers, including multi-phase, multi-vendor, and multi-disease cases.
46 papers · 0 benchmarks
AliMeeting (Multi-Channel Multi-Party Meeting Transcription Challenge)
AliMeeting corpus consists of 120 hours of recorded Mandarin meeting data, including far-field data collected by 8-channel microphone array as well as near-field data collected by headset microphone.
46 papers · 1 benchmark
A new large dataset with over 100,000 examples consisting of Java classes from online code repositories, and develop a new encoder-decoder architecture that models the interaction between the method documentation and the class environment.
46 papers · 1 benchmark
CoS-E (Commonsense Explanations Dataset)
CoS-E consists of human explanations for commonsense reasoning in the form of natural language sequences and highlighted annotations Source: Explain Yourself!
46 papers · 0 benchmarks
Criteo (Display Advertising Challenge)
Criteo contains 7 days of click-through data, which is widely used for CTR prediction benchmarking.
46 papers · 1 benchmark
FEVEROUS (Fact Extraction and VERification Over Unstructured and Structured information)
FEVEROUS (Fact Extraction and VERification Over Unstructured and Structured information) is a fact verification dataset which consists of 87,026 verified claims.
46 papers · 0 benchmarks
FLIC (Frames Labelled in Cinema)
The FLIC dataset contains 5003 images from popular Hollywood movies.
46 papers · 2 benchmarks
GeoQA (Geometric Question Answering)
GeoQA is a dataset for automatic geometric problem solving containing 5,010 geometric problems with corresponding annotated programs, which illustrate the solving process of the given problems Compared with another publicly available…
46 papers · 1 benchmark
A new large-scale dataset for referring expressions, based on MS-COCO.
46 papers · 2 benchmarks
InternVid is a large-scale video-centric multimodal dataset that enables learning powerful and transferable video-text representations for multimodAL understanding and generation.
46 papers · 0 benchmarks
Legal General Language Understanding Evaluation (LexGLUE) benchmark is a collection of datasets for evaluating model performance across a diverse set of legal NLU tasks in a standardized way.
46 papers · 1 benchmark
MAMS (Multi Aspect Multi-Sentiment)
MAMS is a challenge dataset for aspect-based sentiment analysis (ABSA), in which each sentences contain at least two aspects with different sentiment polarities.
46 papers · 1 benchmark
MeViS (Motion expressions Video Segmentation)
MeViS is a large-scale dataset for motion expressions guided video segmentation, which focuses on segmenting objects in video content based on a sentence describing the motion of the objects.
46 papers · 1 benchmark
OpenLane is the first real-world and the largest scaled 3D lane dataset to date.
46 papers · 2 benchmarks
Wikipedia abstracts automatically annotated with WikiData entities and relations that are entailed by the text.
46 papers · 2 benchmarks
The Stanford Background dataset contains 715 RGB images and the corresponding label images.
46 papers · 0 benchmarks
Synscapes is a synthetic dataset for street scene parsing created using photorealistic rendering techniques, and show state-of-the-art results for training and validation as well as new types of analysis.
46 papers · 1 benchmark
TAP-Vid is a benchmark which contains both real-world videos with accurate human annotations of point tracks, and synthetic videos with perfect ground-truth point tracks.
46 papers · 1 benchmark
UDC (Ubuntu Dialogue Corpus)
Ubuntu Dialogue Corpus (UDC) is a dataset containing almost 1 million multi-turn dialogues, with a total of over 7 million utterances and 100 million words.
46 papers · 8 benchmarks
Questions is an interaction graph of users of a question-answering website based on data provided by Yandex Q.
46 papers · 1 benchmark
BEDLAM is a large-scale synthetic video dataset designed to train and test algorithms on the task of 3D human pose and shape estimation (HPS).
45 papers · 1 benchmark
The Brain-Score platform aims to yield strong computational models of the ventral stream.
45 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.