Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 34 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1585–1632 of 12,172
DeepFashion2 is a versatile benchmark of four tasks including clothes detection, pose estimation, segmentation, and retrieval.
30 papers · 0 benchmarks
Dex-Net 2.0 is a dataset associating 6.7 million point clouds and analytic grasp quality metrics with parallel-jaw grasps planned using robust quasi-static GWS analysis on a dataset of 1,500 3D object models.
30 papers · 0 benchmarks
ECtHR (European Court of Human Rights Cases)
ECtHR is a dataset comprising European Court of Human Rights cases, including annotations for paragraph-level rationales.
30 papers · 0 benchmarks
EntityQuestions is a dataset of simple, entity-rich questions based on facts from Wikidata (e.g., "Where was Arve Furset born?
30 papers · 1 benchmark
FakeNewsNet is collected from two fact-checking websites: GossipCop and PolitiFact containing news contents with labels annotated by professional journalists and experts, along with social context information.
30 papers · 0 benchmarks
The General-100 dataset is a dataset for image super-resolution.
30 papers · 0 benchmarks
The HELP dataset is an automatically created natural language inference (NLI) dataset that embodies the combination of lexical and logical inferences focusing on monotonicity (i.e., phrase replacement-based reasoning).
30 papers · 1 benchmark
A new large multiview dataset for human body expressions with natural clothing.
30 papers · 0 benchmarks
Holl-E is a dataset containing movie chats wherein each response is explicitly generated by copying and/or modifying sentences from unstructured background knowledge such as plots, comments and reviews about the movie.
30 papers · 0 benchmarks
To fix the defacts of RAVEN dataset, we generate an alternative answer set for each RPM question in RAVEN, forming an improved dataset named Impartial-RAVEN (I-RAVEN for short).
30 papers · 0 benchmarks
InteriorNet is a RGB-D for large scale interior scene understanding and mapping.
30 papers · 0 benchmarks
🤖 Robo3D - The KITTI-C Benchmark KITTI-C is an evaluation benchmark heading toward robust and reliable 3D object detection in autonomous driving.
30 papers · 1 benchmark
N-UCLA (Northwestern-UCLA Multiview Action 3D Dataset)
The Multiview 3D event dataset is capture by me and Xiaohan Nie in UCLA.
30 papers · 2 benchmarks
QuaRTz is a crowdsourced dataset of 3864 multiple-choice questions about open domain qualitative relationships.
30 papers · 0 benchmarks
SMACv2 (StarCraft Multi-Agent Challenge v2) is a new version of the benchmark where scenarios are procedurally generated and require agents to generalise to previously unseen settings (from the same distribution) during evaluation.
30 papers · 0 benchmarks
ShoeV2 is a dataset of 2,000 photos and 6648 sketches of shoes.
30 papers · 0 benchmarks
An interactive, first-person, partially-observed visual environment that uses Google Street View for its photographic content and broad coverage, and give performance baselines for a challenging goal-driven navigation task.
30 papers · 0 benchmarks
The Tox21 data set comprises 12,060 training samples and 647 test samples that represent chemical compounds.
30 papers · 4 benchmarks
URLB (Unsupervised Reinforcement Learning Benchmark)
URLB consists of two phases: reward-free pre-training and downstream task adaptation with extrinsic rewards.
30 papers · 0 benchmarks
VALUE (Video-And-Language Understanding Evaluation)
VALUE is a Video-And-Language Understanding Evaluation benchmark to test models that are generalizable to diverse tasks, domains, and datasets.
30 papers · 0 benchmarks
VGG-SS (VGG Sound Source) is a benchmark for evaluating sound source localisation in videos.
30 papers · 0 benchmarks
Video Instruction Dataset is used to train Video-ChatGPT.
30 papers · 7 benchmarks
VocalSet (VocalSet: A Singing Voice Dataset)
VocalSet is a a singing voice dataset consisting of 10.1 hours of monophonic recorded audio of professional singers demonstrating both standard and extended vocal techniques on all 5 vowels.
30 papers · 2 benchmarks
A new multilingual language model benchmark that is composed of 40+ languages spanning several scripts and linguistic families containing round 40 billion characters and aimed to accelerate the research of multilingual modeling.
30 papers · 3 benchmarks
AVD (Active Vision Dataset)
AVD focuses on simulating robotic vision tasks in everyday indoor environments using real imagery.
29 papers · 1 benchmark
CODAH (COmmonsense Dataset Adversarially-authored by Humans)
The COmmonsense Dataset Adversarially-authored by Humans (CODAH) is an evaluation set for commonsense question-answering in the sentence completion style of SWAG.
29 papers · 2 benchmarks
The COUGHVID dataset provides over 20,000 crowdsourced cough recordings representing a wide range of subject ages, genders, geographic locations, and COVID-19 statuses.
29 papers · 0 benchmarks
Comic2k is a dataset used for cross-domain object detection which contains 2k comic images with image and instance-level annotations.
29 papers · 4 benchmarks
The dataset is manually collected from Google Earth.
29 papers · 1 benchmark
GraspNet-1Billion provides large-scale training data and a standard evaluation platform for the task of general robotic grasping.
29 papers · 1 benchmark
A dataset and evaluation resource that quantifies the extent of of the semantic category membership, that is, type-of relation also known as hyponymy-hypernymy or lexical entailment (LE) relation between 2,616 concept pairs.
29 papers · 0 benchmarks
Konstanz artificially distorted image quality database (KADID-10k) contains 81 pristine images, each degraded by 25 distortions in 5 levels.
29 papers · 2 benchmarks
The MRNet dataset consists of 1,370 knee MRI exams performed at Stanford University Medical Center.
29 papers · 1 benchmark
MaRVL (Multicultural Reasoning over Vision and Language)
Multicultural Reasoning over Vision and Language (MaRVL) is a dataset based on an ImageNet-style hierarchy representative of many languages and cultures (Indonesian, Mandarin Chinese, Swahili, Tamil, and Turkish).
29 papers · 1 benchmark
MassiveText is a collection of large English-language text datasets from multiple sources: web pages, books, news articles, and code.
29 papers · 0 benchmarks
A machine reading comprehension (MRC) dataset with discourse structure built over multiparty dialog.
29 papers · 2 benchmarks
Multiface consists of high quality recordings of the faces of 13 identities, each captured in a multi-view capture stage performing various facial expressions.
29 papers · 0 benchmarks
The NVGesture dataset focuses on touchless driver controlling.
29 papers · 1 benchmark
RGB-D dataset of synthetic indoor scenes with color, noisy depth map, etc.
29 papers · 0 benchmarks
OCID (Object Clutter Indoor Dataset)
Developing robot perception systems for handling objects in the real-world requires computer vision algorithms to be carefully scrutinized with respect to the expected operating domain.
29 papers · 1 benchmark
Omni-Realm Benchmark (OmniBenchmark) is a diverse (21 semantic realm-wise datasets) and concise (realm-wise datasets have no concepts overlapping) benchmark for evaluating pre-trained model generalization across semantic…
29 papers · 1 benchmark
Useful for through two applications - automatic readability assessment and automatic text simplification.
29 papers · 0 benchmarks
It is manually annotated, comes with a naturally diverse distribution, and has a large scale.
29 papers · 0 benchmarks
PartImageNet is a large, high-quality dataset with part segmentation annotations.
29 papers · 0 benchmarks
Pushshift makes available all the submissions and comments posted on Reddit between June 2005 and April 2019.
29 papers · 0 benchmarks
Presents a diverse eye-gaze dataset.
29 papers · 1 benchmark
A dataset for rain removal with scene depth information.
29 papers · 1 benchmark
Spring (Spring: A High-Resolution High-Detail Dataset and Benchmark for Scene Flow, Optical Flow and Stereo)
Spring is a large, high-resolution and high-detail, computer-generated benchmark for scene flow, optical flow, and stereo.
29 papers · 3 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.