Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 45 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 2113–2160 of 12,172
MCubeS (Multimodal Material Segmentation Dataset)
Multimodal material segmentation (MCubeS) dataset contains 500 sets of images from 42 street scenes.
19 papers · 1 benchmark
The MECCANO dataset is the first dataset of egocentric videos to study human-object interactions in industrial-like settings.
19 papers · 3 benchmarks
This dataset code generates mathematical question and answer pairs, from a range of question types at roughly school-level difficulty.
19 papers · 1 benchmark
ModaNet is a street fashion images dataset consisting of annotations related to RGB images.
19 papers · 1 benchmark
NNE is a dataset for Nested Named Entity Recognition in English Newswire
19 papers · 1 benchmark
OVEN (Open-domain Visual Entity Recognition)
In this project, we formally present the task of Open-domain Visual Entity recognitioN (OVEN), where a model need to link an image onto a Wikipedia entity with respect to a text query.
19 papers · 1 benchmark
PKLot (A Robust Dataset for Parking Lot Classification)
The PKLot dataset contains 12,417 images of parking lots and 695,899 images of parking spaces segmented from them, which were manually checked and labeled.
19 papers · 1 benchmark
PRID 2011 is a person reidentification dataset that provides multiple person trajectories recorded from two different static surveillance cameras, monitoring crosswalks and sidewalks.
19 papers · 2 benchmarks
Physion is a visual and physical prediction benchmark to measure the performance of machine learning models on making predictions about commonplace real world physical events.
19 papers · 0 benchmarks
QAMPARI is an ODQA benchmark, where question answers are lists of entities, spread across many paragraphs.
19 papers · 0 benchmarks
A dataset on asking Questions for Lack of Clarity in open-domain information-seeking conversations.
19 papers · 0 benchmarks
SHD (Spiking Heidelberg Digits)
The Spiking Heidelberg Digits (SHD) dataset is an audio-based classification dataset of 1k spoken digits ranging from zero to nine in the English and German languages.
19 papers · 1 benchmark
SIDER contains information on marketed medicines and their recorded adverse drug reactions.
19 papers · 3 benchmarks
SSP-3D (Sports Shape and Pose 3D)
SSP-3D is an evaluation dataset consisting of 311 images of sportspersons in tight-fitted clothes, with a variety of body shapes and poses.
19 papers · 1 benchmark
STARSS22 (Sony-TAu Realistic Spatial Soundscapes 2022)
The Sony-TAu Realistic Spatial Soundscapes 2022(STARSS22) dataset consists of recordings of real scenes captured with high channel-count spherical microphone array (SMA).
19 papers · 1 benchmark
SUTD-TrafficQA (Singapore University of Technology and Design - Traffic Question Answering) is a dataset which takes the form of video QA based on 10,080 in-the-wild videos and annotated 62,535 QA pairs, for benchmarking the cognitive…
19 papers · 1 benchmark
ScribbleKITTI is a scribble-annotated dataset for LiDAR semantic segmentation.
19 papers · 2 benchmarks
There are now many computer programs for automatically determining the sense of a word in context (Word Sense Disambiguation or WSD).
19 papers · 0 benchmarks
Node classification on Squirrel with the fixed 48%/32%/20% splits provided by Geom-GCN.
19 papers · 2 benchmarks
Node classification on Squirrel with 60%/20%/20% random splits for training/validation/test.
19 papers · 1 benchmark
StaQC (Stack Overflow Question-Code pairs) is a large dataset of around 148K Python and 120K SQL domain question-code pairs, which are automatically mined from StackOverflow.
19 papers · 0 benchmarks
dataset of 400 image pairs
19 papers · 1 benchmark
Taskmaster-1 is a dialog dataset consisting of 13,215 task-based dialogs in English, including 5,507 spoken and 7,708 written dialogs created with two distinct procedures.
19 papers · 0 benchmarks
Torque is an English reading comprehension benchmark built on 3.2k news snippets with 21k human-generated questions querying temporal relationships.
19 papers · 1 benchmark
Touchdown is a corpus for executing navigation instructions and resolving spatial descriptions in visual real-world environments.
19 papers · 1 benchmark
With social media becoming increasingly popular on which lots of news and real-time events are reported, developing automated question answering systems is critical to the effectiveness of many applications that rely on real-time knowledge.
19 papers · 1 benchmark
UMLS (Unified Medical Language System)
The Unified Medical Language System (UMLS) is a comprehensive resource that integrates and disseminates essential terminology, classification standards, and coding systems.
19 papers · 1 benchmark
Contains data from three platforms, i.e., synthetic drones, satellites and ground cameras of 1,652 university buildings around the world.
19 papers · 2 benchmarks
Person re-identification (Reid) is now an active research topic for AI-based video surveillance applications such as specific person search, but the practical issue that the target person(s) may change clothes (clothes inconsistency…
19 papers · 2 benchmarks
A dataset with fully annotated attention targets in video for attention target estimation.
19 papers · 1 benchmark
WanJuan is a large-scale training corpus that includes multiple modalities.
19 papers · 0 benchmarks
The ClinTox dataset compares drugs approved by the FDA and drugs that have failed clinical trials for toxicity reasons.
19 papers · 3 benchmarks
The rounD dataset introduces a fresh compilation of natural road user trajectory data from German roundabouts, gathered using drone technology to navigate past usual challenges such as occlusions inherent in traditional traffic data…
19 papers · 0 benchmarks
2010 i2b2/VA is a biomedical dataset for relation classification and entity typing.
18 papers · 4 benchmarks
Collects high quality 360 datasets with ground truth depth annotations, by re-using recently released large scale 3D datasets and re-purposing them to 360 via rendering.
18 papers · 0 benchmarks
AI-TOD (Tiny Object Detection in Aerial Images)
AI-TOD comes with 700,621 object instances for eight categories across 28,036 aerial images.
18 papers · 2 benchmarks
ActivityNet-Entities, augments the challenging ActivityNet Captions dataset with 158k bounding box annotations, each grounding a noun phrase.
18 papers · 0 benchmarks
The AggreFact dataset is a benchmark for evaluating the factuality of summaries generated by different summarization models.
18 papers · 1 benchmark
A large-scale aerial farmland image dataset for semantic segmentation of agricultural patterns.
18 papers · 0 benchmarks
This dataset includes reviews (ratings, text, helpfulness votes), product metadata (descriptions, category information, price, brand, and image features), and links (also viewed/also bought graphs).
18 papers · 3 benchmarks
BOBSL (BBC-Oxford British Sign Language)
BOBSL is a large-scale dataset of British Sign Language (BSL).
18 papers · 1 benchmark
BURST is a benchmark suite built upon TAO that requires tracking and segmenting multiple objects from camera video.
18 papers · 5 benchmarks
Bridge Data is a large multi-domain and multi-task dataset, with 7,200 demonstrations constituting 71 tasks across 10 environments.
18 papers · 0 benchmarks
CASIA-B is a large multiview gait database, which is created in January 2005.
18 papers · 1 benchmark
Car Crash Dataset (CCD) is collected for traffic accident analysis.
18 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.