Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 39 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1825–1872 of 12,172
InterHuman is a multimodal dataset, named InterHuman.
24 papers · 1 benchmark
KELM is a large-scale synthetic corpus of Wikidata KG as natural text.
24 papers · 0 benchmarks
The Segmenting and Tracking Every Pixel (STEP) benchmark consists of 21 training sequences and 29 test sequences.
24 papers · 2 benchmarks
KeypointNet is a large-scale and diverse 3D keypoint dataset that contains 83,231 keypoints and 8,329 3D models from 16 object categories, by leveraging numerous human annotations, based on ShapeNet models.
24 papers · 0 benchmarks
LOL-v2-real contains 689 low-/normal-light image pairs for training and 100 pairs for testing.
24 papers · 1 benchmark
MCScript is used as the official dataset of SemEval2018 Task11.
24 papers · 0 benchmarks
MMSE-HR (Multimodal Spontaneous Expression-Heart Rate dataset)
The MMSE-HR benchmark consists of a dataset of 102 videos from 40 subjects recorded at 1040x1392 raw resolution at 25fps.
24 papers · 1 benchmark
The dataset aims to find the algorithms that produce the most visually pleasant image possible and generalize well to a broad range of content.
24 papers · 1 benchmark
Moral Stories is a crowd-sourced dataset of structured narratives that describe normative and norm-divergent actions taken by individuals to accomplish certain intentions in concrete situations, and their respective consequences.
24 papers · 0 benchmarks
NAS-Bench-1Shot1 draws on the recent large-scale tabular benchmark NAS-Bench-101 for cheap anytime evaluations of one-shot NAS methods.
24 papers · 0 benchmarks
The Natural Stories dataset consists of English texts edited to contain many low-frequency syntactic constructions while still sounding fluent to native speakers.
24 papers · 0 benchmarks
A new video dataset for aerial view concurrent human action detection.
24 papers · 1 benchmark
QM7 dataset is a subset of the GDB-13 database.
24 papers · 1 benchmark
RADIATE (RAdar Dataset In Adverse weaThEr)
RADIATE (RAdar Dataset In Adverse weaThEr) is new automotive dataset created by Heriot-Watt University which includes Radar, Lidar, Stereo Camera and GPS/IMU.
24 papers · 2 benchmarks
REALY (Region-aware benchmark based on the LYHM)
The REALY benchmark aims to introduce a region-aware evaluation pipeline to measure the fine-grained normalized mean square error (NMSE) of 3D face reconstruction methods from under-controlled image sets.
24 papers · 2 benchmarks
Reddit12k contains 11929 graphs each corresponding to an online discussion thread where nodes represent users, and an edge represents the fact that one of the two users responded to the comment of the other user.
24 papers · 1 benchmark
ROPES (Reasoning Over Paragraph Effects in Situations)
ROPES is a QA dataset which tests a system's ability to apply knowledge from a passage of text to a new situation.
24 papers · 0 benchmarks
RecipeQA is a dataset for multimodal comprehension of cooking recipes.
24 papers · 1 benchmark
SIMMC (Situated and Interactive Multimodal Conversations)
Situated Interactive MultiModal Conversations (SIMMC) is the task of taking multimodal actions grounded in a co-evolving multimodal input content in addition to the dialog history.
24 papers · 0 benchmarks
SQUID (Stereo Quantitative Underwater Image Dataset)
A dataset of images taken in different locations with varying water properties, showing color charts in the scenes.
24 papers · 0 benchmarks
In SpokenSQuAD, the document is in spoken form, the input question is in the form of text and the answer to each question is always a span in the document.
24 papers · 1 benchmark
Here, we take a key step in this direction and release a new benchmark, TempQuestions, containing 1,271 questions, that are all temporal in nature, paired with their answers.
24 papers · 1 benchmark
Toronto-3D is a large-scale urban outdoor point cloud dataset acquired by an MLS system in Toronto, Canada for semantic segmentation.
24 papers · 2 benchmarks
The Ubuntu IRC dataset is a valuable resource for research in natural language understanding and dialogue systems.
24 papers · 2 benchmarks
VocalSound is a free dataset consisting of 21,024 crowdsourced recordings of laughter, sighs, coughs, throat clearing, sneezes, and sniffs from 3,365 unique subjects.
24 papers · 1 benchmark
News translation is a recurring WMT task.
24 papers · 0 benchmarks
The Wiki-ZSL (Wiki Zero-Shot Learning) dataset contains 113 relations and 94,383 instances from Wikipedia.
24 papers · 1 benchmark
eLife (Scientific Lay Summarization)
This dataset contains 4,828 full biomedical articles paired with non-technical lay summaries derived from the eLife scientific journal.
24 papers · 2 benchmarks
node classification on twitch-gamers
24 papers · 2 benchmarks
4D-DRESS (A 4D Dataset of Real-world Human Clothing with Semantic Annotations)
4D-DRESS is the first real-world 4D dataset of human clothing, capturing 64 human outfits in more than 520 motion sequences.
23 papers · 4 benchmarks
AIM-500 (Automatic Image Matting-500)
AIM-500 is the first natural image matting test set, contains 500 high-resolution real-world natural images from three types of images (salient opaque foregrounds, salient transparent/meticulous foregrounds, non-salient foregrounds), and…
23 papers · 1 benchmark
ARCTIC (Articulated Objects in Free-form Hand Interaction)
ARCTIC is a dataset of free-form interactions of hands and articulated objects.
23 papers · 0 benchmarks
ATOM3D is a unified collection of datasets concerning the three-dimensional structure of biomolecules, including proteins, small molecules, and nucleic acids.
23 papers · 0 benchmarks
We propose EMAGE, a framework to generate full-body human gestures from audio and masked gestures, encompassing facial, local body, hands, and global movements.
23 papers · 2 benchmarks
CADC (Canadian Adverse Driving Conditions)
Collected with the Autonomoose autonomous vehicle platform, based on a modified Lincoln MKZ.
23 papers · 0 benchmarks
CIFAR10-DVS is an event-stream dataset for object classification.
23 papers · 2 benchmarks
CharXiv is a comprehensive evaluation suite for testing the chart understanding capabilities of Multimodal Large Language Models (MLLMs)¹².
23 papers · 0 benchmarks
CoSQA (Code Search and Question Answering)
CoSQA (Code Search and Question Answering) It includes 20,604 labels for pairs of natural language queries and codes, each annotated by at least 3 human annotators.
23 papers · 0 benchmarks
DSEC (A Stereo Event Camera Dataset for Driving Scenarios)
DSEC is a stereo camera dataset in driving scenarios that contains data from two monochrome event cameras and two global shutter color cameras in favorable and challenging illumination conditions.
23 papers · 2 benchmarks
The DeepWeeds dataset consists of 17,509 images capturing eight different weed species native to Australia in situ with neighbouring flora.
23 papers · 0 benchmarks
DocUNet (Document Image Unwarping via a Stacked U-Net)
Various documents dataset.
23 papers · 3 benchmarks
ELD (Extreme Low-light Denoising dataset)
Extreme low-light denoising (ELD) dataset that covers 10 indoor scenes and 4 camera devices from multiple brands (SonyA7S2, NikonD850, CanonEOS70D, CanonEOS700D).
23 papers · 2 benchmarks
EMOPIA (A Multi-Modal Pop Piano Dataset For Emotion Recognition and Emotion-based Music Generation)
EMOPIA (pronounced ‘yee-mò-pi-uh’) dataset is a shared multi-modal (audio and MIDI) database focusing on perceived emotion in pop piano music, to facilitate research on various tasks related to music emotion.
23 papers · 0 benchmarks
EURLEX57K is a new publicly available legal LMTC dataset, dubbed EURLEX57K, containing 57k English EU legislative documents from the EUR-LEX portal, tagged with ∼4.3k labels (concepts) from the European Vocabulary (EUROVOC).
23 papers · 1 benchmark
We construct a fine-grained video dataset organized by both semantic and temporal structures, where each structure contains two-level annotations.
23 papers · 1 benchmark
A large-scale 4D egocentric dataset with rich annotations, to catalyze the research of category-level human-object interaction.
23 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.