Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 38 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1777–1824 of 12,172
WikiReading is a large-scale natural language understanding task and publicly-available dataset with 18 million instances.
26 papers · 0 benchmarks
XStoryCloze consists of the professionally translated version of the English StoryCloze dataset (Spring 2016 version) to 10 non-English languages.
26 papers · 0 benchmarks
ASNQ (Answer Sentence Natural Questions)
A large scale dataset to enable the transfer step, exploiting the Natural Questions dataset.
25 papers · 1 benchmark
BAM! (Behance Artistic Media)
The Behance Artistic Media dataset (BAM!) is a large-scale dataset of contemporary artwork from Behance, a website containing millions of portfolios from professional and commercial artists.
25 papers · 0 benchmarks
BDD-A (Berkeley DeepDrive Attention)
Dataset Statistics: The statistics of our dataset are summarized and compared with the largest existing dataset (DR(eye)VE) [1] in Table 1.
25 papers · 0 benchmarks
BIPED (Barcelona Images for Perceptual Edge Detection)
Details It contains 250 outdoor images of 1280×720 pixels each.
25 papers · 1 benchmark
BioRED is a first-of-its-kind biomedical relation extraction dataset with multiple entity types (e.g.
25 papers · 3 benchmarks
CH-SIMS is a Chinese single- and multimodal sentiment analysis dataset which contains 2,281 refined video segments in the wild with both multimodal and independent unimodal annotations.
25 papers · 1 benchmark
Chest ImaGenome is a dataset with a scene graph data structure to describe 242,072 images.
25 papers · 0 benchmarks
CrisisMMD is a large multi-modal dataset collected from Twitter during different natural disasters.
25 papers · 0 benchmarks
CrossWOZ is the first large-scale Chinese Cross-Domain Wizard-of-Oz task-oriented dataset.
25 papers · 0 benchmarks
DIOR-RSVG is a large-scale benchmark dataset of remote sensing data (RSVG).
25 papers · 0 benchmarks
DND (Darmstadt Noise Dataset)
Benchmarking Denoising Algorithms with Real Photographs This dataset consists of 50 pairs of noisy and (nearly) noise-free images captured with four consumer cameras.
25 papers · 2 benchmarks
ELEVATER (Evaluation of Language-augmented Visual Task-level Transfer)
The ELEVATER benchmark is a collection of resources for training, evaluating, and analyzing language-image models on image classification and object detection.
25 papers · 2 benchmarks
ESOL is a water solubility prediction dataset consisting of 1128 samples.
25 papers · 4 benchmarks
HONEST (Hurtful Sentence Completion in English Language Models)
The HONEST dataset is a template-based corpus for testing the hurtfulness of sentence completions in language models (e.g., BERT) in six different languages (English, Italian, French, Portuguese, Romanian, and Spanish).
25 papers · 1 benchmark
A three million frame, multi-view, furniture assembly video dataset that includes depth, atomic actions, object segmentation, and human pose.
25 papers · 1 benchmark
A parallel corpus of over 300 languages with around 100 thousand parallel sentences per language pair on average.
25 papers · 0 benchmarks
LargeST (LargeST: A Benchmark Dataset for Large-Scale Traffic Forecasting)
In this work, we propose LargeST as a new benchmark dataset (see Figure 1), with the goal of facilitating the development of accurate and efficient methods in the context of large-scale traffic forecasting.
25 papers · 1 benchmark
This is a dataset for a video super-resolution task.
25 papers · 1 benchmark
The MULTEXT-East resources are a multilingual dataset for language engineering research and development.
25 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
25 papers · 0 benchmarks
OASST1 (OpenAssistant Conversations Dataset)
license: apache-2.0 tags: human-feedback sizecategories: 100K Languages with under 1000 messages Vietnamese: 952 Basque: 947 Polish: 886 Hungarian: 811 Arabic: 666 Dutch: 628 Swedish: 512 Turkish: 454 Finnish: 386 Czech: 372 Danish: 358…
25 papers · 0 benchmarks
PeMS07 is a traffic forecasting benchmark.
25 papers · 1 benchmark
ROSE (Retinal OCTA SEgmentation dataset)
Retinal OCTA SEgmentation dataset (ROSE) consists of 229 OCTA images with vessel annotations at either centerline-level or pixel level.
25 papers · 4 benchmarks
The signing is recorded by a stationary color camera placed in front of the sign language interpreters.
25 papers · 1 benchmark
ScanNet++ (ScanNet++: A High-Fidelity Dataset of 3D Indoor Scenes)
ScanNet++ is a large scale dataset with 450+ 3D indoor scenes containing sub-millimeter resolution laser scans, registered 33-megapixel DSLR images, and commodity RGB-D streams from iPhone.
25 papers · 5 benchmarks
TOPv2 (Task Oriented Parsing v2)
Task Oriented Parsing v2 (TOPv2) representations for intent-slot based dialog systems.
25 papers · 0 benchmarks
TinyPerson is a benchmark for tiny object detection in a long distance and with massive backgrounds.
25 papers · 0 benchmarks
ZESHEL is a zero-shot entity linking dataset, which places more emphasis on understanding the unstructured descriptions of entities to resolve the ambiguity of mentions on four unseen domains.
25 papers · 1 benchmark
pathbased is a 3-cluster data set.
25 papers · 2 benchmarks
Description: 105,941 Images Natural Scenes OCR Data of 12 Languages.
24 papers · 0 benchmarks
AFLW-19 (The 19 landmark variant of AFLW.)
The original AFLW provides at most 21 points for each face, but excluding coordinates for invisible landmarks, causing difficulties for training most of the existing baseline approaches.
24 papers · 1 benchmark
Composed of 1,395 questions posed by crowdworkers on Wikipedia articles, and a machine translation of the Stanford Question Answering Dataset (Arabic-SQuAD).
24 papers · 0 benchmarks
A benchmark dataset for the Aspect Sentiment Triplet Extraction, an updated version of ASTE-Data-V1.
24 papers · 1 benchmark
Amazon Sports (Amazon Sports 5-core)
24 papers · 1 benchmark
Dataset for face anti-spoofing in terms of both subjects and modalities.
24 papers · 0 benchmarks
CCPD (Chinese City Parking Dataset)
The Chinese City Parking Dataset (CCPD) is a dataset for license plate detection and recognition.
24 papers · 0 benchmarks
DADA-seg is a pixel-wise annotated accident dataset, which contains a variety of critical scenarios from traffic accidents.
24 papers · 1 benchmark
DHF1K is a video saliency dataset which contains a ground-truth map of binary pixel-wise gaze fixation points and a continuous map of the fixation points after being blurred by a gaussian filter.
24 papers · 1 benchmark
FineAction contains 103K temporal instances of 106 action categories, annotated in 17K untrimmed videos.
24 papers · 3 benchmarks
FlickrStyle10K is collected and built on Flickr30K image caption dataset.
24 papers · 2 benchmarks
The FreeSolv database offers a curated collection of experimental and calculated hydration-free energies for small molecules in water.
24 papers · 2 benchmarks
GeoS is a dataset for automatic math problem solving.
24 papers · 1 benchmark
The Google Landmarks dataset contains 1,060,709 images from 12,894 landmarks, and 111,036 additional query images.
24 papers · 0 benchmarks
HarMeme is a benchmark dataset for hateful meme classification containing 3, 544 memes related to COVID-19 collected from the Internet
24 papers · 1 benchmark
The INRIA Person dataset is a dataset of images of persons used for pedestrian detection.
24 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.