Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 86 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 4081–4128 of 12,172
ShipSG (Ship Segmentation and Georeferencing Dataset)
The ShipSG dataset is the first public dataset of its kind for ship segmentation and georeferencing.
6 papers · 0 benchmarks
SimJEB (Simulated Jet Engine Bracket)
Simulated Jet Engine Bracket Dataset (SimJEB) is a public collection of crowdsourced mechanical brackets and high-fidelity structural simulations designed specifically for surrogate modeling.
6 papers · 0 benchmarks
SimpleQuestionsWikidata maps SimpleQuestions to Wikidata.
6 papers · 1 benchmark
SocNav1 is a dataset for social navigation conventions.
6 papers · 0 benchmarks
The data and audio included here were collected for the Soundscape Attributes Translation Project (SATP).
6 papers · 0 benchmarks
SpeechMatrix is a large-scale multilingual corpus of speech-to-speech translations mined from real speech of European Parliament recordings.
6 papers · 0 benchmarks
A set of approximately 100K podcast episodes comprised of raw audio files along with accompanying ASR transcripts.
6 papers · 0 benchmarks
SuperCLUE is a Chinese language model evaluation benchmark named after another popular Chinese LLM benchmark CLUE.
6 papers · 0 benchmarks
Swords (Stanford Word Substitution benchmark)
Swords (Standford Word Substitution) is a benchmark for lexical substitution, the task of finding appropriate substitutes for a target word in a context.
6 papers · 0 benchmarks
TREC-10 (TREC-10 Question Classification)
A question type classification dataset with 6 classes for questions about a person, location, numeric information, etc.
6 papers · 1 benchmark
TRECVID is a yearly set of competitions centered on video retrieval and indexing, hosting a variety of video data sets.
6 papers · 1 benchmark
TURL (Twitter News URL Corpus)
Twitter News URL Corpus is a human-labeled paraphrase corpus to date of 51,524 sentence pairs and the first cross-domain benchmarking for automatic paraphrase identification.
6 papers · 1 benchmark
The TalkSumm dataset contains 1705 automatically-generated summaries of scientific papers from ACL, NAACL, EMNLP, SIGDIAL (2015-2018), and ICML (2017-2018).
6 papers · 0 benchmarks
This dataset contains information collected by the U.S Census Service concerning housing in the area of Boston Mass.
6 papers · 0 benchmarks
Collected from top 10 most popular clothing/wearable brandname logos captured in rich visual context.
6 papers · 1 benchmark
The dataset comprises multiple independent events, where each event contains simulated measurements (essentially 3D points) of particles generated in a collision between proton bunches at the Large Hadron Collider at CERN.
6 papers · 0 benchmarks
TutorialBank is a publicly available dataset which aims to facilitate NLP education and research.
6 papers · 0 benchmarks
A dataset for building models that detect people Looking At Each Other (LAEO) in video sequences.
6 papers · 0 benchmarks
UDD is an underwater open-sea farm object detection dataset.
6 papers · 0 benchmarks
UIT-ViIC contains manually written captions for images from Microsoft COCO dataset relating to sports played with ball.
6 papers · 0 benchmarks
We present a further analysis of visual modality incompleteness, benchmarking latest MMEA models on our proposed dataset MMEA-UMVM.
6 papers · 3 benchmarks
UTSD (Unified Time Series Dataset)
Unified Time Series Dataset (UTSD) includes 7 domains with up to 1 billion time points with hierarchical capacities to facilitate research of large models in the field of time series.
6 papers · 0 benchmarks
Ulm-TSST (Ulm-Trier Social Stress Dataset)
Ulm-TSST is a dataset continuous emotion (valence and arousal) prediction and physiological-emotion' prediction.
6 papers · 0 benchmarks
Ultra-high definition benchmark (UHDBench) includes 2293 images at 2k resolution sourced from the ground-truth test sets of HRSOD, LIU4k, UAVid, UHDM, and UHRSD.
6 papers · 1 benchmark
The Urban Environments dataset is a dataset of 20 land use classes across 300 European cities paired with satellite imagery data.
6 papers · 0 benchmarks
V2C (Video-to-Commonsense)
6 papers · 0 benchmarks
VDD (Varied Drone Dataset for Semantic Segmentation)
Semantic segmentation of drone images is critical for various aerial vision tasks as it provides essential seman- tic details to understand scenes on the ground.
6 papers · 1 benchmark
VEDAI (Vehicle Detection in Aerial Imagery)
VEDAI is a dataset for Vehicle Detection in Aerial Imagery, provided as a tool to benchmark automatic target recognition algorithms in unconstrained environments.
6 papers · 1 benchmark
VIPER is a benchmark suite for visual perception.
6 papers · 0 benchmarks
VMRD (Visual Manipulation Relationship Dataset)
VMRD is a multi-object grasp dataset.
6 papers · 0 benchmarks
ViNLI (Vietnamese Natural Language Inference Dataset)
A large-scale and high-quality corpus is necessary for studies on NLI for Vietnamese, which can be considered a low-resource language.
6 papers · 1 benchmark
VideoCube is a high-quality and large-scale benchmark to create a challenging real-world experimental environment for Global Instance Tracking (GIT).
6 papers · 1 benchmark
VideoLT is a large-scale long-tailed video recognition dataset that contains 256,218 untrimmed videos, annotated into 1,004 classes with a long-tailed distribution.
6 papers · 0 benchmarks
VisPro dataset contains coreference annotation of 29,722 pronouns from 5,000 dialogues.
6 papers · 0 benchmarks
WDC Products is an entity matching benchmark which provides for the systematic evaluation of matching systems along combinations of three dimensions while relying on real-word data.
6 papers · 4 benchmarks
WHOI-Plankton is a collection of annotated plankton images.
6 papers · 0 benchmarks
WNLaMPro (WordNet Language Model Probing)
The WordNet Language Model Probing (WNLaMPro) dataset consists of relations between keywords and words.
6 papers · 0 benchmarks
WNUT 2020 (WNUT-2020 Task 1 Overview: Extracting Entities and Relations from Wet Lab Protocols)
The training and development dataset for our task was taken from previous work on wet lab corpus (Kulkarni et al., 2018) that consists of from the 623 protocols.
6 papers · 2 benchmarks
WSJ0-2mix-extr is a speech extraction dataset
6 papers · 1 benchmark
Multi-level Benchmark of Watermarks for Large Language Models
6 papers · 0 benchmarks
Aims to facilitate research in caricature recognition.
6 papers · 0 benchmarks
WebLINX (Real-World Website Navigation with Multi-Turn)
WebLINX is a large-scale benchmark of 100K interactions across 2300 expert demonstrations of conversational web navigation.
6 papers · 1 benchmark
An unsupervised dataset for co-reference resolution.
6 papers · 0 benchmarks
WikiNEuRal is a high-quality automatically-generated dataset for Multilingual Named Entity Recognition.
6 papers · 0 benchmarks
The WikiNews Arabic Diacritization dataset is a test set composed of 70 WikiNews articles (majority are from 2013 and 2014) that cover a variety of themes, namely: politics, economics, health, science and technology, sports, arts, and…
6 papers · 0 benchmarks
Wild-Time is a benchmark of 5 datasets that reflect temporal distribution shifts arising in a variety of real-world applications, including patient prognosis and news classification.
6 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.