Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 30 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1393–1440 of 12,172
MAFL (Multi-Attribute Facial Landmark)
The MAFL dataset contains manually annotated facial landmark locations for 19,000 training and 1,000 test images.
36 papers · 1 benchmark
METR-LA is a dataset for traffic prediction.
36 papers · 3 benchmarks
The MultiModal Safety Benchmark (MM-SafetyBench) is a comprehensive framework designed for conducting safety-critical evaluations of Multimodal Large Language Models (MLLMs).
36 papers · 0 benchmarks
The MSLR-WEB10K dataset consists of 10,000 search queries over the documents from search results.
36 papers · 0 benchmarks
The Microsoft Malware Classification Challenge was announced in 2015 along with a publication of a huge dataset of nearly 0.5 terabytes, consisting of disassembly and bytecode of more than 20K malware samples.
36 papers · 1 benchmark
The Open Table-and-Text Question Answering (OTT-QA) dataset contains open questions which require retrieving tables and text from the web to answer.
36 papers · 1 benchmark
PanLex translates words in thousands of languages.
36 papers · 0 benchmarks
Enables detailed human body model reconstruction in clothing from a single monocular RGB video without requiring a pre scanned template or manually clicked points.
36 papers · 0 benchmarks
RAP (Richly Annotated Pedestrian)
The Richly Annotated Pedestrian (RAP) dataset is a dataset for pedestrian attribute recognition.
36 papers · 1 benchmark
Our project (STPLS3D) aims to provide a large-scale aerial photogrammetry dataset with synthetic and real annotated 3D point clouds for semantic and instance segmentation tasks.
36 papers · 3 benchmarks
SciTSR is a large-scale table structure recognition dataset, which contains 15,000 tables in PDF format and their corresponding structure labels obtained from LaTeX source files.
36 papers · 0 benchmarks
TEACh (Task-driven Embodied Agents that Chat)
Robots operating in human spaces must be able to engage in natural language interaction with people, both understanding and executing instructions, and using conversation to resolve ambiguity and recover from mistakes.
36 papers · 0 benchmarks
A large-scale V2X perception dataset using CARLA and OpenCDA
36 papers · 1 benchmark
The Visual Object Tracking (VOT) dataset is a collection of video sequences used for evaluating and benchmarking visual object tracking algorithms.
36 papers · 0 benchmarks
VisualMRC (VisualMRC: Machine Reading Comprehension on Document Images)
VisualMRC is a visual machine reading comprehension dataset that proposes a task: given a question and a document image, a model produces an abstractive answer.
36 papers · 1 benchmark
stocknet-dataset This repository releases a comprehensive dataset for stock movement prediction from tweets and historical stock prices.
36 papers · 1 benchmark
Activity recognition research has shifted focus from distinguishing full-body motion patterns to recognizing complex interactions of multiple entities.
35 papers · 2 benchmarks
Attribution, Relation, and Order (ARO) benchmark to systematically evaluate the ability of VLMs to understand different types of relationships, attributes, and order information.
35 papers · 0 benchmarks
description withheld: archive row vandalised before snapshot
35 papers · 6 benchmarks
Amazon Review is a dataset to tackle the task of identifying whether the sentiment of a product review is positive or negative.
35 papers · 1 benchmark
Arxiv HEP-TH (high energy physics theory) citation graph is from the e-print arXiv and covers all the citations within a dataset of 27,770 papers with 352,807 edges.
35 papers · 5 benchmarks
The BirdSong dataset consists of audio recordings of bird songs at the H.
35 papers · 0 benchmarks
CIRCO (Composed Image Retrieval on Common Objects in context)
CIRCO (Composed Image Retrieval on Common Objects in context) is an open-domain benchmarking dataset for Composed Image Retrieval (CIR) based on real-world images from COCO 2017 unlabeled set.
35 papers · 1 benchmark
CMNLI (Chinese Multi-Genre NLI)
The CMNLI dataset is part of the Chinese Language Understanding Evaluation (CLUE) benchmark.
35 papers · 0 benchmarks
Contains hundreds of frontal view X-rays and is the largest public resource for COVID-19 image and prognostic data, making it a necessary resource to develop and evaluate tools to aid in the treatment of COVID-19.
35 papers · 1 benchmark
CaRB (Crowdsourced automatic open Relation extraction Benchmark)
CaRB [Bhardwaj et al., 2019] is developed by re-annotating the dev and test splits of OIE2016 via crowd-sourcing.
35 papers · 1 benchmark
DeepCAD is a CAD dataset consisting of 179,133 models and their CAD construction sequences.
35 papers · 2 benchmarks
EgoBody dataset is a novel large-scale dataset for egocentric 3D human pose, shape and motions under interactions in complex 3D scenes.
35 papers · 1 benchmark
This dataset contains complex tables from the annual reports of S&P 500 companies with detailed table structure annotations to help table structure recognition and table data extraction.
35 papers · 0 benchmarks
This is the second version of the Google Landmarks dataset (GLDv2), which contains images annotated with labels representing human-made and natural landmarks.
35 papers · 4 benchmarks
Is a dataset for many-hop evidence extraction and fact verification.
35 papers · 0 benchmarks
Imagenet64 is a massive dataset of small images called the down-sampled version of Imagenet.
35 papers · 1 benchmark
MARC (Multilingual Amazon Reviews Corpus)
Multilingual Amazon Reviews Corpus (MARC) is a large-scale collection of Amazon reviews for multilingual text classification.
35 papers · 0 benchmarks
This dataset includes time-series data generated by accelerometer and gyroscope sensors (attitude, gravity, userAcceleration, and rotationRate).
35 papers · 0 benchmarks
MuCo-3DHP is a large scale training data set showing real images of sophisticated multi-person interactions and occlusions.
35 papers · 0 benchmarks
The Open Entity dataset is a collection of about 6,000 sentences with fine-grained entity types annotations.
35 papers · 2 benchmarks
PHYRE (PHYsical REasoning)
Benchmark for physical reasoning that contains a set of simple classical mechanics puzzles in a 2D physical environment.
35 papers · 2 benchmarks
PST900 is a dataset of 894 synchronized and calibrated RGB and Thermal image pairs with per pixel human annotations across four distinct classes from the DARPA Subterranean Challenge.
35 papers · 1 benchmark
This dataset was designed for contextual investigations, with related works making considerable usage of said context.
35 papers · 3 benchmarks
STRING is a collection of protein-protein interaction (PPI) networks.
35 papers · 0 benchmarks
SVT (Street View Text Dataset)
The Street View Text (SVT) dataset was harvested from Google Street View.
35 papers · 1 benchmark
Spot-the-diff is a dataset consisting of 13,192 image pairs along with corresponding human provided text annotations stating the differences between the two images.
35 papers · 0 benchmarks
V2V4Real is a large-scale real-world multi-modal dataset for V2V perception.
35 papers · 0 benchmarks
VSPW (Video Scene Parsing in the Wild)
A Large-scale Dataset for Video Scene Parsing in the Wild
35 papers · 1 benchmark
A new dataset of goal-oriented dialogues that are grounded in the associated documents.
35 papers · 0 benchmarks
A dataset for robot grasp planning based on physics simulation.
34 papers · 0 benchmarks
BioGRID (Biological General Repository for Interaction Datasets)
BioGRID is a biomedical interaction repository with data compiled through comprehensive curation efforts.
34 papers · 2 benchmarks
CLEAR is a continual image classification benchmark dataset with a natural temporal evolution of visual concepts in the real world that spans a decade (2004-2014).
34 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.