Home › Datasets › task › Anomaly Detection
Anomaly Detection datasets
archive 2025-07-28
119 datasets carry the task tag "Anomaly Detection" (the task itself: Anomaly Detection), ordered by the archive's paper count. Page 1 of 3: 48 shown of 119. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Anomaly Detection datasets 1–48 of 119
description withheld: archive row vandalised before snapshot
16,145 papers · 91 benchmarks
The CIFAR-100 dataset (Canadian Institute for Advanced Research, 100 classes) is a subset of the Tiny Images dataset and consists of 60000 32x32 color images.
9,045 papers · 51 benchmarks
The MNIST database (Modified National Institute of Standards and Technology database) is a large collection of handwritten digits.
7,651 papers · 44 benchmarks
SVHN (Street View House Numbers)
Street View House Numbers (SVHN) is a digit classification benchmark dataset that contains 600,000 32×32 RGB images of printed digits (from 0 to 9) cropped from pictures of house number plates.
3,406 papers · 12 benchmarks
Fashion-MNIST is a dataset comprising of 28×28 grayscale images of 70,000 fashion products from 10 categories, with 7,000 images per category.
3,202 papers · 15 benchmarks
STL-10 (Self-Taught Learning 10)
The STL-10 is an image dataset derived from ImageNet and popularly used to evaluate algorithms of unsupervised feature learning or self-taught learning.
1,092 papers · 18 benchmarks
AG News (AG’s News Corpus) is a subdataset of AG's corpus of news articles constructed by assembling titles and description fields of articles from the 4 largest classes (“World”, “Sports”, “Business”, “Sci/Tech”) of AG’s Corpus.
969 papers · 9 benchmarks
MVTecAD (MVTEC ANOMALY DETECTION DATASET)
MVTec AD is a dataset for benchmarking anomaly detection methods with a focus on industrial inspection.
402 papers · 4 benchmarks
The Shanghaitech dataset is a large-scale crowd counting dataset.
277 papers · 5 benchmarks
The ShanghaiTech Campus dataset has 13 scenes with complex light conditions and camera angles.
207 papers · 4 benchmarks
The UCF-Crime dataset is a large-scale dataset of 128 hours of videos.
142 papers · 3 benchmarks
MSL (Mars Science Laboratory)
This dataset contains expert-labeled telemetry anomaly data from the Mars Science Laboratory (MSL) rover, Curiosity.
130 papers · 1 benchmark
An open access benchmark dataset comprising of 13,975 CXR images across 13,870 patient cases, with the largest number of publicly available COVID-19 positive cases to the best of the authors' knowledge.
96 papers · 1 benchmark
The UCSD Anomaly Detection Dataset was acquired with a stationary camera mounted at an elevation, overlooking pedestrian walkways.
92 papers · 4 benchmarks
VisA (Visual Anomaly Dataset)
The VisA dataset contains 12 subsets corresponding to 12 different objects as shown in the above figure.
86 papers · 3 benchmarks
The Yelp Dataset is a valuable resource for academic research, teaching, and learning.
86 papers · 15 benchmarks
The original ionosphere dataset from UCI machine learning repository is a binary classification dataset with dimensionality 34.
67 papers · 2 benchmarks
NAB (Numenta Anomaly Benchmark)
The First Temporal Benchmark Designed to Evaluate Real-time Anomaly Detectors Benchmark The growth of the Internet of Things has created an abundance of streaming data.
66 papers · 1 benchmark
BTAD (beanTech Anomaly Detection)
The BTAD ( beanTech Anomaly Detection) dataset is a real-world industrial anomaly dataset.
61 papers · 2 benchmarks
Lost and Found is a novel lost-cargo image sequence dataset comprising more than two thousand frames with pixelwise annotations of obstacle and free-space and provide a thorough comparison to several stereo-based baseline methods.
57 papers · 1 benchmark
This dataset contains images of unusual dangers which can be encountered by a vehicle on the road – animals, rocks, traffic cones and other obstacles.
56 papers · 1 benchmark
Fishyscapes is a public benchmark for uncertainty estimation in a real-world task of semantic segmentation for urban driving.
51 papers · 2 benchmarks
Avenue Dataset contains 16 training and 21 testing video clips.
48 papers · 2 benchmarks
UBnormal (University of Bucharest Abnormal Videos)
UBnormal is a new supervised open-set benchmark composed of multiple virtual scenes for video anomaly detection.
47 papers · 3 benchmarks
MVTec 3D Anomaly Detection Dataset (MVTec 3D-AD) is a comprehensive 3D dataset for the task of unsupervised anomaly detection and localization.
45 papers · 4 benchmarks
MPDD (Metal Parts Defect Detection Dataset)
MPDD is a dataset aimed at benchmarking visual defect detection methods in industrial metal parts manufacturing.
44 papers · 1 benchmark
A large dataset of musculoskeletal radiographs containing 40,561 images from 14,863 studies, where each study is manually labeled by radiologists as either normal or abnormal.
43 papers · 0 benchmarks
CASIA-FASD is a small face anti-spoofing dataset containing 50 subjects.
42 papers · 0 benchmarks
Sound Dataset for Malfunctioning Industrial Machine Investigation and Inspection (MIMII) is a sound dataset of industrial machine sounds.
40 papers · 0 benchmarks
MVTec Logical Constraints Anomaly Detection (MVTec LOCO AD) dataset is intended for the evaluation of unsupervised anomaly localization algorithms.
32 papers · 1 benchmark
The MIT-BIH Arrhythmia Database contains 48 half-hour excerpts of two-channel ambulatory ECG recordings, obtained from 47 subjects studied by the BIH Arrhythmia Laboratory between 1975 and 1979.
31 papers · 5 benchmarks
ADNI (Alzheimer's Disease NeuroImaging Initiative)
Alzheimer's Disease Neuroimaging Initiative (ADNI) is a multisite study that aims to improve clinical trials for the prevention and treatment of Alzheimer’s disease (AD).[1] This cooperative study combines expertise and funding from the…
28 papers · 5 benchmarks
Street Scene is a dataset for video anomaly detection.
26 papers · 3 benchmarks
Real 3D-AD is the first point cloud anomaly detection dataset for industrial products.
23 papers · 2 benchmarks
Encourages machine learning research in this area and to help facilitate further work in understanding and mitigating the effects of climate change.
21 papers · 0 benchmarks
LAG (Large-scale Attention based Glaucoma)
Includes 5,824 fundus images labeled with either positive glaucoma (2,392) or negative glaucoma (3,432).
21 papers · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
17 papers · 3 benchmarks
NINCO (No ImageNet Class Objects)
The NINCO (No ImageNet Class Objects) dataset is introduced in the ICML 2023 paper In or Out?
17 papers · 0 benchmarks
RoadAnomaly21 is a dataset for anomaly segmentation, the task of identify the image regions containing objects that have never been seen during training.
15 papers · 0 benchmarks
ToyADMOS dataset is a machine operating sounds dataset of approximately 540 hours of normal machine operating sounds and over 12,000 samples of anomalous sounds collected with four microphones at a 48kHz sampling rate, prepared by Yuma…
15 papers · 0 benchmarks
The UCR Anomaly Archive is a collection of 250 uni-variate time series collected in human medicine, biology, meteorology and industry.
14 papers · 2 benchmarks
UPFD (User Preference-aware Fake News Detection)
For benchmarking, please refer to its variant UPFD-POL and UPFD-GOS.
13 papers · 0 benchmarks
Yelp-Fraud (Multi-relational Graph Dataset for Yelp Spam Review Detection)
Yelp-Fraud is a multi-relational graph dataset built upon the Yelp spam review dataset, which can be used in evaluating graph-based node classification, fraud detection, and anomaly detection models.
13 papers · 3 benchmarks
HyperKvasir dataset contains 110,079 images and 374 videos where it captures anatomical landmarks and pathological and normal findings.
12 papers · 2 benchmarks
AeBAD (Aero-engine Blade Anomaly Detection Dataset)
Unlike previous datasets that focus on detecting the diversity of defect categories (like MVTec AD and VisA), AeBAD is centered on the diversity of domains within the same data category.
11 papers · 3 benchmarks
CATS (Color and Thermal Stereo Benchmark)
A dataset consisting of stereo thermal, stereo color, and cross-modality image pairs with high accuracy ground truth (< 2mm) generated from a LiDAR.
11 papers · 2 benchmarks
Darpa is a dataset consisting of communications between source IPs and destination IPs.
11 papers · 0 benchmarks
Predicting forest cover type from cartographic variables only (no remotely sensed data).
11 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.