Home › Datasets › task › Metric Learning
Metric Learning datasets
archive 2025-07-28
33 datasets carry the task tag "Metric Learning" (the task itself: Metric Learning), ordered by the archive's paper count. Page 1 of 1: 33 shown of 33. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Metric Learning datasets 1–33 of 33
The Caltech-UCSD Birds-200-2011 (CUB-200-2011) dataset is the most widely-used dataset for fine-grained visual categorization task.
2,235 papers · 47 benchmarks
The CASIA-WebFace dataset is used for face verification and face identification tasks.
415 papers · 2 benchmarks
Stanford Online Products (SOP) dataset has 22,634 classes with 120,053 product images.
231 papers · 5 benchmarks
In-Shop (In-shop Clothes Retrieval Benchmark)
In-shop Clothes Retrieval Benchmark evaluates the performance of in-shop Clothes Retrieval.
154 papers · 2 benchmarks
VIPeR (Viewpoint Invariant Pedestrian Recognition)
The Viewpoint Invariant Pedestrian Recognition (VIPeR) dataset includes 632 people and two outdoor cameras under different viewpoints and light conditions.
134 papers · 0 benchmarks
Structured3D is a large-scale photo-realistic dataset containing 3.5K house designs (a) created by professional designers with a variety of ground truth 3D structure annotations (b) and generate photo-realistic 2D images (c).
85 papers · 7 benchmarks
ABO (Amazon Berkeley Objects)
ABO is a large-scale dataset designed for material prediction and multi-view retrieval experiments.
82 papers · 0 benchmarks
WildDeepfake is a dataset for real-world deepfakes detection which consists of 7,314 face sequences extracted from 707 deepfake videos that are collected completely from the internet.
49 papers · 0 benchmarks
CARS196 is composed of 16,185 car images of 196 classes.
43 papers · 4 benchmarks
Veri-Wild is the largest vehicle re-identification dataset (as of CVPR 2019).
40 papers · 3 benchmarks
This is the second version of the Google Landmarks dataset (GLDv2), which contains images annotated with labels representing human-made and natural landmarks.
35 papers · 4 benchmarks
Consists of 20k English biomedical entity mentions from Reddit expert-annotated with links to SNOMED CT, a widely-used medical knowledge graph.
21 papers · 0 benchmarks
Provides detailed, graph-based annotations of social situations depicted in movie clips.
13 papers · 0 benchmarks
The Airport dataset is a dataset for person re-identification which consists of 39,902 images and 9,651 identities across six cameras.
8 papers · 0 benchmarks
The Hotels-50K dataset consists of over 1 million images from 50,000 different hotels around the world.
7 papers · 0 benchmarks
Collected from top 10 most popular clothing/wearable brandname logos captured in rich visual context.
6 papers · 1 benchmark
N-Digit MNIST is a multi-digit MNIST-like dataset.
5 papers · 0 benchmarks
The BirdVox-full-night dataset contains 6 audio recordings, each about ten hours in duration.
3 papers · 0 benchmarks
Goldfinch is a dataset for fine-grained recognition challenges.
3 papers · 0 benchmarks
Large Age-Gap (LAG) is a dataset for face verification, The dataset contains 3,828 images of 1,010 celebrities.
3 papers · 0 benchmarks
The OpeReid dataset is a person re-identification dataset that consists of 7,413 images of 200 persons.
3 papers · 0 benchmarks
DyML-Animal is based on animal images selected from ImageNet-5K [1].
2 papers · 1 benchmark
DyML-Product is derived from iMaterialist-2019, a hierarchical online product dataset.
2 papers · 1 benchmark
DyML-Vehicle merges two vehicle re-ID datasets PKU VehicleID [1], VERI-Wild [1].
2 papers · 1 benchmark
The Freiburg Spatial Relations dataset features 546 scenes each containing two out of 25 household objects.
2 papers · 0 benchmarks
IMEMNET (Image-MusicEmotion-Matching-Net)
The Image-MusicEmotion-Matching-Net (IMEMNet) dataset is a dataset for continuous emotion-based image and music matching.
2 papers · 0 benchmarks
The Live Comment Dataset is a large-scale dataset with 2,361 videos and 895,929 live comments that were written while the videos were streamed.
2 papers · 0 benchmarks
VideoForensicsHQ is a benchmark dataset for face video forgery detection, providing high quality visual manipulations.
2 papers · 0 benchmarks
A View From Somewhere (AVFS)—a dataset of 638,180 face similarity judgments over 4,921 faces.
1 paper · 0 benchmarks
HAM (Human-annotated Mappings)
HAM is a dataset for molecular graph partitioning.
1 paper · 0 benchmarks
This is the supporting dataset for the ECCV 2024 paper "MARs: Multi-view Attention Regularizations for Patch-based Feature Recognition of Space Terrain".
1 paper · 0 benchmarks
S3O4D (Stanford 3D Objects for Disentangling)
The data consists of 100,000 renderings each of the Bunny and Dragon objects from the Stanford 3D Scanning Repository.
1 paper · 0 benchmarks
Tsinghua Dogs is a fine-grained classification dataset for dogs, over 65% of whose images are collected from people's real life.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.