Home › Datasets › task › Unsupervised Domain Adaptation

Unsupervised Domain Adaptation datasets

archive 2025-07-28

34 datasets carry the task tag "Unsupervised Domain Adaptation" (the task itself: Unsupervised Domain Adaptation), ordered by the archive's paper count. Page 1 of 1: 34 shown of 34. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Unsupervised Domain Adaptation datasets 1–34 of 34

Office-Home is a benchmark dataset for domain adaptation which contains 4 domains where each domain consists of 65 categories.
1,074 papers · 11 benchmarks
Market-1501 is a large-scale public benchmark dataset for person re-identification.
873 papers · 9 benchmarks
PACS (Photo-Art-Cartoon-Sketch)
PACS is an image dataset for domain generalization.
668 papers · 10 benchmarks
Office-31 (Office Dataset)
The Office dataset contains 31 object categories in three domains: Amazon, DSLR and Webcam.
643 papers · 7 benchmarks
ImageNet-C is an open source data set that consists of algorithmically generated corruptions (blur, noise) applied to the ImageNet test-set.
602 papers · 4 benchmarks
SYNTHIA (SYNTHetic Collection of Imagery and Annotations)
The SYNTHIA dataset is a synthetic dataset that consists of 9400 multi-viewpoint photo-realistic frames rendered from a virtual city and comes with pixel-level semantic annotations for 13 classes.
538 papers · 10 benchmarks
ImageNet-R (ImageNet-Rendition)
ImageNet-R(endition) contains art, cartoons, deviantart, graffiti, embroidery, graphics, origami, paintings, patterns, plastic objects, plush objects, sculptures, sketches, tattoos, toys, and video game renditions of ImageNet classes.
481 papers · 5 benchmarks
The ImageNet-A dataset consists of real-world, unmodified, and naturally occurring examples that are misclassified by ResNet models.
431 papers · 5 benchmarks
GTA5 (Grand Theft Auto 5)
The GTA5 dataset contains 24966 synthetic images with pixel level semantic annotation.
412 papers · 7 benchmarks
Foggy Cityscapes is a synthetic foggy dataset which simulates fog on real scenes.
249 papers · 7 benchmarks
VisDA-2017 is a simulation-to-real dataset for domain adaptation with over 280,000 images across 12 categories in the training, validation and testing domains.
223 papers · 6 benchmarks
This paper introduces the pipeline to scale the largest dataset in egocentric vision EPIC-KITCHENS.
162 papers · 6 benchmarks
VehicleID (PKU VehicleID)
The “VehicleID” dataset contains CARS captured during the daytime by multiple real-world surveillance cameras distributed in a small city in China.
134 papers · 8 benchmarks
SIM10k is a synthetic dataset containing 10,000 images, which is rendered from the video game Grand Theft Auto V (GTA5).
92 papers · 3 benchmarks
VeRi-776 is a vehicle re-identification dataset which contains 49,357 images of 776 vehicles from 20 cameras.
79 papers · 1 benchmark
The Oxford RobotCar Dataset contains over 100 repetitions of a consistent route through Oxford, UK, captured over a period of over a year.
42 papers · 3 benchmarks
SynLiDAR is a large-scale synthetic LiDAR sequential point cloud dataset with point-wise annotations.
20 papers · 1 benchmark
Jester Gesture Recognition dataset includes 148,092 labeled video clips of humans performing basic, pre-defined hand gestures in front of a laptop camera or webcam.
16 papers · 6 benchmarks
Adaptiope is a domain adaptation dataset with 123 classes in the three domains synthetic, product and real life.
9 papers · 0 benchmarks
The Cross-dataset Testbed is a Decaf7 based cross-dataset image classification dataset, which contains 40 categories of images from 3 domains: 3,847 images in Caltech256, 4,000 images in ImageNet, and 2,626 images for SUN.
7 papers · 0 benchmarks
MegaAge is a large dataset that consists of 41,941 faces annotated with age posterior distributions.
7 papers · 0 benchmarks
The ClonedPerson dataset is a large-scale synthetic person re-identification dataset introduced in the paper "Cloning Outfits from Real-World Images to 3D Characters for Generalizable Person Re-Identification" in CVPR 2022.
6 papers · 4 benchmarks
Modern Office-31 is a refurbished version of the commonly used Office-31 dataset.
6 papers · 0 benchmarks
OOD-CV (Out Of Distribution Generalization in Computer Vision)
Enhancing the robustness of vision algorithms in real-world scenarios is challenging.
6 papers · 1 benchmark
Libri-Adapt aims to support unsupervised domain adaptation research on speech recognition models.
5 papers · 0 benchmarks
This dataset contains 114 individuals including 1824 images captured from two disjoint camera views.
5 papers · 1 benchmark
RoCoG-v2 (Robot Control Gestures)
RoCoG-v2 (Robot Control Gestures) is a dataset intended to support the study of synthetic-to-real and ground-to-air video domain adaptation.
5 papers · 1 benchmark
The Sims4Action Dataset: a videogame-based dataset for Synthetic→Real domain adaptation for human activity recognition.
5 papers · 0 benchmarks
The Five-Billion-Pixels dataset contains more than 5 billion labeled pixels of 150 high-resolution Gaofen-2 (4 m) satellite images, annotated in a 24-category system covering artificial-constructed, agricultural, and natural classes.
3 papers · 0 benchmarks
MSDA (Multi-source domain adaptation dataset for text recognition)
5 domains: synthetic domain, document domain, street view domain, handwritten domain, and car license domain over five million images
2 papers · 2 benchmarks
CFC-DAOD (Caltech Fish Counting – Domain Adaptive Object Detection)
CFC-DAOD is a domain adaptation extension to the Caltech Fish Counting domain generalization benchmark.
1 paper · 1 benchmark
UDA-CH (Unsupervised Domain Adaptation on Cultural Heritage)
UDA-CH contains 16 objects that cover a variety of artworks which can be found in a museum like sculptures, paintings and books.
1 paper · 1 benchmark
A cross-city UDA benchmark built upon nuScenes.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.