Home › Datasets › task › Autonomous Driving

Autonomous Driving datasets

archive 2025-07-28

70 datasets carry the task tag "Autonomous Driving" (the task itself: Autonomous Driving), ordered by the archive's paper count. Page 1 of 2: 48 shown of 70. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Autonomous Driving datasets 1–48 of 70

CARLA (Car Learning to Act)
CARLA (CAR Learning to Act) is an open simulator for urban driving, developed as an open-source layer over Unreal Engine 4.
1,345 papers · 4 benchmarks
The Waymo Open Dataset is comprised of high resolution sensor data collected by autonomous vehicles operated by the Waymo Driver in a wide variety of conditions.
481 papers · 16 benchmarks
AirSim is a simulator for drones, cars and more, built on Unreal Engine.
285 papers · 0 benchmarks
Virtual KITTI is a photo-realistic synthetic video dataset designed to learn and evaluate computer vision models for several video understanding tasks: object detection and multi-object tracking, scene-level and instance-level semantic…
133 papers · 0 benchmarks
IDD (Indian Driving Dataset)
IDD is a dataset for road scene understanding in unstructured environments used for semantic segmentation and object detection for autonomous driving.
98 papers · 1 benchmark
TORCS (The Open Racing Car Simulator)
TORCS (The Open Racing Car Simulator) is a driving simulator.
96 papers · 0 benchmarks
The INTERACTION dataset contains naturalistic motions of various traffic participants in a variety of highly interactive driving scenarios from different countries.
81 papers · 1 benchmark
ApolloScape is a large dataset consisting of over 140,000 video frames (73 street scene videos) from various locations in China under varying weather conditions.
74 papers · 4 benchmarks
The SemanticPOSS dataset for 3D semantic segmentation contains 2988 various and complicated LiDAR scans with large quantity of dynamic instances.
71 papers · 1 benchmark
Lost and Found is a novel lost-cargo image sequence dataset comprising more than two thousand frames with pixelwise annotations of obstacle and free-space and provide a thorough comparison to several stereo-based baseline methods.
57 papers · 1 benchmark
Fisheye cameras are commonly employed for obtaining a large field of view in surveillance, augmented reality and in particular automotive applications.
55 papers · 1 benchmark
PandaSet is a dataset produced by a complete, high-precision autonomous vehicle sensor kit with a no-cost commercial license.
54 papers · 0 benchmarks
UAVid is a high-resolution UAV semantic segmentation dataset as a complement, which brings new challenges, including large scale variation, moving object recognition and temporal consistency preservation.
54 papers · 2 benchmarks
DrivingStereo contains over 180k images covering a diverse set of driving scenarios, which is hundreds of times larger than the KITTI Stereo dataset.
50 papers · 0 benchmarks
BDD-X (Berkeley Deep Drive-X (eXplanation))
Berkeley Deep Drive-X (eXplanation) is a dataset is composed of over 77 hours of driving within 6,970 videos.
47 papers · 0 benchmarks
WildDash is a benchmark evaluation method is presented that uses the meta-information to calculate the robustness of a given algorithm with respect to the individual hazards.
47 papers · 2 benchmarks
The Talk2Car dataset finds itself at the intersection of various research domains, promoting the development of cross-disciplinary solutions for improving the state-of-the-art in grounding natural language into visual space.
45 papers · 0 benchmarks
DDD17 (DAVIS Driving Dataset 2017)
DDD17 has over 12 h of a 346x260 pixel DAVIS sensor recording highway and city driving in daytime, evening, night, dry and wet weather conditions, along with vehicle speed, GPS position, driver steering, throttle, and brake captured from…
41 papers · 1 benchmark
KITTI Road is road and lane estimation benchmark that consists of 289 training and 290 test images.
41 papers · 0 benchmarks
The Argoverse 2 Motion Forecasting Dataset is a curated collection of 250,000 scenarios for training and validation.
39 papers · 0 benchmarks
H3D (Honda Research Institute 3D)
The H3D is a large scale full-surround 3D multi-object detection and tracking dataset.
39 papers · 0 benchmarks
HDD (Honda Research Institute Driving Dataset)
Honda Research Institute Driving Dataset (HDD) is a dataset to enable research on learning driver behavior in real-life environments.
38 papers · 0 benchmarks
The A3D dataset is a step forward to make autonomous driving safer for pedestrians and the public in the real world.
37 papers · 0 benchmarks
DR(eye)VE is a large dataset of driving scenes for which eye-tracking annotations are available.
34 papers · 0 benchmarks
ROAD (ROAD: The ROad event Awareness Dataset for Autonomous Driving)
ROAD is designed to test an autonomous vehicle's ability to detect road events, defined as triplets composed by an active agent, the action(s) it performs and the corresponding scene locations.
27 papers · 0 benchmarks
Toronto-3D is a large-scale urban outdoor point cloud dataset acquired by an MLS system in Toronto, Canada for semantic segmentation.
24 papers · 2 benchmarks
CARRADA is a dataset of synchronized camera and radar recordings with range-angle-Doppler annotations.
22 papers · 0 benchmarks
ApolloCar3DT is a dataset that contains 5,277 driving images and over 60K car instances, where each car is fitted with an industry-grade 3D CAD model with absolute model size and semantically labelled keypoints.
17 papers · 14 benchmarks
4Seasons is adataset covering seasonal and challenging perceptual conditions for autonomous driving.
15 papers · 0 benchmarks
TITAN consists of 700 labeled video-clips (with odometry) captured from a moving vehicle on highly interactive urban traffic scenes in Tokyo.
15 papers · 0 benchmarks
The KITTI-Depth dataset includes depth maps from projected LiDAR point clouds that were matched against the depth estimation from the stereo cameras.
14 papers · 0 benchmarks
UrbanLoco is a mapping/localization dataset collected in highly-urbanized environments with a full sensor-suite.
14 papers · 0 benchmarks
SynWoodScape (Synthetic Surround-view Fisheye Camera Dataset for Autonomous Driving)
SynWoodScape is a synthetic version of the surround-view dataset covering many of its weaknesses and extending it.
13 papers · 0 benchmarks
ONCE-3DLanes (Monocular 3D Lane Detection Dataset)
ONCE-3DLanes is a real-world autonomous driving dataset with lane layout annotation in 3D space.
12 papers · 0 benchmarks
SODA10M is a large-scale object detection benchmark for standardizing the evaluation of different self-supervised and semi-supervised approaches by learning from raw data.
12 papers · 0 benchmarks
DAWN emphasizes a diverse traffic environment (urban, highway and freeway) as well as a rich variety of traffic flow.
10 papers · 0 benchmarks
Annotated using images taken by a drone in 501 separate flights, totalling in over 62 hours of trajectory data.
10 papers · 0 benchmarks
BLVD is a large scale 5D semantics dataset collected by the Visual Cognitive Computing and Intelligent Vehicles Lab.
9 papers · 0 benchmarks
SynthCity is a 367.9M point synthetic full colour Mobile Laser Scanning point cloud.
8 papers · 0 benchmarks
DurLAR (A High-Fidelity 128-Channel LiDAR Dataset with Panoramic Ambient and Reflectivity Imagery)
DurLAR is a high-fidelity 128-channel 3D LiDAR dataset with panoramic ambient (near infrared) and reflectivity imagery for multi-modal autonomous driving applications.
5 papers · 0 benchmarks
ELAS is a dataset for lane detection.
5 papers · 0 benchmarks
PedX is a large-scale multi-modal collection of pedestrians at complex urban intersections.
5 papers · 0 benchmarks
SODA-D is a large-scale dataset tailored for small object detection in driving scenario, which is built on top of MVD dataset and owned data, where the former is a dataset dedicated to pixel-level understanding of street scenes, and the…
5 papers · 1 benchmark
This self-driving dataset collected in Brno, Czech Republic contains data from four WUXGA cameras, two 3D LiDARs, inertial measurement unit, infrared camera and especially differential RTK GNSS receiver with centimetre accuracy.
4 papers · 0 benchmarks
Unsupervised Domain Adaptation demonstrates great potential to mitigate domain shifts by transferring models from labeled source domains to unlabeled target domains.
4 papers · 3 benchmarks
MUAD (Multiple Uncertainties for Autonomous Driving)
The MUAD dataset (Multiple Uncertainties for Autonomous Driving), consisting of 10,413 realistic synthetic images with diverse adverse weather conditions (night, fog, rain, snow), out-of-distribution objects, and annotations for semantic…
4 papers · 0 benchmarks
Swiss3DCities is a dataset that is manually annotated for semantic segmentation with per-point labels, and is built using photogrammetry from images acquired by multirotors equipped with high-resolution cameras.
4 papers · 0 benchmarks
TCG (Traffic Control Gesture)
The TCG dataset is used to evaluate Traffic Control Gesture recognition for autonomous driving.
4 papers · 1 benchmark

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.