Home › Datasets › task › Visual Navigation
Visual Navigation datasets
archive 2025-07-28
19 datasets carry the task tag "Visual Navigation" (the task itself: Visual Navigation), ordered by the archive's paper count. Page 1 of 1: 19 shown of 19. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Visual Navigation datasets 1–19 of 19
The Replica Dataset is a dataset of high quality reconstructions of a variety of indoor spaces.
414 papers · 4 benchmarks
AI2-Thor is an interactive environment for embodied AI.
243 papers · 1 benchmark
SUNCG is a large-scale dataset of synthetic 3D scenes with dense volumetric annotations.
186 papers · 0 benchmarks
R2R is a dataset for visually-grounded natural language navigation in real buildings.
174 papers · 2 benchmarks
The 2D-3D-S dataset provides a variety of mutually registered modalities from 2D, 2.5D and 3D domains, with instance-level semantic and geometric annotations.
147 papers · 6 benchmarks
The HELP dataset is an automatically created natural language inference (NLI) dataset that embodies the combination of lexical and logical inferences focusing on monotonicity (i.e., phrase replacement-based reasoning).
30 papers · 1 benchmark
AVD (Active Vision Dataset)
AVD focuses on simulating robotic vision tasks in everyday indoor environments using real imagery.
29 papers · 1 benchmark
MINOS is a simulator designed to support the development of multisensory models for goal-directed navigation in complex indoor environments.
22 papers · 0 benchmarks
The Habitat-Matterport 3D Semantics Dataset (HM3DSem) is the largest-ever dataset of 3D real-world and indoor spaces with densely annotated semantics that is available to the academic community.
13 papers · 0 benchmarks
A rich, extensible and efficient environment that contains 45,622 human-designed 3D scenes of visually realistic houses, ranging from single-room studios to multi-storied houses, equipped with a diverse set of fully labeled 3D objects,…
11 papers · 0 benchmarks
Talk The Walk is a large-scale dialogue dataset grounded in action and perception.
11 papers · 0 benchmarks
IQUAD (Interactive Question Answering Dataset)
IQUAD is a dataset for Visual Question Answering in interactive environments.
7 papers · 0 benchmarks
MineRLis an imitation learning dataset with over 60 million frames of recorded human player data.
3 papers · 0 benchmarks
BASEPROD (The Bardenas Semi-Desert Planetary Rover Dataset)
BASEPROD provides comprehensive rover sensor data collected over a 1.7 km traverse, accompanied by high-resolution 2D and 3D drone maps of the terrain.
2 papers · 0 benchmarks
MIDGARD is an open-source simulator for autonomous robot navigation in outdoor unstructured environments.
2 papers · 0 benchmarks
Talk2Nav is a large-scale dataset with verbal navigation instructions.
2 papers · 0 benchmarks
This is the supporting dataset for the ECCV 2024 paper "MARs: Multi-view Attention Regularizations for Patch-based Feature Recognition of Space Terrain".
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
A dataset for Image-Goal Navigation in Habitat based on Gibson scenes.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.