Browse State-of-the-Art › Visual Place Recognition
Visual Place Recognition
141 papers with code · 40 benchmarks · 28 datasets archive 2025-07-28
Visual Place Recognition is the task of matching a view of a place with a different view of the same place taken at a different time.
Source: Visual place recognition using landmark distribution descriptors
Image credit: Visual place recognition using landmark distribution descriptors
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
40 leaderboard tables shown for this task, 40 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 40 until expanded.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
28 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
3 subtasks in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 141 papers with code (297 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
29 Apr 2021 32 repositories listed Syntology ran 5 of 20 samples · 15 unverified · 2 pointer-only (licence)In this paper, we question if self-supervised learning provides new properties to Vision Transformer (ViT) that stand out compared to convolutional networks (convnets).
-
14 Apr 2023 26 repositories listed Syntology ran 21 of 46 samples · 25 unverified · 12 pointer-only (licence)The recent breakthroughs in natural language processing for model pretraining on large quantities of data have opened the way for similar foundation models in computer vision.
-
26 Nov 2019 19 repositories listed Syntology ran 6 of 22 samples · 16 unverified · 6 pointer-only (licence)This paper introduces SuperGlue, a neural network that matches two sets of local features by jointly finding correspondences and rejecting non-matchable points.
-
23 Nov 2015 14 repositories listed Syntology ran 0 of 12 samples · 12 unverifiedWe tackle the problem of large scale visual place recognition, where the task is to quickly and accurately recognize the location of a given query photograph.
-
10 Apr 2018 6 repositories listedThis is largely due to the difficulty in extracting local feature descriptors from a point cloud that can subsequently be encoded into a global descriptor for the retrieval task.
-
24 Feb 2025 4 repositories listed Syntology ran 1 of 6 samples · 5 unverifiedWe find that MegaLoc (1) achieves state of the art on a large number of Visual Place Recognition datasets, (2) impressive results on common Landmark Retrieval datasets, and (3) sets a new state of the art for Visual…
-
21 Aug 2023 4 repositories listed Syntology ran 11 of 13 samples · 2 unverified · 1 pointer-only (licence)Visual Place Recognition is a task that aims to predict the place of an image (called query) based solely on its visual features.
-
2 Mar 2021 4 repositories listed Syntology ran 4 of 5 samples · 1 unverifiedVisual Place Recognition is a challenging task for robotics and autonomous systems, which must deal with the twin problems of appearance and viewpoint change in an always changing world.
-
27 Sep 2023 3 repositories listed Syntology ran 5 of 13 samples · 8 unverifiedWorldwide Geo-localization aims to pinpoint the precise location of images taken anywhere on Earth.
-
9 Dec 2018 3 repositories listed Syntology ran 1 of 9 samples · 8 unverified · 2 pointer-only (licence)In this paper we propose HF-Net, a hierarchical localization approach based on a monolithic CNN that simultaneously predicts local features and global descriptors for accurate 6-DoF localization.
-
20 Nov 2018 3 repositories listedOur approaches rely on local features with an encoding technique to represent an image as a single vector.
-
13 Mar 2024 2 repositories listedFeature point detection and description is the backbone for various computer vision applications, such as Structure-from-Motion, visual SLAM, and visual place recognition.
-
5 Jun 2023 2 repositories listed Syntology ran 4 of 6 samples · 2 unverifiedThis paper proposes a novel thermal geo-localization framework using satellite RGB imagery, which includes multiple domain adaptation methods to address the limited availability of paired thermal and satellite images.
-
4 Apr 2023 2 repositories listed Syntology ran 20 of 29 samples · 9 unverified · 15 pointer-only (licence)Line segments are powerful features complementary to points.
-
5 Apr 2022 2 repositories listed Syntology ran 6 of 10 samples · 4 unverified · 2 pointer-only (licence)Visual Geo-localization (VG) is the task of estimating the position where a given photo was taken by comparing it with a large database of images of known locations.
-
28 Mar 2022 2 repositories listed Syntology ran 9 of 16 samples · 7 unverified · 16 pointer-only (licence)Natural language-based communication with mobile devices and home appliances is becoming increasingly popular and has the potential to become natural for communicating with mobile robots in the future.
-
14 Oct 2020 2 repositories listed Syntology ran 1 of 8 samples · 7 unverifiedWe address the task of cross-domain visual place recognition, where the goal is to geolocalize a given query image against a labeled gallery, in the case where the query and the gallery belong to different visual…
-
11 Dec 2018 2 repositories listedPoint cloud based place recognition is still an open issue due to the difficulty in extracting local features from the raw 3D point cloud and generating the global descriptor, and it's even harder in the large-scale…
-
16 Jun 2018 2 repositories listedThis document describes G2D, a software that enables capturing videos from Grand Theft Auto V (GTA V), a popular role playing game set in an expansive virtual city.
-
7 Nov 2016 2 repositories listedThe paper presents an approach to indoor personal localization on a mobile device based on visual place recognition.
-
19 Jun 2025 1 repository listedStand-alone Visual Place Recognition (VPR) systems have little defence against a well-designed adversarial attack, which can lead to disastrous consequences when deployed for robot navigation.
-
20 May 2025 1 repository listedThe unified framework of leading-edge place recognition methods, i.
-
29 Apr 2025 1 repository listedEach pipeline is tailored to process high-resolution ISS imagery, identifying both natural and man-made geographical features.
-
15 Apr 2025 1 repository listedIn many applications this information is already present or can be acquired with low effort.
-
14 Apr 2025 1 repository listed Syntology ran 3 of 3 samples · 0 unverifiedIn this paper, we propose the Focus on Local (FoL) approach to stimulate the performance of image retrieval and re-ranking in VPR simultaneously by mining and exploiting reliable discriminative local regions in images…
-
8 Apr 2025 1 repository listedVisual Place Recognition (VPR) is a critical task in computer vision, traditionally enhanced by re-ranking retrieval results with image matching.
-
27 Mar 2025 1 repository listedThis paper introduces a novel training paradigm to improve the performance of existing VPR networks by enhancing multi-view diversity within current datasets through uncertainty estimation and NeRF-based data…
-
9 Mar 2025 1 repository listedVisual Place Recognition (VPR) is a crucial capability for long-term autonomous robots, enabling them to identify previously visited locations using visual information.
-
9 Mar 2025 1 repository listedMost deep learning-based methods in an end-to-end manner cannot extract global features with sufficient semantic information from RGB images.
-
Image-Based Relocalization and Alignment for Long-Term Monitoring of Dynamic Underwater Environments6 Mar 2025 1 repository listedEffective monitoring of underwater ecosystems is crucial for tracking environmental changes, guiding conservation efforts, and ensuring long-term ecosystem health.
Syntology lines on 15 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections