Browse State-of-the-Art › Semantic correspondence
Semantic correspondence
88 papers with code · 6 benchmarks · 8 datasets archive 2025-07-28
The task of semantic correspondence aims to establish reliable visual correspondence between different instances of the same object category.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
6 leaderboard tables shown for this task, 6 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| SPair-71k (22 rows) | GeoAware-SC (Supervised, AP-10K P.T.) | Telling Left from Right: Identifying Geometry-Aware Semantic Correspondence | code | Syntology ran 11 of 15 samples · 4 unverified | Compare |
| PF-PASCAL (15 rows) | DINOv2 | Semantic Correspondence: Unified Benchmarking and a Strong Baseline | code | — | Compare |
| PF-WILLOW (8 rows) | LDMCorrespondences | Unsupervised Semantic Correspondence Using Stable Diffusion | code | Syntology ran 4 of 12 samples · 8 unverified | Compare |
| Caltech-101 (2 rows) | HPF | Hyperpixel Flow: Semantic Correspondence with Multi-layer Neural Features | code | Syntology ran 0 of 14 samples · 14 unverified | Compare |
| AP-10K (1 row) | DINOv2 | Semantic Correspondence: Unified Benchmarking and a Strong Baseline | code | — | Compare |
| CUB-200-2011 (1 row) | LDM Correspondences | Unsupervised Semantic Correspondence Using Stable Diffusion | code | Syntology ran 4 of 12 samples · 8 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
8 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 88 papers with code (175 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
15 Jun 2021 3 repositories listed Syntology ran 2 of 9 samples · 7 unverifiedIn this paper, we present a fast exemplar-based image colorization approach using color embeddings named Color2Embed.
-
13 May 2021 3 repositories listedWe introduce DiscoBox, a novel framework that jointly learns instance segmentation and semantic correspondence using bounding box supervision.
-
27 Dec 2020 3 repositories listedFor these reasons, we address the computer-assisted search for prior art by creating a training dataset for supervised machine learning called PatentMatch.
-
24 Oct 2018 3 repositories listed Syntology ran 5 of 21 samples · 16 unverifiedSecond, we demonstrate that the model can be trained effectively from weak supervision in the form of matching and non-matching image pairs without the need for costly manual annotation of point to point correspondences.
-
24 May 2024 2 repositories listedBy decoupling the intricate referring semantics into different granularity with a visual-linguistic hierarchy, and dynamic aggregating it with intra- and inter-selection, CoHD boosts multi-granularity comprehension with…
-
6 Jun 2023 2 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedWe propose a simple strategy to extract this implicit knowledge out of diffusion networks as image features, namely DIffusion FeaTures (DIFT), and use them to establish correspondences between real images.
-
18 May 2023 2 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 5 pointer-only (licence)In this paper, we propose a detector with the ability to predict both open-vocabulary objects and their part segmentation.
-
22 Dec 2021 2 repositories listedWe introduce a novel cost aggregation network, dubbed Volumetric Aggregation with Transformers (VAT), to tackle the few-shot segmentation task by using both convolutions and transformers to efficiently handle high…
-
7 Dec 2021 2 repositories listedMakeup transfer is not only to extract the makeup style of the reference image, but also to render the makeup style to the semantic corresponding position of the target image.
-
9 Dec 2020 2 repositories listedAs the main discriminative information of a fine-grained image usually resides in subtle regions, methods along this line are prone to heavy label noise in fine-grained recognition.
-
3 Aug 2018 2 repositories listedSuch image comparison based approach also alleviates the problem of data scarcity and hence enhances scalability of the proposed approach for novel object categories with minimal annotation.
-
19 Dec 2017 2 repositories listedWe tackle the task of semantic alignment where the goal is to compute dense semantic correspondence aligning two images depicting objects of the same category.
-
29 May 2025 1 repository listedHowever, edits requiring significant structural changes, such as non-rigid deformations, object modifications, or content generation, remain challenging.
-
23 May 2025 1 repository listedWe hope this survey serves as a comprehensive reference and consolidated baseline for future development.
-
7 Apr 2025 1 repository listed Syntology ran 1 of 2 samples · 1 unverified · 1 pointer-only (licence)To filter unnecessary similarity interactions and decrease trainable parameters in the Interactive Similarity Aggregation (ISA) module, we design a Similarity Reorganization (SR) module to identify attentive…
-
10 Feb 2025 1 repository listedThe evaluation of image captions, looking at both linguistic fluency and semantic correspondence to visual contents, has witnessed a significant effort.
-
Common3D: Self-Supervised Learning of 3D Morphable Models for Common Objects in Neural Feature Space1 Jan 2025 1 repository listedGiven a single test image, 3DMMs can be used to solve various tasks, such as predicting the 3D shape, pose, semantic correspondence, and instance segmentation of an object.
-
4 Dec 2024 1 repository listed Syntology ran 0 of 5 samples · 5 unverified · 5 pointer-only (licence)Internal features from large-scale pre-trained diffusion models have recently been established as powerful semantic descriptors for a wide range of downstream tasks.
-
4 Dec 2024 1 repository listedIn this paper, we argue that measure at such a level may not be effective enough to generalize from base to novel classes when using only a few images.
-
3 Oct 2024 1 repository listed Syntology ran 4 of 4 samples · 0 unverified · 4 pointer-only (licence)The Diffusion Model has not only garnered noteworthy achievements in the realm of image generation but has also demonstrated its potential as an effective pretraining method utilizing unlabeled data.
-
8 Jul 2024 1 repository listedOn the other hand, we propose semantic-geometric consistency for outlier removal, which makes full use of semantic information and significantly improves the quality of correspondences.
-
11 Jun 2024 1 repository listed Syntology ran 2 of 13 samples · 11 unverifiedImage editing serves as a practical yet challenging task considering the diverse demands from users, where one of the hardest parts is to precisely describe how the edited image should look like.
-
27 May 2024 1 repository listed Syntology ran 3 of 3 samples · 0 unverifiedWe introduce the first zero-shot approach for Video Semantic Segmentation (VSS) based on pre-trained diffusion models.
-
16 May 2024 1 repository listedIn this work, we present Semantic Gesticulator, a novel framework designed to synthesize realistic gestures accompanying speech with strong semantic correspondence.
-
15 May 2024 1 repository listedIn Stage 1, we introduce factuality-guided contrastive learning for visual representation by maximizing the semantic correspondence between radiographs and corresponding factual descriptions.
-
1 Jan 2024 1 repository listedEstablishing precise semantic correspondence across object instances in different images is a fundamental and challenging task in computer vision.
-
13 Dec 2023 1 repository listedRecently, patch-wise contrastive learning is drawing attention for the image translation by exploring the semantic correspondence between the input and output images.
-
4 Dec 2023 1 repository listed Syntology ran 8 of 11 samples · 3 unverified · 11 pointer-only (licence)Given a clothing image and a person image, an image-based virtual try-on aims to generate a customized image that appears natural and accurately reflects the characteristics of the clothing image.
-
30 Nov 2023 1 repository listedThis paper builds on the hypothesis that there is an inherent data-hungry matter in learning semantic correspondences and uncovers the models can be more trained by employing densified training pairs.
-
28 Nov 2023 1 repository listed Syntology ran 11 of 15 samples · 4 unverified · 15 pointer-only (licence)This paper identifies the importance of being geometry-aware for semantic correspondence and reveals a limitation of the features of current foundation models under simple post-processing.
Syntology lines on 11 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections