Browse State-of-the-Art › Scene Parsing
Scene Parsing
80 papers with code · 2 benchmarks · 4 datasets archive 2025-07-28
Scene parsing is to segment and parse an image into different image regions associated with semantic categories, such as sky, road, person, and bed. MIT Description
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
2 leaderboard tables shown for this task, 2 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| PGDP5K (2 rows) | PGDPNet | Plane Geometry Diagram Parsing | code | Syntology ran 1 of 3 samples · 2 unverified | Compare |
| Cityscapes test (1 row) | VCD No Coarse | Variational Context-Deformable ConvNets for Indoor Scene Parsing | — | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
4 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
9 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 80 papers with code (199 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
4 Dec 2016 67 repositories listed Syntology ran 7 of 29 samples · 22 unverified · 5 pointer-only (licence)Scene parsing is challenging for unrestricted open vocabulary and diverse scenes.
-
18 Aug 2016 22 repositories listed Syntology ran 1 of 3 samples · 2 unverified · 1 pointer-only (licence)Scene parsing, or recognizing and segmenting objects and stuff in an image, is one of the key problems in computer vision.
-
3 Jan 2018 9 repositories listedWe propose and study a task we name panoptic segmentation (PS).
-
15 Jan 2021 8 repositories listedThe proposed deep dual-resolution networks (DDRNets) are composed of two deep branches between which multiple bilateral fusions are performed.
-
4 Sep 2018 8 repositories listedTo capture richer context information, we further combine our interlaced sparse self-attention scheme with the conventional multi-scale context schemes including pyramid pooling~\citep{zhao2017pyramid} and atrous…
-
1 Apr 2019 7 repositories listed Syntology ran 1 of 13 samples · 12 unverifiedScene understanding of high resolution aerial images is of great importance for the task of automated monitoring in various remote sensing applications.
-
24 Feb 2020 6 repositories listed Syntology ran 1 of 8 samples · 7 unverified · 1 pointer-only (licence)A common practice to improve the performance is to attain high resolution feature maps with strong semantic representation.
-
10 Nov 2022 4 repositories listed Syntology ran 0 of 5 samples · 5 unverifiedHowever, such panoptic architectures do not truly unify image segmentation because they need to be trained individually on the semantic, instance, or panoptic segmentation to achieve the best performance.
-
1 Sep 2018 4 repositories listedWe notice information flow in convolutional neural networks is restricted inside local neighborhood regions due to the physical design of convolutional filters, which limits the overall understanding of complex scenes.
-
20 Jun 2020 3 repositories listedThis work introduces pyramidal convolution (PyConv), which is capable of processing the input at multiple filter scales.
-
21 Jul 2022 2 repositories listedA key algorithm for understanding the world is material segmentation, which assigns a label (metal, glass, etc.)
-
3 Dec 2021 2 repositories listedIn this paper, we propose a series of modular operations for effective geometric feature learning from 3D triangle meshes.
-
18 Jan 2021 2 repositories listedThis mental model captures geometric and semantic aspects of the scene, describes the environment at multiple levels of abstractions (e.
-
17 Nov 2020 2 repositories listedThis paper proposes minimal solvers that use combinations of imaged translational symmetries and parallel scene lines to jointly estimate lens undistortion with either affine rectification or focal length and absolute…
-
18 Jul 2020 2 repositories listed Syntology ran 4 of 14 samples · 10 unverifiedIn this paper, we propose a novel operator called malleable 2.
-
6 May 2020 2 repositories listed Syntology ran 2 of 7 samples · 5 unverifiedIn this paper, we propose a novel approach to address the high-resolution segmentation problem without using any high-resolution training data.
-
30 Mar 2020 2 repositories listed Syntology ran 2 of 11 samples · 9 unverifiedSpatial pooling has been proven highly effective in capturing long-range contextual information for pixel-wise prediction tasks, such as scene parsing.
-
1 Oct 2019 2 repositories listedDMNet is composed of multiple Dynamic Convolutional Modules (DCMs) arranged in parallel, each of which exploits context-aware filters to estimate semantic representation for a specific scale.
-
25 Jul 2019 2 repositories listedDrones or general Unmanned Aerial Vehicles (UAVs), endowed with computer vision function by on-board cameras and embedded systems, have become popular in a wide range of applications.
-
3 Apr 2019 2 repositories listedSemantic segmentation generates comprehensive understanding of scenes through densely predicting the category for each pixel.
-
6 Dec 2018 2 repositories listedLearning to insert an object instance into an image in a semantically coherent manner is a challenging and interesting problem.
-
19 Oct 2018 2 repositories listedWe introduce Synscapes -- a synthetic dataset for street scene parsing created using photorealistic rendering techniques, and show state-of-the-art results for training and validation as well as new types of analysis.
-
2 Jul 2024 1 repository listedThe existing contrastive learning methods mainly focus on single-grained representation learning, e.
-
15 Jun 2024 1 repository listedIn this paper, we leverage Prompt Images Guidance (PIG) to enhance UDA with supplementary night knowledge.
-
4 May 2024 1 repository listedBy leveraging pre-trained neural networks, accurate semantic segmentation of fruit in the field is achieved with only a few labeled images.
-
4 Apr 2024 1 repository listedIn this study, we take one step toward this new research area by exploring a feasible strategy to fully exploit VFM features for RGB-thermal scene parsing.
-
15 Mar 2024 1 repository listedA RANSAC estimator guided by a neural network fits these primitives to a depth map.
-
5 Feb 2024 1 repository listedThere are two challenges presented in parsing road scenes from UAV images: the complexity of processing high-resolution images and the dependency on extensive manual annotations required by traditional supervised deep…
-
30 Jan 2024 1 repository listedResearch on inter-network data connectivity is scant.
-
3 Dec 2023 1 repository listedMore importantly, we innovatively propose to learn to merge the over-divided clusters based on the local low-level geometric property similarities and the learned high-level feature similarities supervised by weak…
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections