Browse State-of-the-Art › Thermal Image Segmentation
Thermal Image Segmentation
70 papers with code · 7 benchmarks · 4 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
7 leaderboard tables shown for this task, 7 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| MFN Dataset (55 rows) | RoadFormer+ (ConvNeXt-L) | RoadFormer+: Delivering RGB-X Scene Parsing through Scale-Aware... | — | — | Compare |
| PST900 (22 rows) | SHIFNet | Unveiling the Potential of Segment Anything Model 2 for... | code | — | Compare |
| RGB-T-Glass-Segmentation (22 rows) | RGB-T-Glass-Segmentation | Glass Segmentation with RGB-Thermal Image Pairs | code | — | Compare |
| Noisy RS RGB-T Dataset (6 rows) | CMNeXt (B4) | Delivering Arbitrary-Modal Semantic Segmentation | code | Syntology ran 6 of 7 samples · 1 unverified | Compare |
| KP day-night (5 rows) | HAPNet | HAPNet: Toward Superior RGB-Thermal Scene Parsing via Hybrid,... | code | — | Compare |
| SCUT-Seg Dataset (1 row) | FTNet | FTNet: Feature Transverse Network for Thermal Image Semantic Segmentation | code | — | Compare |
| SODA Dataset (1 row) | FTNet | FTNet: Feature Transverse Network for Thermal Image Semantic Segmentation | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
4 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 70 papers with code (84 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
18 May 2015 487 repositories listed Syntology ran 510 of 757 samples · 247 unverified · 426 pointer-only (licence)There is large consent that successful training of deep networks requires many thousand annotated training samples.
-
25 Mar 2021 80 repositories listed Syntology ran 108 of 207 samples · 99 unverified · 43 pointer-only (licence)This paper presents a new vision Transformer, called Swin Transformer, that capably serves as a general-purpose backbone for computer vision.
-
17 Jun 2017 77 repositories listed Syntology ran 3 of 7 samples · 4 unverified · 3 pointer-only (licence)To handle the problem of segmenting objects at multiple scales, we design modules which employ atrous convolution in cascade or in parallel to capture multi-scale context by adopting multiple atrous rates.
-
2 Nov 2015 74 repositories listed Syntology ran 9 of 44 samples · 35 unverified · 10 pointer-only (licence)We show that SegNet provides good performance with competitive inference time and more efficient inference memory-wise as compared to other architectures.
-
4 Dec 2016 67 repositories listed Syntology ran 7 of 29 samples · 22 unverified · 5 pointer-only (licence)Scene parsing is challenging for unrestricted open vocabulary and diverse scenes.
-
14 Nov 2014 51 repositories listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)Convolutional networks are powerful visual models that yield hierarchies of features.
-
20 Aug 2019 42 repositories listed Syntology ran 3 of 34 samples · 31 unverified · 16 pointer-only (licence)High-resolution representations are essential for position-sensitive vision problems, such as human pose estimation, semantic segmentation, and object detection.
-
18 Jul 2018 34 repositories listed Syntology ran 5 of 28 samples · 23 unverified · 2 pointer-only (licence)Implementation of different kinds of Unet Models for Image Segmentation - Unet , RCNN-Unet, Attention Unet, RCNN-Attention Unet, Nested Unet
-
31 May 2021 28 repositories listed Syntology ran 48 of 86 samples · 38 unverified · 15 pointer-only (licence)We present SegFormer, a simple, efficient yet powerful semantic segmentation framework which unifies Transformers with lightweight multilayer perception (MLP) decoders.
-
12 Feb 2019 24 repositories listed Syntology ran 0 of 32 samples · 32 unverifiedThe encoder-decoder framework is state-of-the-art for offline semantic image segmentation.
-
2 Aug 2018 21 repositories listed Syntology ran 6 of 18 samples · 12 unverified · 1 pointer-only (licence)Semantic segmentation requires both rich spatial information and sizeable receptive field.
-
27 Apr 2017 18 repositories listedWe focus on the challenging task of real-time semantic segmentation in this paper.
-
14 Jun 2017 14 repositories listed Syntology ran 3 of 4 samples · 1 unverified · 2 pointer-only (licence)As a result they are huge in terms of parameters and number of operations; hence slow too.
-
9 Oct 2017 13 repositories listedA comprehensive set of experiments on the publicly available Cityscapes dataset demonstrates that our system achieves an accuracy that is similar to the state of the art, while being orders of magnitude faster to…
-
8 Jan 2019 12 repositories listed Syntology ran 7 of 12 samples · 5 unverifiedIn this work, we perform a detailed study of this minimally extended version of Mask R-CNN with FPN, which we refer to as Panoptic FPN, and show it is a robust and accurate baseline for both tasks.
-
9 Sep 2018 12 repositories listed Syntology ran 0 of 7 samples · 7 unverified · 5 pointer-only (licence)Specifically, we append two types of attention modules on top of traditional dilated FCN, which model the semantic interdependencies in spatial and channel dimensions respectively.
-
23 Mar 2018 12 repositories listed Syntology ran 1 of 8 samples · 7 unverified · 7 pointer-only (licence)In this paper, we explore the impact of global contextual information in semantic segmentation by introducing the Context Encoding Module, which captures the semantic context of scenes and selectively highlights…
-
12 May 2021 8 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedIn this paper we introduce Segmenter, a transformer model for semantic segmentation.
-
27 Feb 2017 5 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedThis framework 1) effectively enlarges the receptive fields (RF) of the network to aggregate global information; 2) alleviates what we call the "gridding issue" caused by the standard dilated convolution operation.
-
13 Nov 2020 4 repositories listedIn this paper, we propose an efficient and robust RGB-D segmentation approach that can be optimized to a high degree using NVIDIA TensorRT and, thus, is well suited as a common initial processing step in a complex…
-
28 Nov 2018 4 repositories listed Syntology ran 9 of 14 samples · 5 unverifiedCompared with the non-local block, the proposed recurrent criss-cross attention module requires 11x less GPU memory usage.
-
19 Mar 2018 4 repositories listed Syntology ran 1 of 11 samples · 10 unverifiedConvolutional neural networks (CNN) are limited by the lack of capability to handle geometric information due to the fixed grid kernel structure.
-
24 Nov 2016 4 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Therefore, additional processing steps have to be performed in order to obtain pixel-accurate segmentation masks at the full image resolution.
-
8 Aug 2023 3 repositories listed Syntology ran 13 of 27 samples · 14 unverified · 11 pointer-only (licence)We first conduct systematic analyses about the components of image fusion, investigating the correlation with segmentation robustness under adversarial perturbations.
-
18 Aug 2021 3 repositories listed Syntology ran 8 of 14 samples · 6 unverified · 14 pointer-only (licence)To effectively fuse cross-modal features in the shared learning network, we propose a cross-enhanced integration module (CIM) and then propagate the fused feature to the next layer for integrating cross-level…
-
25 Apr 2018 3 repositories listed Syntology ran 0 of 8 samples · 8 unverifiedMost existing methods of semantic segmentation still suffer from two aspects of challenges: intra-class inconsistency and inter-class indistinction.
-
4 Aug 2023 2 repositories listed Syntology ran 19 of 26 samples · 7 unverified · 13 pointer-only (licence)Multi-modality image fusion and segmentation play a vital role in autonomous driving and robotic operation.
-
25 Apr 2021 2 repositories listed Syntology ran 7 of 11 samples · 4 unverified · 11 pointer-only (licence)We also develop a token-based multi-task decoder to simultaneously perform saliency and boundary detection by introducing task-related tokens and a novel patch-task-attention mechanism.
-
23 Jul 2020 2 repositories listed Syntology ran 1 of 2 samples · 1 unverified · 2 pointer-only (licence)The explicitly extracted edge information goes together with saliency to give more emphasis to the salient regions and object boundaries.
-
17 Jul 2020 2 repositories listedDepth information has proven to be a useful cue in the semantic segmentation of RGB-D images for providing a geometric counterpart to the RGB representation.
Syntology lines on 26 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections