Papers › U-Net Ensemble for Enhanced Semantic Segmentation in Remote Sensing Imagery

U-Net Ensemble for Enhanced Semantic Segmentation in Remote Sensing Imagery

8 Jun 2024Remote Sensing 2024 6archive 2025-07-28

Ivica Dimitrovski, Vlatko Spasev, Suzana Loshkovska, Ivan Kitanovski

Semantic segmentation of remote sensing imagery stands as a fundamental task within the domains of both remote sensing and computer vision. Its objective is to generate a comprehensive pixel-wise segmentation map of an image, assigning a specific label to each pixel. This facilitates in-depth analysis and comprehension of the Earth’s surface. In this paper, we propose an approach for enhancing semantic segmentation performance by employing an ensemble of U-Net models with three different backbone networks: Multi-Axis Vision Transformer, ConvFormer, and EfficientNet. The final segmentation maps are generated through a geometric mean ensemble method, leveraging the diverse representations learned by each backbone network. The effectiveness of the base U-Net models and the proposed ensemble is evaluated on multiple datasets commonly used for semantic segmentation tasks in remote sensing imagery, including LandCover.ai, LoveDA, INRIA, UAVid, and ISPRS Potsdam datasets. Our experimental results demonstrate that the proposed approach achieves state-of-the-art performance, showcasing its effectiveness and robustness in accurately capturing the semantic information embedded within remote sensing images.

PaperPDF

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

SegmentationSegmentation Of Remote Sensing ImagerySemantic Segmentation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Semantic Segmentation ISPRS Potsdam U-Net (ConvFormer-M36) Mean IoU 89.45 #20 of 20 Archive leaderboard report
Semantic Segmentation LandCover.ai U-Net (ConvFormer-M36) mIoU 87.64 #1 of 1 Archive leaderboard report
Semantic Segmentation LoveDA U-Net (MaxViT-S) Category mIoU 56.16 #1 of 19 Archive leaderboard report
Semantic Segmentation UAVid U-Net Ensemble Mean IoU 73.34 #1 of 10 Archive leaderboard report
Semantic Segmentation UAVid U-Net (MaxViT-S) Mean IoU 71.88 #2 of 10 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

1x1 ConvolutionAbsolute Position EncodingsAdamAttentionAverage PoolingBASEBPEBatch NormalizationConcatenated Skip ConnectionConvolutionDense ConnectionsDepthwise ConvolutionDepthwise Separable ConvolutionDropoutEfficientNetInverted Residual BlockLabel SmoothingLayer NormalizationLinear LayerMax PoolingMulti-Head AttentionPointwise ConvolutionPosition-Wise Feed-Forward LayerRMSPropReLUResidual ConnectionSigmoid ActivationSoftmaxSqueeze-and-Excitation BlockTransformerU-NetVision Transformer

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections