Papers › Revisiting the Encoding of Satellite Image Time Series

Revisiting the Encoding of Satellite Image Time Series

3 May 2023arXiv:2305.02086archive 2025-07-28

Xin Cai, Yaxin Bi, Peter Nicholl, Roy Sterritt

Satellite Image Time Series (SITS) representation learning is complex due to high spatiotemporal resolutions, irregular acquisition times, and intricate spatiotemporal interactions. These challenges result in specialized neural network architectures tailored for SITS analysis. The field has witnessed promising results achieved by pioneering researchers, but transferring the latest advances or established paradigms from Computer Vision (CV) to SITS is still highly challenging due to the existing suboptimal representation learning framework. In this paper, we develop a novel perspective of SITS processing as a direct set prediction problem, inspired by the recent trend in adopting query-based transformer decoders to streamline the object detection or image segmentation pipeline. We further propose to decompose the representation learning process of SITS into three explicit steps: collect-update-distribute, which is computationally efficient and suits for irregularly-sampled and asynchronous temporal satellite observations. Facilitated by the unique reformulation, our proposed temporal learning backbone of SITS, initially pre-trained on the resource efficient pixel-set format and then fine-tuned on the downstream dense prediction tasks, has attained new state-of-the-art (SOTA) results on the PASTIS benchmark dataset. Specifically, the clear separation between temporal and spatial components in the semantic/panoptic segmentation pipeline of SITS makes us leverage the latest advances in CV, such as the universal image segmentation architecture, resulting in a noticeable 2.5 points increase in mIoU and 8.8 points increase in PQ, respectively, compared to the best scores reported so far.

PaperPDFCode

Code

TotalVariation/Exchanger4SITS officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Image SegmentationObject DetectionPanoptic SegmentationRepresentation LearningSegmentationSemantic SegmentationTime Seriesobject-detection

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Panoptic Segmentation PASTIS Exchanger+Mask2Former PQ 52.6 #1 of 3 Archive leaderboard report
Panoptic Segmentation PASTIS Exchanger+Mask2Former RQ 61.6 #1 of 3 Archive leaderboard report
Panoptic Segmentation PASTIS Exchanger+Mask2Former SQ 84.6 #1 of 3 Archive leaderboard report
Panoptic Segmentation PASTIS Exchanger+Unet+PaPs PQ 47.8 #2 of 3 Archive leaderboard report
Panoptic Segmentation PASTIS Exchanger+Unet+PaPs RQ 58.9 #2 of 3 Archive leaderboard report
Panoptic Segmentation PASTIS Exchanger+Unet+PaPs SQ 80.3 #2 of 3 Archive leaderboard report
Semantic Segmentation PASTIS Exchanger+Mask2Former Mean IoU (test) 67.9 #1 of 3 Archive leaderboard report
Semantic Segmentation PASTIS Exchanger+Unet Mean IoU (test) 66.8 #2 of 3 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections