Papers › Spatial-Temporal Correlation and Topology Learning for Person Re-Identification in Videos

Spatial-Temporal Correlation and Topology Learning for Person Re-Identification in Videos

15 Apr 2021CVPR 2021 1arXiv:2104.08241archive 2025-07-28

Jiawei Liu, Zheng-Jun Zha, Wei Wu, Kecheng Zheng, Qibin Sun

Video-based person re-identification aims to match pedestrians from video sequences across non-overlapping camera views. The key factor for video person re-identification is to effectively exploit both spatial and temporal clues from video sequences. In this work, we propose a novel Spatial-Temporal Correlation and Topology Learning framework (CTL) to pursue discriminative and robust representation by modeling cross-scale spatial-temporal correlation. Specifically, CTL utilizes a CNN backbone and a key-points estimator to extract semantic local features from human body at multiple granularities as graph nodes. It explores a context-reinforced topology to construct multi-scale graphs by considering both global contextual information and physical connections of human body. Moreover, a 3D graph convolution and a cross-scale graph convolution are designed, which facilitate direct cross-spacetime and cross-scale information propagation for capturing hierarchical spatial-temporal dependencies and structural information. By jointly performing the two convolutions, CTL effectively mines comprehensive clues that are complementary with appearance information to enhance representational capacity. Extensive experiments on two video benchmarks have demonstrated the effectiveness of the proposed method and the state-of-the-art performance.

PaperPDFConference PDF

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Person Re-IdentificationVideo DeinterlacingVideo-Based Person Re-Identification

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Video Deinterlacing MSU Deinterlacer Benchmark ST-Deint FPS on CPU 2.7 #10 of 31 Archive leaderboard report
Video Deinterlacing MSU Deinterlacer Benchmark ST-Deint PSNR 40.869 #10 of 31 Archive leaderboard report
Video Deinterlacing MSU Deinterlacer Benchmark ST-Deint SSIM 0.964 #10 of 31 Archive leaderboard report
Video Deinterlacing MSU Deinterlacer Benchmark ST-Deint Subjective 0.550 #10 of 31 Archive leaderboard report
Video Deinterlacing MSU Deinterlacer Benchmark ST-Deint VMAF 94.36 #10 of 31 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

Convolution

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections