Papers › Multi-scale Context-aware Network with Transformer for Gait Recognition

Multi-scale Context-aware Network with Transformer for Gait Recognition

7 Apr 2022ICCV 2021 10arXiv:2204.03270archive 2025-07-28

Duowang Zhu, Xiaohu Huang, Xinggang Wang, Bo Yang, Botao He, Wenyu Liu, Bin Feng

Although gait recognition has drawn increasing research attention recently, since the silhouette differences are quite subtle in spatial domain, temporal feature representation is crucial for gait recognition. Inspired by the observation that humans can distinguish gaits of different subjects by adaptively focusing on clips of varying time scales, we propose a multi-scale context-aware network with transformer (MCAT) for gait recognition. MCAT generates temporal features across three scales, and adaptively aggregates them using contextual information from both local and global perspectives. Specifically, MCAT contains an adaptive temporal aggregation (ATA) module that performs local relation modeling followed by global relation modeling to fuse the multi-scale features. Besides, in order to remedy the spatial feature corruption resulting from temporal operations, MCAT incorporates a salient spatial feature learning (SSFL) module to select groups of discriminative spatial features. Extensive experiments conducted on three datasets demonstrate the state-of-the-art performance. Concretely, we achieve rank-1 accuracies of 98.7%, 96.2% and 88.7% under normal walking, bag-carrying and coat-wearing conditions on CASIA-B, 97.5% on OU-MVLP and 50.6% on GREW. The source code will be available at https://github.com/zhuduowang/MCAT.git.

PaperPDFConference PDF

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Gait RecognitionMultiview Gait Recognition

1 archive task tag without a task page not shown.

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Gait Recognition OUMVLP CSTL Averaged rank-1 acc(%) 91.0 #3 of 7 Archive leaderboard report
Multiview Gait Recognition CASIA-B CSTL Accuracy (Cross-View, Avg) 94.5 #3 of 12 Archive leaderboard report
Multiview Gait Recognition CASIA-B CSTL BG#1-2 94.8 #3 of 12 Archive leaderboard report
Multiview Gait Recognition CASIA-B CSTL CL#1-2 88.7 #3 of 12 Archive leaderboard report
Multiview Gait Recognition CASIA-B CSTL NM#5-6 98.7 #3 of 12 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections