Papers › Skeleton-based Action Recognition via Temporal-Channel Aggregation
Skeleton-based Action Recognition via Temporal-Channel Aggregation
Shengqin Wang, Yongji Zhang, Minghao Zhao, Hong Qi, Kai Wang, Fenglin Wei, Yu Jiang
Skeleton-based action recognition methods are limited by the semantic extraction of spatio-temporal skeletal maps. However, current methods have difficulty in effectively combining features from both temporal and spatial graph dimensions and tend to be thick on one side and thin on the other. In this paper, we propose a Temporal-Channel Aggregation Graph Convolutional Networks (TCA-GCN) to learn spatial and temporal topologies dynamically and efficiently aggregate topological features in different temporal and channel dimensions for skeleton-based action recognition. We use the Temporal Aggregation module to learn temporal dimensional features and the Channel Aggregation module to efficiently combine spatial dynamic channel-wise topological features with temporal dynamic topological features. In addition, we extract multi-scale skeletal features on temporal modeling and fuse them with an attention mechanism. Extensive experiments show that our model results outperform state-of-the-art methods on the NTU RGB+D, NTU RGB+D 120, and NW-UCLA datasets.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Skeleton Based Action Recognition | N-UCLA | TCA-GCN | Accuracy | 97.0 | #10 of 25 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D | TCA-GCN | Accuracy (CS) | 92.8 | #28 of 135 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D | TCA-GCN | Accuracy (CV) | 97.0 | #28 of 135 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D | TCA-GCN | Ensembled Modalities | 4 | #28 of 135 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D 120 | TCA-GCN | Accuracy (Cross-Setup) | 90.8 | #20 of 83 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D 120 | TCA-GCN | Accuracy (Cross-Subject) | 89.4 | #20 of 83 | Archive leaderboard | report |
| Skeleton Based Action Recognition | NTU RGB+D 120 | TCA-GCN | Ensembled Modalities | 4 | #20 of 83 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections