Papers › A Large-scale Varying-view RGB-D Action Dataset for Arbitrary-view Human Action Recognition
A Large-scale Varying-view RGB-D Action Dataset for Arbitrary-view Human Action Recognition
Yanli Ji, Feixiang Xu, Yang Yang, Fumin Shen, Heng Tao Shen, Wei-Shi Zheng
Current researches of action recognition mainly focus on single-view and multi-view recognition, which can hardly satisfies the requirements of human-robot interaction (HRI) applications to recognize actions from arbitrary views. The lack of datasets also sets up barriers. To provide data for arbitrary-view action recognition, we newly collect a large-scale RGB-D action dataset for arbitrary-view action analysis, including RGB videos, depth and skeleton sequences. The dataset includes action samples captured in 8 fixed viewpoints and varying-view sequences which covers the entire 360 degree view angles. In total, 118 persons are invited to act 40 action categories, and 25,600 video samples are collected. Our dataset involves more participants, more viewpoints and a large number of samples. More importantly, it is the first dataset containing the entire 360 degree varying-view sequences. The dataset provides sufficient data for multi-view, cross-view and arbitrary-view action analysis. Besides, we propose a View-guided Skeleton CNN (VS-CNN) to tackle the problem of arbitrary-view action recognition. Experiment results show that the VS-CNN achieves superior performance.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
No code repository is listed for this paper in the archive or in Syntology's graph.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Datasets
Introduced by this paper, per the archive.
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Skeleton Based Action Recognition | Varying-view RGB-D Action-Skeleton | VS-CNN | Accuracy (AV I) | 57% | #1 of 7 | Archive leaderboard | report |
| Skeleton Based Action Recognition | Varying-view RGB-D Action-Skeleton | VS-CNN | Accuracy (AV II) | 75% | #1 of 7 | Archive leaderboard | report |
| Skeleton Based Action Recognition | Varying-view RGB-D Action-Skeleton | VS-CNN | Accuracy (CS) | 76% | #1 of 7 | Archive leaderboard | report |
| Skeleton Based Action Recognition | Varying-view RGB-D Action-Skeleton | VS-CNN | Accuracy (CV I) | 29% | #1 of 7 | Archive leaderboard | report |
| Skeleton Based Action Recognition | Varying-view RGB-D Action-Skeleton | VS-CNN | Accuracy (CV II) | 71% | #1 of 7 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections