Datasets › NTU RGB+D 120

NTU RGB+D 120

Introduced by Jun Liu et al. in NTU RGB+D 120: A Large-Scale Benchmark for 3D Human Activity Understanding archive 2025-07-28

NTU RGB+D 120 is a large-scale dataset for RGB+D human action recognition, which is collected from 106 distinct subjects and contains more than 114 thousand video samples and 8 million frames. This dataset contains 120 different action classes including daily, mutual, and health-related activities.

Source: NTU RGB+D 120: A Large-Scale Benchmark for 3D Human Activity Understanding

Benchmarks archive 2025-07-28

All 8 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

30 shown of 105 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 137. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Zero-shot Skeleton-based Action Recognition with Prototype-guided Feature Alignment 1 2 1 Jul 2025 not harvested
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation 1 1 16 Apr 2025 not harvested
DSTSA-GCN: Advancing Skeleton-Based Gesture Recognition with Semantic-Aware Spatio-Temporal Topology Modeling 1 2 21 Jan 2025 not harvested
MSA-GCN: Exploiting Multi-Scale Temporal Dynamics With Adaptive Graph Convolution for Skeleton-Based Action Recognition 0 1 19 Dec 2024 not harvested
USDRL: Unified Skeleton-Based Dense Representation Learning with Multi-Grained Feature Decorrelation 1 1 12 Dec 2024 not harvested
Revealing Key Details to See Differences: A Novel Prototypical Perspective for Skeleton-based Action Recognition 1 1 28 Nov 2024 not harvested
TDSM: Triplet Diffusion for Skeleton-Text Matching in Zero-Shot Action Recognition 1 1 16 Nov 2024 not harvested
Joint Mixing Data Augmentation for Skeleton-based Action Recognition 1 1 13 Oct 2024 not harvested
CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition 1 1 9 Oct 2024 ran 2 of 2 samples (0 unverified)
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition 1 1 23 Sep 2024 not harvested
EPAM-Net: An Efficient Pose-driven Attention-guided Multimodal Network for Video Action Recognition 1 1 10 Aug 2024 not harvested
Joint-Partition Group Attention for skeleton-based action recognition 1 2 30 Jul 2024 not harvested
Multi-Modality Co-Learning for Efficient Skeleton-based Action Recognition 1 1 22 Jul 2024 ran 9 of 10 samples (1 unverified; 10 pointer-only for licence)
SA-DVAE: Improving Zero-Shot Skeleton-Based Action Recognition by Disentangled Variational Autoencoders 1 4 18 Jul 2024 not harvested
Shap-Mix: Shapley Value Guided Mixing for Long-Tailed Skeleton Based Action Recognition 1 1 17 Jul 2024 not harvested
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition 1 1 19 Jun 2024 ran 3 of 3 samples (0 unverified)
Fine-Grained Side Information Guided Dual-Prompts for Zero-Shot Skeleton Action Recognition 0 1 11 Apr 2024 not harvested
LLMs are Good Action Recognizers 0 1 31 Mar 2024 not harvested
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition 1 2 14 Mar 2024 ran 3 of 3 samples (0 unverified)
Explore Human Parsing Modality for Action Recognition 1 1 4 Jan 2024 ran 7 of 10 samples (3 unverified; 10 pointer-only for licence)
MaskCLR: Attention-Guided Contrastive Learning for Robust Action Representation Learning 0 1 1 Jan 2024 not harvested
BlockGCN: Redefine Topology Awareness for Skeleton-Based Action Recognition 1 1 1 Jan 2024 not harvested
A Dense-Sparse Complementary Network for Human Action Recognition based on RGB and Skeleton Modalities 1 1 28 Dec 2023 not harvested
DVANet: Disentangling View and Action Features for Multi-View Action Recognition 1 1 10 Dec 2023 not harvested
STEP CATFormer: Spatial-Temporal Effective Body-Part Cross Attention Transformer for Skeleton-based Action Recognition 1 1 6 Dec 2023 ran 10 of 13 samples (3 unverified)
Just Add π! Pose Induced Video Transformers for Understanding Activities of Daily Living 1 2 30 Nov 2023 not harvested
Multi-Semantic Fusion Model for Generalized Zero-Shot Skeleton-Based Action Recognition 1 1 18 Sep 2023 not harvested
Masked Motion Predictors are Strong 3D Action Representation Learners 1 1 14 Aug 2023 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
Zero-shot Skeleton-based Action Recognition via Mutual Information Estimation and Maximization 1 1 7 Aug 2023 not harvested
Integrating Human Parsing and Pose Network for Human Action Recognition 1 1 16 Jul 2023 not harvested

The full list of 105 is in the JSON twin.

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Custom (research-only)

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • NTU RGB+D 120

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections