Home › Datasets › task › Activity Recognition

Activity Recognition datasets

archive 2025-07-28

30 datasets carry the task tag "Activity Recognition" (the task itself: Activity Recognition), ordered by the archive's paper count. Page 1 of 1: 30 shown of 30. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Activity Recognition datasets 1–30 of 30

Augments the video-description dataset TACoS with short and single sentence descriptions.
45 papers · 1 benchmark
GazeFollow is a large-scale dataset annotated with the location of where people in images are looking.
38 papers · 1 benchmark
This dataset includes time-series data generated by accelerometer and gyroscope sensors (attitude, gravity, userAcceleration, and rotationRate).
35 papers · 0 benchmarks
The Drive&Act dataset is a state of the art multi modal benchmark for driver behavior recognition.
26 papers · 1 benchmark
MEVA (Multiview Extended Video with Activities)
Large-scale dataset for human activity recognition.
17 papers · 0 benchmarks
A database with 2,000 videos captured by surveillance cameras in real-world scenes.
16 papers · 1 benchmark
First-Person Hand Action Benchmark is a collection of RGB-D video sequences comprised of more than 100K frames of 45 daily hand action categories, involving 26 different objects in several hand configurations.
15 papers · 2 benchmarks
MoVi (Large Multipurpose Motion and Video Dataset)
Contains 60 female and 30 male actors performing a collection of 20 predefined everyday actions and sports movements, and one self-chosen movement.
12 papers · 1 benchmark
EgoHOS (Fine-Grained Egocentric Hand-Object Segmentation Dataset)
EgoHOS is a labeled dataset consisting of 11243 egocentric images with per-pixel segmentation labels of hands and objects being interacted with during a diverse array of daily activities.
9 papers · 0 benchmarks
Home Action Genome is a large-scale multi-view video database of indoor daily activities.
9 papers · 2 benchmarks
MISAW (MIcro-Surgical Anastomose Workflow recognition on training sessions)
The MISAW data set is composed of 27 sequences of micro-surgical anastomosis on artificial blood vessels performed by 3 surgeons and 3 engineering students.
8 papers · 1 benchmark
OPERAnet is a multimodal activity recognition dataset acquired from radio frequency and vision-based sensors.
8 papers · 0 benchmarks
EYTH (EgoYouTubeHands)
Includes egocentric videos containing hands in the wild.
7 papers · 0 benchmarks
This is a 3D action recognition dataset, also known as 3D Action Pairs dataset.
7 papers · 1 benchmark
A dataset which provides detailed annotations for activity recognition.
5 papers · 1 benchmark
Simitate is a hybrid benchmarking suite targeting the evaluation of approaches for imitation learning.
5 papers · 0 benchmarks
The Sims4Action Dataset: a videogame-based dataset for Synthetic→Real domain adaptation for human activity recognition.
5 papers · 0 benchmarks
Includes 11,771 samples of both human activities and falls performed by 30 subjects of ages ranging from 18 to 60 years.
4 papers · 0 benchmarks
A multi-sensor, multi-modal dataset, implemented to benchmark Human Activity Recognition(HAR) and Multi-modal Fusion algorithms.
3 papers · 0 benchmarks
MMDB (Multimodal Dyadic Behavior)
Multimodal Dyadic Behavior (MMDB) dataset is a unique collection of multimodal (video, audio, and physiological) recordings of the social and communicative behavior of toddlers.
3 papers · 0 benchmarks
PETRAW (PEg TRAnsfer Workflow recognition by different modalities)
PETRAW data set was composed of 150 sequences of peg transfer training sessions.
3 papers · 6 benchmarks
Contains annotations of human activity with different sub-actions, e.g., activity Ping-Pong with four sub-actions which are pickup-ball, hit, bounce-ball and serve.
2 papers · 0 benchmarks
MPHOI-72 (Multi-person Human-object Interaction Dataset 72)
MPHOI-72 is a multi-person human-object interaction dataset that can be used for a wide variety of HOI/activity recognition and pose estimation/object tracking tasks.
2 papers · 0 benchmarks
Stanford40 (Stanford 40 Actions)
The Stanford 40 Action Dataset contains images of humans performing 40 actions.
2 papers · 1 benchmark
CLAD (Complex and Long Activities Dataset)
CLAD (Compled and Long Activities Dataset) is an activity dataset which exhibits real-life and diverse scenarios of complex, temporally-extended human activities and actions.
1 paper · 0 benchmarks
HASCD (Human Activity Segmentation Challenge Dataset)
HASCD (Human Activity Segmentation Challenge Dataset) contains 250 annotated multivariate time series capturing 10.7 h of real-world human motion smartphone sensor data from 15 bachelor computer science students.
1 paper · 0 benchmarks
INDRA (INdian Dataset for RoAd crossing)
INDRA is a dataset capturing videos of Indian roads from the pedestrian point-of-view.
1 paper · 0 benchmarks
MOSAD (Mobile Sensing Human Activity Data Set)
MOSAD (Mobile Sensing Human Activity Data Set) is a multi-modal, annotated time series (TS) data set that contains 14 recordings of 9 triaxial smartphone sensor measurements (126 TS) from 6 human subjects performing (in part) 3 motion…
1 paper · 0 benchmarks
DAHLIA (DAily Human Life Activity)
DAHLIA dataset [1] is devoted to human activity recognition, which is a major issue for adapting smart-home services such as user assistance.
0 papers · 0 benchmarks
InfiniteRep is a synthetic, open-source dataset for fitness and physical therapy (PT) applications.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.