Home › Datasets › task › Action Segmentation
Action Segmentation datasets
archive 2025-07-28
18 datasets carry the task tag "Action Segmentation" (the task itself: Action Segmentation), ordered by the archive's paper count. Page 1 of 1: 18 shown of 18. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Action Segmentation datasets 1–18 of 18
The Breakfast Actions Dataset comprises of 10 actions related to breakfast preparation, performed by 52 different individuals in 18 different kitchens.
179 papers · 6 benchmarks
GTEA (Georgia Tech Egocentric Activity)
The Georgia Tech Egocentric Activities (GTEA) dataset contains seven types of daily activities such as making sandwich, tea, or coffee.
120 papers · 2 benchmarks
The COIN dataset (a large-scale dataset for COmprehensive INstructional video analysis) consists of 11,827 videos related to 180 different tasks in 12 domains (e.g., vehicles, gadgets, etc.) related to our daily life.
105 papers · 2 benchmarks
JIGSAWS (JHU-ISI Gesture and Skill Assessment Working Set)
The JHU-ISI Gesture and Skill Assessment Working Set (JIGSAWS) is a surgical activity dataset for human motion modeling.
105 papers · 3 benchmarks
Assembly101 is a new procedural activity dataset featuring 4321 videos of people assembling and disassembling 101 "take-apart" toy vehicles.
57 papers · 4 benchmarks
Activity recognition research has shifted focus from distinguishing full-body motion patterns to recognizing complex interactions of multiple entities.
35 papers · 2 benchmarks
A large-scale 4D egocentric dataset with rich annotations, to catalyze the research of category-level human-object interaction.
23 papers · 0 benchmarks
The Watch-n-Patch dataset was created with the focus on modeling human activities, comprising multiple actions in a completely unsupervised setting.
13 papers · 0 benchmarks
EgoExoLearn is a fascinating dataset designed to bridge the gap between egocentric and exocentric views of procedural activities.
12 papers · 3 benchmarks
EgoProceL is a large-scale dataset for procedure learning.
9 papers · 0 benchmarks
We address the problem of automatically learning the main steps to complete a certain task, such as changing a car tire, from a set of narrated instruction videos.
8 papers · 2 benchmarks
The TUM Kitchen dataset is an action recognition dataset that contains 20 video sequences captured by 4 cameras with overlapping views.
7 papers · 0 benchmarks
A dataset which provides detailed annotations for activity recognition.
5 papers · 1 benchmark
HA-ViD (HA-ViD: A Human Assembly Video Dataset)
Understanding comprehensive assembly knowledge from videos is critical for futuristic ultra-intelligent industry.
1 paper · 0 benchmarks
InHARD (Industrial Human Action Recognition Dataset in the Context of Industrial Collaborative Robotics)
We introduce a RGB+S dataset named “Industrial Human Action Recognition Dataset” (InHARD) from a real-world setting for industrial human action recognition with over 2 million frames, collected from 16 distinct subjects.
1 paper · 0 benchmarks
LARa (Logistic Activity Recognition Challenge)
LARa is the first freely accessible logistics-dataset for human activity recognition.
1 paper · 0 benchmarks
UW IOM (University of Washington Indoor Object Manipulation)
Comprises twenty individuals picking up and placing objects of varying weights to and from cabinet and table locations at various heights.
1 paper · 0 benchmarks
SICS-155 (Phase Recognition in Small Incision Cataract Surgery Videos)
Cataract is the leading cause of blindness worldwide, most affecting life in low- and middle-income countries (LMICs).
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.