Datasets › EGTEA
EGTEA (EGTEA Gaze+)
Extended GTEA Gaze+ EGTEA Gaze+ is a large-scale dataset for FPV actions and gaze. It subsumes GTEA Gaze+ and comes with HD videos (1280x960), audios, gaze tracking data, frame-level action annotations, and pixel-level hand masks at sampled frames. Specifically, EGTEA Gaze+ contains 28 hours (de-identified) of cooking activities from 86 unique sessions of 32 subjects. These videos come with audios and gaze tracking (30Hz). We have further provided human annotations of actions (human-object interactions) and hand masks.
The action annotations include 10325 instances of fine-grained actions, such as "Cut bell pepper" or "Pour condiment (from) condiment container into salad".
The hand annotations consist of 15,176 hand masks from 13,847 frames from the videos.
Source: http://cbs.ic.gatech.edu/fpv/ Image Source: http://cbs.ic.gatech.edu/fpv/
Benchmarks archive 2025-07-28
All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Egocentric Activity Recognition | EGTEA | LaViLa (Finetuned, TimeSformer-L) Average Accuracy 81.75 | Learning Video Representations from Large Language Models | facebookresearch/lavila +2 | 6 | Compare |
| Action Anticipation | EGTEA | UADT Top-1 Accuracy 68.4 | Uncertainty-aware Action Decoupling Transformer for... | — | 3 | Compare |
| Long-tail Learning | EGTEA | CDB-loss (3D- ResNeXt101) Average Precision 63.86 | Class-Wise Difficulty-Balanced Loss for Solving Class-Imbalance | hitachi-rd-cv/CDB-loss | 3 | Compare |
Papers archive 2025-07-28
12 shown of 12 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 100. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
No licence recorded in the archive. Absence here is not a statement about the dataset's terms.
Modalities archive 2025-07-28
No modality tagged.
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
No variants listed.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections