Datasets › LIP

LIP (Look into Person)

Introduced by Ke Gong et al. in Look into Person: Self-supervised Structure-sensitive Learning and A New Benchmark for Human Parsing1 Jan 2017 archive 2025-07-28

The LIP (Look into Person) dataset is a large-scale dataset focusing on semantic understanding of a person. It contains 50,000 images with elaborated pixel-wise annotations of 19 semantic human part labels and 2D human poses with 16 key points. The images are collected from real-world scenarios and the subjects appear with challenging poses and view, heavy occlusions, various appearances and low resolution.

Source: http://sysu-hcp.net/lip/ Image Source: http://sysu-hcp.net/lip/

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Semantic Segmentation LIP val Hulk(Finetune, ViT-L) mIoU 66.02% Hulk: A Universal Knowledge Translator for Human-Centric Tasks opengvlab/humanbench +1 13 Compare

Papers archive 2025-07-28

10 shown of 10 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 61. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Hulk: A Universal Knowledge Translator for Human-Centric Tasks 2 2 4 Dec 2023 ran 12 of 24 samples (12 unverified)
Beyond Appearance: a Semantic Controllable Self-Supervised Learning Framework for Human-Centric Visual Tasks 4 1 30 Mar 2023 ran 1 of 2 samples (1 unverified; 1 pointer-only for licence)
UniHCP: A Unified Model for Human-Centric Perceptions 1 1 6 Mar 2023 ran 7 of 13 samples (6 unverified)
Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation 11 3 24 Sep 2019 ran 4 of 9 samples (5 unverified; 1 pointer-only for licence)
High-Resolution Representations for Labeling Pixels and Regions 39 1 9 Apr 2019 ran 3 of 18 samples (15 unverified; 5 pointer-only for licence)
Devil in the Details: Towards Accurate Single and Multiple Human Parsing 2 1 17 Sep 2018 ran 1 of 1 samples (0 unverified; 1 pointer-only for licence)
Mutual Learning to Adapt for Joint Human Parsing and Pose Estimation 0 1 1 Sep 2018 not harvested
Macro-Micro Adversarial Network for Human Parsing 1 1 22 Jul 2018 not harvested
Look into Person: Joint Body Parsing & Pose Estimation Network and A New Benchmark 3 1 5 Apr 2018 not harvested
Look into Person: Self-supervised Structure-sensitive Learning and A New Benchmark for Human Parsing 1 1 16 Mar 2017 not harvested

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Custom (research, non-research, non-commercial)

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • LIP val
  • LIP

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections