Browse State-of-the-Art › Human Parsing
Human Parsing
62 papers with code · 2 benchmarks · 3 datasets archive 2025-07-28
Human parsing is the task of segmenting a human image into different fine-grained semantic parts such as head, torso, arms and legs.
( Image credit: Multi-Human-Parsing (MHP) )
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
2 leaderboard tables shown for this task, 2 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| 4D-DRESS (6 rows) | Graphonomy_Inner | Graphonomy: Universal Image Parsing via Graph Reasoning and Transfer | code | — | Compare |
| PASCAL Context (1 row) | InvPT | InvPT: Inverted Pyramid Multi-task Transformer for Dense Scene... | code | Syntology ran 2 of 4 samples · 2 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
3 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 62 papers with code (125 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
30 Mar 2023 4 repositories listed Syntology ran 1 of 2 samples · 1 unverified · 1 pointer-only (licence)Unlike the existing self-supervised learning methods, prior knowledge from human images is utilized in SOLIDER to build pseudo semantic labels and import more semantic information into the learned representation.
-
28 Nov 2018 4 repositories listed Syntology ran 9 of 14 samples · 5 unverifiedCompared with the non-local block, the proposed recurrent criss-cross attention module requires 11x less GPU memory usage.
-
7 Nov 2022 3 repositories listedFirstly, individual body part appearance is not as discriminative as global appearance (two distinct IDs might have the same local appearance), this means standard ReID training objectives using identity labels are not…
-
1 Oct 2022 3 repositories listedIn addition, we propose a network named M2Net, which integrates multi-modality features from the RGB images, contour images and human parsing images.
-
5 Apr 2018 3 repositories listedTo further explore and take advantage of the semantic correlation of these two tasks, we propose a novel joint human parsing and pose estimation network to explore efficient context modeling, which can simultaneously…
-
29 Jun 2023 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)In this work, we propose milliFlow, a novel deep learning approach to estimate scene flow as complementary motion information for mmWave point cloud, serving as an intermediate level of features and directly benefiting…
-
1 Jan 2023 2 repositories listedHuman parsing aims to partition humans in image or video into multiple pixel-level semantic parts.
-
31 May 2022 2 repositories listedIn this work, we present a text-driven controllable framework, Text2Human, for a high-quality and diverse human generation.
-
8 Mar 2021 2 repositories listedA recent pioneering work employed knowledge distillation to reduce the dependency of human parsing, where the try-on images produced by a parser-based method are used as supervisions to train a "student" network without…
-
26 Jan 2021 2 repositories listedPrior highly-tuned image parsing models are usually studied in a certain domain with a specific set of semantic labels and can hardly be adapted into other scenarios (e.
-
17 Dec 2019 2 repositories listed Syntology ran 1 of 15 samples · 14 unverifiedDespite great success in human parsing, progress for parsing other deformable articulated objects, like animals, is still limited by the lack of labeled data.
-
22 Oct 2019 2 repositories listed Syntology ran 2 of 3 samples · 1 unverifiedTo tackle the problem of learning with label noises, this work introduces a purification strategy, called Self-Correction for Human Parsing (SCHP), to progressively promote the reliability of the supervised labels as…
-
30 Nov 2018 2 repositories listed Syntology ran 2 of 7 samples · 5 unverifiedModels need to distinguish different human instances in the image panel and learn rich features to represent the details of each instance.
-
17 Sep 2018 2 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Human parsing has received considerable interest due to its wide application potentials.
-
10 Apr 2018 2 repositories listedDespite the noticeable progress in perceptual tasks like detection, instance segmentation and human parsing, computers still perform unsatisfactorily on visually understanding humans in crowded scenes, such as group…
-
19 May 2017 2 repositories listedTo address the multi-human parsing problem, we introduce a new multi-human parsing (MHP) dataset and a novel multi-human parsing model named MH-Parser.
-
16 Mar 2025 1 repository listedFinally, we propose a Limb-aware Texture Fusion (LTF) module to estimate high-quality details in limb regions by fusing textures of the clothing and the human body with the guidance of explicit limb-aware features.
-
16 Dec 2024 1 repository listedThe gait, as a kind of soft biometric characteristic, can reflect the distinct walking patterns of individuals at a distance, exhibiting a promising technique for unrestrained human identification.
-
16 Nov 2024 1 repository listedIn particular, the GCM aims to enhance the quality of parsing features by leveraging global features from silhouettes, while the PCM aligns the dynamics of human parts between silhouette and parsing features using the…
-
2 Jul 2024 1 repository listedMulti-task dense scene understanding, which learns a model for multiple dense prediction tasks, has a wide range of application scenarios.
-
29 Apr 2024 1 repository listed Syntology ran 5 of 9 samples · 4 unverified · 9 pointer-only (licence)Addressing this gap, we introduce 4D-DRESS, the first real-world 4D dataset advancing human clothing research with its high-quality 4D textured scans and garment meshes.
-
31 Jan 2024 1 repository listedUnlike mainstream approaches using global features for simultaneous multi-task learning of ReID and human parsing, or relying on semantic information for attention guidance, DROP argues that the inferior performance of…
-
4 Jan 2024 1 repository listed Syntology ran 7 of 10 samples · 3 unverified · 10 pointer-only (licence)Multimodal-based action recognition methods have achieved high success using pose and RGB modality.
-
15 Dec 2023 1 repository listedIn addition, existing occluded person ReID benchmarks utilize occluded samples as queries, which will amplify the role of alleviating occlusion interference and underestimate the impact of the feature absence issue.
-
13 Oct 2023 1 repository listedMulti-human parsing is an image segmentation task necessitating both instance-level and fine-grained category-level information.
-
31 Aug 2023 1 repository listedFurthermore, due to the lack of suitable datasets, we build the first parsing-based dataset for gait recognition in the wild, named Gait3D-Parsing, by extending the large-scale and challenging Gait3D dataset.
-
26 Aug 2023 1 repository listedAdditionally, we propose Virtual Try-on-guided Pose for Data Synthesis to address the limited pose variation observed in training images.
-
16 Jul 2023 1 repository listedWe propose an Integrating Human Parsing and Pose Network (IPP-Net) for action recognition, which is the first to leverage both skeletons and human parsing feature maps in dual-branch approach.
-
22 Apr 2023 1 repository listedWe instead present a high-performance Single-stage Multi-human Parsing (SMP) deep architecture that decouples the multi-human parsing problem into two fine-grained sub-problems, i.
-
10 Mar 2023 1 repository listed Syntology ran 15 of 25 samples · 10 unverifiedSpecifically, we propose a \textbf{HumanBench} based on existing datasets to comprehensively evaluate on the common ground the generalization abilities of different pretraining methods on 19 datasets from 6 diverse…
Syntology lines on 10 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections