Browse State-of-the-Art › 2D Pose Estimation
2D Pose Estimation
46 papers with code · 9 benchmarks · 14 datasets archive 2025-07-28
detective pose
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
9 leaderboard tables shown for this task, 9 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| iRodent (8 rows) | fine-tuned HRNetw32 pretrained on SuperAnimal (1 fac of data) | SuperAnimal pretrained pose estimation models for behavioral analysis | code | — | Compare |
| MP-100 (6 rows) | CapeLLM | CapeLLM: Support-Free Category-Agnostic Pose Estimation with... | — | — | Compare |
| 300W (1 row) | UniPose | X-Pose: Detecting Any Keypoints | code | Syntology ran 11 of 13 samples · 2 unverified | Compare |
| Animal Kingdom (1 row) | UniPose | X-Pose: Detecting Any Keypoints | code | Syntology ran 11 of 13 samples · 2 unverified | Compare |
| Desert Locust (1 row) | UniPose | X-Pose: Detecting Any Keypoints | code | Syntology ran 11 of 13 samples · 2 unverified | Compare |
| HARPER (1 row) | HRNet | Deep High-Resolution Representation Learning for Human Pose Estimation | code | Syntology ran 8 of 25 samples · 17 unverified | Compare |
| Human3.6M (1 row) | UniHCP (finetune) | UniHCP: A Unified Model for Human-Centric Perceptions | code | Syntology ran 7 of 13 samples · 6 unverified | Compare |
| MacaquePose (1 row) | UniPose | X-Pose: Detecting Any Keypoints | code | Syntology ran 11 of 13 samples · 2 unverified | Compare |
| Vinegar Fly (1 row) | UniPose | X-Pose: Detecting Any Keypoints | code | Syntology ran 11 of 13 samples · 2 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
14 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
2 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 46 papers with code (90 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
24 Nov 2016 61 repositories listed Syntology ran 4 of 23 samples · 19 unverified · 4 pointer-only (licence)We present an approach to efficiently detect the 2D pose of multiple people in an image.
-
18 Dec 2018 51 repositories listed Syntology ran 3 of 16 samples · 13 unverified · 2 pointer-only (licence)OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
-
25 Feb 2019 39 repositories listed Syntology ran 8 of 25 samples · 17 unverifiedWe start from a high-resolution subnetwork as the first stage, gradually add high-to-low resolution subnetworks one by one to form more stages, and connect the mutli-resolution subnetworks in parallel.
-
1 Dec 2023 35 repositories listed Syntology ran 18 of 62 samples · 44 unverified · 28 pointer-only (licence)Foundation models, now powering most of the exciting applications in deep learning, are almost universally based on the Transformer architecture and its core attention module.
-
2 Jul 2021 6 repositories listedPixel-wise regression is probably the most common problem in fine-grained computer vision tasks, such as estimating keypoint heatmaps and segmentation masks.
-
8 Apr 2017 6 repositories listedWe propose a weakly-supervised transfer learning method that uses mixed 2D and 3D labels in a unified deep neutral network that presents two-stage cascaded structure.
-
14 Mar 2022 3 repositories listedWe illustrate the utility of our models in behavioral classification in mice and gait analysis in horses.
-
22 Aug 2024 2 repositories listed Syntology ran 0 of 11 samples · 11 unverifiedWe present Sapiens, a family of models for four fundamental human-centric vision tasks -- 2D pose estimation, body-part segmentation, depth estimation, and surface normal prediction.
-
29 Nov 2023 2 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Traditional 2D pose estimation models are limited by their category-specific design, making them suitable only for predefined object categories.
-
12 Oct 2023 2 repositories listed Syntology ran 11 of 13 samples · 2 unverified · 13 pointer-only (licence)This work aims to address an advanced keypoint detection problem: how to accurately detect any keypoints in complex real-world scenarios, which involves massive, messy, and open-ended objects as well as their associated…
-
12 Oct 2022 2 repositories listed Syntology ran 0 of 17 samples · 17 unverifiedThe state-of-the-art for monocular 3D human pose estimation in videos is dominated by the paradigm of 2D-to-3D pose uplifting.
-
24 Dec 2020 2 repositories listed Syntology ran 1 of 3 samples · 2 unverified · 3 pointer-only (licence)We further show that this change in orientation can be used to impose an additional motion constraint in Siamese tracking through imposing restriction on the change in orientation between two consecutive frames.
-
20 Aug 2020 2 repositories listedComputer vision (CV) has achieved great success in interpreting semantic meanings from images, yet CV algorithms can be brittle for tasks with adverse vision conditions and the ones suffering from data/label pair…
-
2 Aug 2019 2 repositories listedHere we explore two variations of synthetic data for this challenging problem; a dataset with purely synthetic humans and a real dataset augmented with synthetic humans.
-
16 Jun 2025 1 repository listedThe paper presents a family of transformer-based models capable of performing multi-person 2D pose estimation in real-time.
-
25 Nov 2024 1 repository listedCategory-Agnostic Pose Estimation (CAPE) localizes keypoints across diverse object categories with a single model, using one or a few annotated support images.
-
20 Aug 2024 1 repository listedEstimating 3D human poses from 2D images is challenging due to occlusions and projective acquisition.
-
11 Jul 2024 1 repository listedIn this work, we present RTMW (Real-Time Multi-person Whole-body pose estimation models), a series of high-performance models for 2D/3D whole-body pose estimation.
-
1 Jun 2024 1 repository listedWe validate our novel approach using the MP-100 benchmark, a comprehensive dataset spanning over 100 categories and 18, 000 images.
-
9 Nov 2023 1 repository listedHuman silhouette extraction is a fundamental task in computer vision with applications in various downstream tasks.
-
31 Oct 2023 1 repository listed Syntology ran 3 of 6 samples · 3 unverified · 6 pointer-only (licence)To further capture human characteristics, we propose a structure-invariant alignment loss that enforces different masked views, guided by the human part prior, to be closely aligned for the same image.
-
13 Mar 2023 1 repository listedRecent studies on 2D pose estimation have achieved excellent performance on public benchmarks, yet its application in the industrial community still suffers from heavy model parameters and high latency.
-
6 Mar 2023 1 repository listed Syntology ran 7 of 13 samples · 6 unverifiedWhen adapted to a specific task, UniHCP achieves new SOTAs on a wide range of human-centric tasks, e.
-
1 Jan 2023 1 repository listedTo calibrate the inaccurate matching results, we introduce a two-stage framework, where matched keypoints from the first stage are viewed as similarity-aware position proposals.
-
18 Dec 2022 1 repository listedWe present a graph convolutional network with 2D pose estimation for the first time on child action recognition task achieving on par results with an RGB modality based model on a novel benchmark dataset containing…
-
29 Nov 2022 1 repository listedPrevious video-based human pose estimation methods have shown promising results by leveraging aggregated features of consecutive frames.
-
22 Oct 2022 1 repository listedIn addition to the benchmark, we propose a cross-modality training framework that leverages the ground-truth 2D keypoints representing human body joints for training, which are systematically generated from the…
-
21 Jul 2022 1 repository listed Syntology ran 0 of 3 samples · 3 unverifiedIn this paper, we introduce the task of Category-Agnostic Pose Estimation (CAPE), which aims to create a pose estimation model capable of detecting the pose of any class of object given only a few samples with keypoint…
-
1 Apr 2022 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedIn this paper, we investigate the problem of domain adaptive 2D pose estimation that transfers knowledge learned on a synthetic source domain to a target domain without supervision.
-
2 Nov 2021 1 repository listedIn this paper we introduce a novel method to estimate the head pose of people in single images starting from a small set of head keypoints.
Syntology lines on 13 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections