{"url":"/task/keypoint-detection","name":"Keypoint Detection","slug":"keypoint-detection","description_markdown":"**Keypoint Detection** is essential for analyzing and interpreting images in computer vision. It involves simultaneously detecting and localizing interesting points in an image. Keypoints, also known as interest points, are spatial locations or points in the image that define what is interesting or what stands out. They are invariant to image rotation, shrinkage, translation, distortion, etc. Keypoints examples are body joints, facial landmarks, or any other salient points in objects. Keypoints have uses in problems such as pose estimation, object detection and tracking, facial analysis, and augmented reality.\r\n\r\n<span style=\"color:grey; opacity: 0.6\">( Image credit: [PifPaf: Composite Fields for Human Pose Estimation](https://github.com/vita-epfl/openpifpaf); \"Learning to surf\" by fotologic, license: CC-BY-2.0 )</span>","categories":[{"name":"Computer Vision","url":"/area/computer-vision"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"archive_url"},"counts":{"papers_tagged":339,"papers_with_code":180,"benchmarks":9,"benchmark_tables_in_archive":9,"benchmark_tables_shown":9,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":13,"subtasks":0,"parent_tasks":1},"benchmarks":[{"leaderboard":"/sota/keypoint-detection-on-coco","slug":"keypoint-detection-on-coco","dataset":"COCO (Common Objects in Context)","dataset_url":"/dataset/coco","rows_in_archive":24,"metrics":["Test AP","Validation AP","FPS"],"first_row_in_archive_order":{"model":"4xRSN-50(384×288)","paper_title":"Learning Delicate Local Representations for Multi-Person Pose Estimation","paper_url":"/paper/learning-delicate-local-representations-for","paper_date":"2020-03-09","arxiv_id":"2003.04030","code_links":[{"title":"open-mmlab/mmpose","url":"https://github.com/open-mmlab/mmpose"},{"title":"chenyilun95/tf-cpn","url":"https://github.com/chenyilun95/tf-cpn"},{"title":"caiyuanhao1998/RSN","url":"https://github.com/caiyuanhao1998/RSN"},{"title":"HuangJunJie2017/UDP-Pose","url":"https://github.com/HuangJunJie2017/UDP-Pose"}],"syntology":{"n":3,"n_ran":2,"n_unverified":1,"n_pointer_only":0}}},{"leaderboard":"/sota/keypoint-detection-on-coco-test-dev","slug":"keypoint-detection-on-coco-test-dev","dataset":"COCO test-dev","dataset_url":"/dataset/coco","rows_in_archive":16,"metrics":["APM","APL","AP50","AP75","AR","AR50","AR75","ARM","ARL","AP"],"first_row_in_archive_order":{"model":"HRNet*","paper_title":"Deep High-Resolution Representation Learning for Human Pose Estimation","paper_url":"/paper/deep-high-resolution-representation-learning","paper_date":"2019-02-25","arxiv_id":"1902.09212","code_links":[{"title":"open-mmlab/mmdetection","url":"https://github.com/open-mmlab/mmdetection"},{"title":"PaddlePaddle/PaddleDetection","url":"https://github.com/PaddlePaddle/PaddleDetection"},{"title":"open-mmlab/mmpose","url":"https://github.com/open-mmlab/mmpose"},{"title":"leoxiaobin/deep-high-resolution-net.pytorch","url":"https://github.com/leoxiaobin/deep-high-resolution-net.pytorch"},{"title":"HRNet/HRNet-Semantic-Segmentation","url":"https://github.com/HRNet/HRNet-Semantic-Segmentation"},{"title":"osmr/imgclsmob","url":"https://github.com/osmr/imgclsmob"},{"title":"Microsoft/human-pose-estimation.pytorch","url":"https://github.com/Microsoft/human-pose-estimation.pytorch"},{"title":"HRNet/HRNet-Facial-Landmark-Detection","url":"https://github.com/HRNet/HRNet-Facial-Landmark-Detection"},{"title":"HRNet/HRNet-Image-Classification","url":"https://github.com/HRNet/HRNet-Image-Classification"},{"title":"HRNet/HRNet-Object-Detection","url":"https://github.com/HRNet/HRNet-Object-Detection"},{"title":"mindspore-lab/mindone","url":"https://github.com/mindspore-lab/mindone"},{"title":"leeyegy/SimDR","url":"https://github.com/leeyegy/SimDR"},{"title":"leeyegy/simcc","url":"https://github.com/leeyegy/simcc"},{"title":"mks0601/PoseFix_RELEASE","url":"https://github.com/mks0601/PoseFix_RELEASE"},{"title":"HRNet/HRNet-Human-Pose-Estimation","url":"https://github.com/HRNet/HRNet-Human-Pose-Estimation"},{"title":"strivebo/image_segmentation_dl","url":"https://github.com/strivebo/image_segmentation_dl"},{"title":"HRNet/HRNet-MaskRCNN-Benchmark","url":"https://github.com/HRNet/HRNet-MaskRCNN-Benchmark"},{"title":"CASIA-IVA-Lab/ISP-reID","url":"https://github.com/CASIA-IVA-Lab/ISP-reID"},{"title":"NVlabs/PAMTRI","url":"https://github.com/NVlabs/PAMTRI"},{"title":"k-miran/hear","url":"https://github.com/k-miran/hear"},{"title":"Vill-Lab/2022-TIP-HCGA","url":"https://github.com/Vill-Lab/2022-TIP-HCGA"},{"title":"d-shivam/Pose-estimation-based-action-recognition-for-help-Situation-Identification","url":"https://github.com/d-shivam/Pose-estimation-based-action-recognition-for-help-Situation-Identification"},{"title":"chuanqichen/deepcoaching","url":"https://github.com/chuanqichen/deepcoaching"},{"title":"Mary-xl/HRnet_Kaggle_iNat2019_FGVC","url":"https://github.com/Mary-xl/HRnet_Kaggle_iNat2019_FGVC"},{"title":"v1viswan/Domain_adaptation_in_HRNet","url":"https://github.com/v1viswan/Domain_adaptation_in_HRNet"},{"title":"ken724049/action-recognition","url":"https://github.com/ken724049/action-recognition"},{"title":"NU-LL/lighttrack-","url":"https://github.com/NU-LL/lighttrack-"},{"title":"thoughtmachines/Human-Pose-Estimation-using-HRNets","url":"https://github.com/thoughtmachines/Human-Pose-Estimation-using-HRNets"},{"title":"ducongju/HRNet","url":"https://github.com/ducongju/HRNet"},{"title":"thomasslloyd/FitSpatial","url":"https://github.com/thomasslloyd/FitSpatial"},{"title":"laowang666888/HRNET","url":"https://github.com/laowang666888/HRNET"},{"title":"baoshengyu/deep-high-resolution-net.pytorch","url":"https://github.com/baoshengyu/deep-high-resolution-net.pytorch"},{"title":"sdll/hrnet-pose-estimation","url":"https://github.com/sdll/hrnet-pose-estimation"},{"title":"gox-ai/hrnet-pose-api","url":"https://github.com/gox-ai/hrnet-pose-api"},{"title":"anshky/HR-NET","url":"https://github.com/anshky/HR-NET"},{"title":"wsjzha/deep-high-resolution-net.pytorch","url":"https://github.com/wsjzha/deep-high-resolution-net.pytorch"},{"title":"visionNoob/hrnet_pytorch","url":"https://github.com/visionNoob/hrnet_pytorch"},{"title":"abhi1kumar/hrnet_pose_single_gpu","url":"https://github.com/abhi1kumar/hrnet_pose_single_gpu"},{"title":"goutern/PoseEstimation","url":"https://github.com/goutern/PoseEstimation"}],"syntology":{"n":25,"n_ran":8,"n_unverified":17,"n_pointer_only":0}}},{"leaderboard":"/sota/keypoint-detection-on-ochuman","slug":"keypoint-detection-on-ochuman","dataset":"OCHuman","dataset_url":"/dataset/ochuman","rows_in_archive":10,"metrics":["Test AP","Validation AP"],"first_row_in_archive_order":{"model":"BBox-Mask-Pose 2x","paper_title":"Detection, Pose Estimation and Segmentation for Multiple Bodies: Closing the Virtuous Circle","paper_url":"/paper/detection-pose-estimation-and-segmentation-1","paper_date":"2024-12-02","arxiv_id":"2412.01562","code_links":[{"title":"MiraPurkrabek/BBoxMaskPose","url":"https://github.com/MiraPurkrabek/BBoxMaskPose"}],"syntology":null}},{"leaderboard":"/sota/keypoint-detection-on-mpii-multi-person","slug":"keypoint-detection-on-mpii-multi-person","dataset":"MPII Multi-Person","dataset_url":"/dataset/mpii","rows_in_archive":9,"metrics":["mAP@0.5"],"first_row_in_archive_order":{"model":"AlphaPose","paper_title":"RMPE: Regional Multi-person Pose Estimation","paper_url":"/paper/rmpe-regional-multi-person-pose-estimation","paper_date":"2016-12-01","arxiv_id":"1612.00137","code_links":[{"title":"MVIG-SJTU/AlphaPose","url":"https://github.com/MVIG-SJTU/AlphaPose"},{"title":"osmr/imgclsmob","url":"https://github.com/osmr/imgclsmob"},{"title":"mindspore-ai/models","url":"https://github.com/mindspore-ai/models/tree/master/research/cv/AlphaPose"},{"title":"MVIG-SJTU/RMPE","url":"https://github.com/MVIG-SJTU/RMPE"},{"title":"ManifoldFR/recvis-project","url":"https://github.com/ManifoldFR/recvis-project"},{"title":"Fangyh09/pose_nms","url":"https://github.com/Fangyh09/pose_nms"},{"title":"MattyChoi/PoseMachines","url":"https://github.com/MattyChoi/PoseMachines"},{"title":"2023-MindSpore-4/Code8","url":"https://github.com/2023-MindSpore-4/Code8/tree/main/AlphaPose"},{"title":"yangyucheng000/AlphaPose","url":"https://github.com/yangyucheng000/AlphaPose"},{"title":"MindSpore-paper-code-3/code1","url":"https://github.com/MindSpore-paper-code-3/code1/tree/main/AlphaPose"},{"title":"2023-MindSpore-1/ms-code-199","url":"https://github.com/2023-MindSpore-1/ms-code-199"},{"title":"2023-MindSpore-1/ms-code-22","url":"https://github.com/2023-MindSpore-1/ms-code-22"},{"title":"lyqcom/alphapose","url":"https://github.com/lyqcom/alphapose"},{"title":"yuanyuanfyy/yycode","url":"https://github.com/yuanyuanfyy/yycode/tree/mindsporecode/AlphaPose"}],"syntology":null}},{"leaderboard":"/sota/keypoint-detection-on-vicos-towel-dataset","slug":"keypoint-detection-on-vicos-towel-dataset","dataset":"ViCoS Towel Dataset","dataset_url":"/dataset/vicos-towel-dataset","rows_in_archive":9,"metrics":["Best F1"],"first_row_in_archive_order":{"model":"CeDiRNet-3DoF - RGB-D (ConvNext-B)","paper_title":"Center Direction Network for Grasping Point Localization on Cloths","paper_url":"/paper/center-direction-network-for-grasping-point","paper_date":"2024-08-26","arxiv_id":"2408.14456","code_links":[{"title":"vicoslab/cedirnet-3dof","url":"https://github.com/vicoslab/cedirnet-3dof"}],"syntology":null}},{"leaderboard":"/sota/keypoint-detection-on-coco-test-challenge","slug":"keypoint-detection-on-coco-test-challenge","dataset":"COCO test-challenge","dataset_url":"/dataset/coco","rows_in_archive":8,"metrics":["AR","ARM","AP","AP50","AP75","APL","AR50","AR75","ARL"],"first_row_in_archive_order":{"model":"4×RSN-50","paper_title":"Learning Delicate Local Representations for Multi-Person Pose Estimation","paper_url":"/paper/learning-delicate-local-representations-for","paper_date":"2020-03-09","arxiv_id":"2003.04030","code_links":[{"title":"open-mmlab/mmpose","url":"https://github.com/open-mmlab/mmpose"},{"title":"chenyilun95/tf-cpn","url":"https://github.com/chenyilun95/tf-cpn"},{"title":"caiyuanhao1998/RSN","url":"https://github.com/caiyuanhao1998/RSN"},{"title":"HuangJunJie2017/UDP-Pose","url":"https://github.com/HuangJunJie2017/UDP-Pose"}],"syntology":{"n":3,"n_ran":2,"n_unverified":1,"n_pointer_only":0}}},{"leaderboard":"/sota/keypoint-detection-on-pascal3d","slug":"keypoint-detection-on-pascal3d","dataset":"Pascal3D+","dataset_url":"/dataset/pascal3d-2","rows_in_archive":4,"metrics":["Mean PCK"],"first_row_in_archive_order":{"model":"ConvNet + deformable shape model","paper_title":"6-DoF Object Pose from Semantic Keypoints","paper_url":"/paper/6-dof-object-pose-from-semantic-keypoints","paper_date":"2017-03-14","arxiv_id":"1703.04670","code_links":[{"title":"geopavlakos/object3d","url":"https://github.com/geopavlakos/object3d"}],"syntology":null}},{"leaderboard":"/sota/keypoint-detection-on-coco-1","slug":"keypoint-detection-on-coco-1","dataset":"COCO","dataset_url":null,"rows_in_archive":2,"metrics":["Validation AP","Test AP"],"first_row_in_archive_order":{"model":"Mask R-CNN","paper_title":"Mask R-CNN","paper_url":"/paper/mask-r-cnn","paper_date":"2017-03-20","arxiv_id":"1703.06870","code_links":[{"title":"tensorflow/models","url":"https://github.com/tensorflow/models/tree/master/official/vision"},{"title":"facebookresearch/detectron2","url":"https://github.com/facebookresearch/detectron2"},{"title":"facebookresearch/detectron","url":"https://github.com/facebookresearch/detectron"},{"title":"PaddlePaddle/PaddleDetection","url":"https://github.com/PaddlePaddle/PaddleDetection"},{"title":"facebookresearch/maskrcnn-benchmark","url":"https://github.com/facebookresearch/maskrcnn-benchmark"},{"title":"tensorpack/tensorpack","url":"https://github.com/tensorpack/tensorpack/tree/master/examples/FasterRCNN"},{"title":"charlesshang/fastmaskrcnn","url":"https://github.com/charlesshang/fastmaskrcnn"},{"title":"ZQPei/deep_sort_pytorch","url":"https://github.com/ZQPei/deep_sort_pytorch"},{"title":"tryolabs/luminoth","url":"https://github.com/tryolabs/luminoth"},{"title":"multimodallearning/pytorch-mask-rcnn","url":"https://github.com/multimodallearning/pytorch-mask-rcnn"},{"title":"TuSimple/mx-maskrcnn","url":"https://github.com/TuSimple/mx-maskrcnn"},{"title":"mic-dkfz/medicaldetectiontoolkit","url":"https://github.com/mic-dkfz/medicaldetectiontoolkit"},{"title":"open-edge-platform/training_extensions","url":"https://github.com/open-edge-platform/training_extensions"},{"title":"ayoolaolafenwa/PixelLib","url":"https://github.com/ayoolaolafenwa/PixelLib"},{"title":"vimalabs/VIMA","url":"https://github.com/vimalabs/VIMA"},{"title":"ykasten/layered-neural-atlases","url":"https://github.com/ykasten/layered-neural-atlases"},{"title":"longcw/roialign.pytorch","url":"https://github.com/longcw/roialign.pytorch"},{"title":"jremillard/images-to-osm","url":"https://github.com/jremillard/images-to-osm"},{"title":"open-edge-platform/geti","url":"https://github.com/open-edge-platform/geti"},{"title":"bethgelab/siamese-mask-rcnn","url":"https://github.com/bethgelab/siamese-mask-rcnn"},{"title":"Okery/PyTorch-Simple-MaskRCNN","url":"https://github.com/Okery/PyTorch-Simple-MaskRCNN"},{"title":"yubaoliu/rds-slam","url":"https://github.com/yubaoliu/rds-slam"},{"title":"crowdai/crowdai-mapping-challenge-mask-rcnn","url":"https://github.com/crowdai/crowdai-mapping-challenge-mask-rcnn"},{"title":"raymon-tian/hourglass-facekeypoints-detection","url":"https://github.com/raymon-tian/hourglass-facekeypoints-detection"},{"title":"mstfakts/building-detection-maskrcnn","url":"https://github.com/mstfakts/building-detection-maskrcnn"},{"title":"jasjeetIM/Mask-RCNN","url":"https://github.com/jasjeetIM/Mask-RCNN"},{"title":"mindspore-ai/models","url":"https://github.com/mindspore-ai/models/tree/master/official/cv/maskrcnn"},{"title":"SUYEgit/Surgery-Robot-Detection-Segmentation","url":"https://github.com/SUYEgit/Surgery-Robot-Detection-Segmentation"},{"title":"KMnP/fashionpedia-api","url":"https://github.com/KMnP/fashionpedia-api"},{"title":"mirzaevinom/data_science_bowl_2018","url":"https://github.com/mirzaevinom/data_science_bowl_2018"},{"title":"DeNA/Chainer_Mask_R-CNN","url":"https://github.com/DeNA/Chainer_Mask_R-CNN"},{"title":"jianing-sun/Mask-YOLO","url":"https://github.com/jianing-sun/Mask-YOLO"},{"title":"RituYadav92/Radar-RGB-Attentive-Multimodal-Object-Detection","url":"https://github.com/RituYadav92/Radar-RGB-Attentive-Multimodal-Object-Detection"},{"title":"louisyuzhe/car-damage-detector","url":"https://github.com/louisyuzhe/car-damage-detector"},{"title":"AKASH2907/bird-species-classification","url":"https://github.com/AKASH2907/bird-species-classification"},{"title":"AKASH2907/bird_species_classification","url":"https://github.com/AKASH2907/bird_species_classification"},{"title":"NVlabs/industreallib","url":"https://github.com/NVlabs/industreallib"},{"title":"Lopezurrutia/DSB_2018","url":"https://github.com/Lopezurrutia/DSB_2018"},{"title":"lichengunc/mask-faster-rcnn","url":"https://github.com/lichengunc/mask-faster-rcnn"},{"title":"Burf/tfdetection","url":"https://github.com/Burf/tfdetection"},{"title":"EmGarr/kerod","url":"https://github.com/EmGarr/kerod"},{"title":"itsasimiqbal/SeBRe","url":"https://github.com/itsasimiqbal/SeBRe"},{"title":"BupyeongHealer/Mask_RCNN_tf_2.x","url":"https://github.com/BupyeongHealer/Mask_RCNN_tf_2.x"},{"title":"MIC-DKFZ/RegRCNN","url":"https://github.com/MIC-DKFZ/RegRCNN"},{"title":"MIC-DKFZ/DetectionAndRegression","url":"https://github.com/MIC-DKFZ/DetectionAndRegression"},{"title":"yczhang1017/SSD_resnet_pytorch","url":"https://github.com/yczhang1017/SSD_resnet_pytorch"},{"title":"waspinator/deep-learning-explorer","url":"https://github.com/waspinator/deep-learning-explorer"},{"title":"alexander-pv/maskrcnn_tf2","url":"https://github.com/alexander-pv/maskrcnn_tf2"},{"title":"maxfrei750/DeepParticleNet","url":"https://github.com/maxfrei750/DeepParticleNet"},{"title":"dvl-tum/motsynth-baselines","url":"https://github.com/dvl-tum/motsynth-baselines"},{"title":"bowu1004/instance_segmentation_RealSense","url":"https://github.com/bowu1004/instance_segmentation_RealSense"},{"title":"ssanyachetwani/Fashion-Trend-Detection-and-Recommendation-Model","url":"https://github.com/ssanyachetwani/Fashion-Trend-Detection-and-Recommendation-Model"},{"title":"pvdhove/owl-mask-rcnn","url":"https://github.com/pvdhove/owl-mask-rcnn"},{"title":"jylins/core-text","url":"https://github.com/jylins/core-text"},{"title":"SonginCV/MAF_HDA","url":"https://github.com/SonginCV/MAF_HDA"},{"title":"SonginCV/GMPHD_MAF","url":"https://github.com/SonginCV/GMPHD_MAF"},{"title":"ErikGDev/instance-segmentation","url":"https://github.com/ErikGDev/instance-segmentation"},{"title":"SonginCV/GMPHD_SAF","url":"https://github.com/SonginCV/GMPHD_SAF"},{"title":"deolipankaj/Stone_Detection_MRCNN","url":"https://github.com/deolipankaj/Stone_Detection_MRCNN"},{"title":"RaiyaniNirav/Mask-R-CNN-for-water-detection","url":"https://github.com/RaiyaniNirav/Mask-R-CNN-for-water-detection"},{"title":"CarstenIsert/DeepBurn","url":"https://github.com/CarstenIsert/DeepBurn"},{"title":"lukoucky/image_recommendation","url":"https://github.com/lukoucky/image_recommendation"},{"title":"lucylow/salty-wet-man","url":"https://github.com/lucylow/salty-wet-man"},{"title":"collectionslab/book-annotation-classification","url":"https://github.com/collectionslab/book-annotation-classification"},{"title":"Alexander-Whelan/Zeus","url":"https://github.com/Alexander-Whelan/Zeus"},{"title":"collectionslab/annotations-computervision","url":"https://github.com/collectionslab/annotations-computervision"},{"title":"collectionslab/Omniscribe","url":"https://github.com/collectionslab/Omniscribe"},{"title":"SfTI-Robotics/ROS-label-node","url":"https://github.com/SfTI-Robotics/ROS-label-node"},{"title":"houssemjebari/Fruit-Detection","url":"https://github.com/houssemjebari/Fruit-Detection"},{"title":"polospeter/TensorFlow-Advanced-Techniques-Specialization","url":"https://github.com/polospeter/TensorFlow-Advanced-Techniques-Specialization"},{"title":"NiravRaiyani/Mask-R-CNN-for-water-detection","url":"https://github.com/NiravRaiyani/Mask-R-CNN-for-water-detection"},{"title":"lixiaolei1982/DenseNet-base-Mask-RCNN-for-Human-Pose-Estimation","url":"https://github.com/lixiaolei1982/DenseNet-base-Mask-RCNN-for-Human-Pose-Estimation"},{"title":"o-evgeny/MRCNN_DeepFashion2","url":"https://github.com/o-evgeny/MRCNN_DeepFashion2"},{"title":"RajArPatra/Super-OCR","url":"https://github.com/RajArPatra/Super-OCR"},{"title":"xiuyu0000/papers_with_examples","url":"https://github.com/xiuyu0000/papers_with_examples/tree/main/maskrcnn"},{"title":"maxfrei750/FibeR-CNN","url":"https://github.com/maxfrei750/FibeR-CNN"},{"title":"scolocke/Traffic_Sign_ID_GTSRB_GTSDB","url":"https://github.com/scolocke/Traffic_Sign_ID_GTSRB_GTSDB"},{"title":"George-Ogden/Mask-RCNN","url":"https://github.com/George-Ogden/Mask-RCNN"},{"title":"delldu/MaskRCNN","url":"https://github.com/delldu/MaskRCNN"},{"title":"guanfuchen/py-faster-rcnn","url":"https://github.com/guanfuchen/py-faster-rcnn"},{"title":"pj1920/mask-r-cnn","url":"https://github.com/pj1920/mask-r-cnn"},{"title":"chuanqichen/deepcoaching","url":"https://github.com/chuanqichen/deepcoaching"},{"title":"Mrnoorsingh/car-parking","url":"https://github.com/Mrnoorsingh/car-parking"},{"title":"hellohaozheng/maskrcnn-mindspore","url":"https://github.com/hellohaozheng/maskrcnn-mindspore"},{"title":"noelcodes/Mask_RCNN","url":"https://github.com/noelcodes/Mask_RCNN"},{"title":"stanleycelestin1/AirsimDetectron","url":"https://github.com/stanleycelestin1/AirsimDetectron"},{"title":"fsafe/Capstone","url":"https://github.com/fsafe/Capstone"},{"title":"zhangchi9/Airbus_Ship_Detection","url":"https://github.com/zhangchi9/Airbus_Ship_Detection"},{"title":"chenwuperth/ClaRAN","url":"https://github.com/chenwuperth/ClaRAN"},{"title":"Rakeshvd/Semantic-Segmenation-of-MRI-scan-using-Mask-RCNN","url":"https://github.com/Rakeshvd/Semantic-Segmenation-of-MRI-scan-using-Mask-RCNN"},{"title":"Rakeshvd/Instance-Segmenation-of-MRI-scan-using-Mask-RCNN","url":"https://github.com/Rakeshvd/Instance-Segmenation-of-MRI-scan-using-Mask-RCNN"},{"title":"kdethoor/panoptictorch","url":"https://github.com/kdethoor/panoptictorch"},{"title":"quocdat32461997/Mask_RCNN","url":"https://github.com/quocdat32461997/Mask_RCNN"},{"title":"asyrovprog/cs230project","url":"https://github.com/asyrovprog/cs230project"},{"title":"StephenEkaputra/Mask_RCNN-TinyPascalVOC","url":"https://github.com/StephenEkaputra/Mask_RCNN-TinyPascalVOC"},{"title":"infini8-13/MaskRCNN_tryout","url":"https://github.com/infini8-13/MaskRCNN_tryout"},{"title":"code-implementation1/Code9","url":"https://github.com/code-implementation1/Code9/tree/main/MaskRCNN/maskrcnn_mobilenetv1"},{"title":"alexalm4190/Mask_RCNN-Vizzy_Hand","url":"https://github.com/alexalm4190/Mask_RCNN-Vizzy_Hand"},{"title":"jodumagpi/Xray-ObjSep-v1","url":"https://github.com/jodumagpi/Xray-ObjSep-v1"},{"title":"boom85423/Cartoon-style-Stickers-Generator","url":"https://github.com/boom85423/Cartoon-style-Stickers-Generator"},{"title":"jooyounghun/AI-Team-5","url":"https://github.com/jooyounghun/AI-Team-5"},{"title":"samsh19/ML_project","url":"https://github.com/samsh19/ML_project"},{"title":"2023-MindSpore-1/ms-code-217","url":"https://github.com/2023-MindSpore-1/ms-code-217/tree/main/tinydarknet"},{"title":"jiajunhua/facebookresearch-Detectron","url":"https://github.com/jiajunhua/facebookresearch-Detectron"},{"title":"2023-MindSpore-1/ms-code-7","url":"https://github.com/2023-MindSpore-1/ms-code-7/tree/main/tinydarknet"},{"title":"2023-MindSpore-4/Code10","url":"https://github.com/2023-MindSpore-4/Code10/tree/main/MaskRCNN"},{"title":"AshishSingh2261/Pedestrian_Instance_Segmentation","url":"https://github.com/AshishSingh2261/Pedestrian_Instance_Segmentation"},{"title":"yangyucheng000/Mask-RCNN","url":"https://github.com/yangyucheng000/Mask-RCNN"},{"title":"alililia/ascend_maskrcnn_mobilenetv1","url":"https://github.com/alililia/ascend_maskrcnn_mobilenetv1"},{"title":"casiopa/Madrid_Rooftops","url":"https://github.com/casiopa/Madrid_Rooftops"},{"title":"BingkAI-B21CAP0161/C-Mask-Machine-Learning","url":"https://github.com/BingkAI-B21CAP0161/C-Mask-Machine-Learning"},{"title":"miaohua1982/simple_fasterrcnn_pytorch","url":"https://github.com/miaohua1982/simple_fasterrcnn_pytorch"},{"title":"DivJAth/DeepLearning5922","url":"https://github.com/DivJAth/DeepLearning5922"},{"title":"fdac18/ForensicImages","url":"https://github.com/fdac18/ForensicImages"},{"title":"phykn/film-defect-detection","url":"https://github.com/phykn/film-defect-detection"},{"title":"kbardool/mrcnn3","url":"https://github.com/kbardool/mrcnn3"},{"title":"ravenwritingdesk/py-faster-rcnn-master","url":"https://github.com/ravenwritingdesk/py-faster-rcnn-master"},{"title":"ColdNoodler/py-faster-rcnn-cuda10","url":"https://github.com/ColdNoodler/py-faster-rcnn-cuda10"},{"title":"baodi23/hourglass-facekeypoints-detection","url":"https://github.com/baodi23/hourglass-facekeypoints-detection"},{"title":"phtruongan/py-faster-rcnn-docker","url":"https://github.com/phtruongan/py-faster-rcnn-docker"},{"title":"krantirk/py-faster-rcnn","url":"https://github.com/krantirk/py-faster-rcnn"},{"title":"rh01/faster-rcnn","url":"https://github.com/rh01/faster-rcnn"},{"title":"beassssry/U","url":"https://github.com/beassssry/U"},{"title":"sunqiangxtcsun/faster-rcnn","url":"https://github.com/sunqiangxtcsun/faster-rcnn"},{"title":"lincaiming/py-faster-rcnn-windows","url":"https://github.com/lincaiming/py-faster-rcnn-windows"},{"title":"2023-MindSpore-4/Code9","url":"https://github.com/2023-MindSpore-4/Code9/tree/main/MaskRCNN"},{"title":"rickyHong/py-faster-rcnn-repl","url":"https://github.com/rickyHong/py-faster-rcnn-repl"},{"title":"BlackAngel1111/Fast-RCNN","url":"https://github.com/BlackAngel1111/Fast-RCNN"},{"title":"rh01/fast-rnn","url":"https://github.com/rh01/fast-rnn"},{"title":"Arthur-Shi/py-faster-rcnn","url":"https://github.com/Arthur-Shi/py-faster-rcnn"},{"title":"cokowpublic/Nuclei-Segmentation","url":"https://github.com/cokowpublic/Nuclei-Segmentation"},{"title":"evaristr/py-faster_rcnn","url":"https://github.com/evaristr/py-faster_rcnn"},{"title":"MindSpore-paper-code-3/code9","url":"https://github.com/MindSpore-paper-code-3/code9/tree/main/MaskRCNN"},{"title":"rickyHong/py-faster-rcnn-repl-cudnn5-support","url":"https://github.com/rickyHong/py-faster-rcnn-repl-cudnn5-support"},{"title":"soulguy/Faster-rcnn","url":"https://github.com/soulguy/Faster-rcnn"},{"title":"charlesYangM/py-faster-rcnn-80.28","url":"https://github.com/charlesYangM/py-faster-rcnn-80.28"},{"title":"busyboxs/What-I-have-star","url":"https://github.com/busyboxs/What-I-have-star"},{"title":"lincaiming/py-faster-rcnn-update","url":"https://github.com/lincaiming/py-faster-rcnn-update"},{"title":"xzabg/faster-rcnn-with-Caltech","url":"https://github.com/xzabg/faster-rcnn-with-Caltech"},{"title":"zhong110020/py-faster-rcnn","url":"https://github.com/zhong110020/py-faster-rcnn"},{"title":"sbetageri/MaskRCNN","url":"https://github.com/sbetageri/MaskRCNN"},{"title":"leochangzliao/OPBM","url":"https://github.com/leochangzliao/OPBM"},{"title":"qq330488563/TEST","url":"https://github.com/qq330488563/TEST"},{"title":"xunhen/py-faster-rcnn-wjc","url":"https://github.com/xunhen/py-faster-rcnn-wjc"},{"title":"GH3927/Mask-RCNN-applied-to-cranes","url":"https://github.com/GH3927/Mask-RCNN-applied-to-cranes"},{"title":"busyboxs/faster_rcnn_voc","url":"https://github.com/busyboxs/faster_rcnn_voc"},{"title":"TejasBajania/Mtech_pro","url":"https://github.com/TejasBajania/Mtech_pro"},{"title":"godspeedcurry/lung-nodule-detection","url":"https://github.com/godspeedcurry/lung-nodule-detection"},{"title":"paradisetechsoftsolutions/Deep-learning-model-using-Mask-R-CNN-for-predicting-AC-remote","url":"https://github.com/paradisetechsoftsolutions/Deep-learning-model-using-Mask-R-CNN-for-predicting-AC-remote"},{"title":"cwbabel/faster-cnn","url":"https://github.com/cwbabel/faster-cnn"},{"title":"jhihan/rsna_pneumonia_detection","url":"https://github.com/jhihan/rsna_pneumonia_detection"},{"title":"MindSpore-paper-code-2/code3","url":"https://github.com/MindSpore-paper-code-2/code3/tree/main/tinydarknet"},{"title":"jklife3/maskrcnn-impl","url":"https://github.com/jklife3/maskrcnn-impl"},{"title":"chihyanghsu0805/object_detection_yolo","url":"https://github.com/chihyanghsu0805/object_detection_yolo"},{"title":"Krupal09/MaskRCNN-Demo","url":"https://github.com/Krupal09/MaskRCNN-Demo"},{"title":"nikhithakarennagari/Git","url":"https://github.com/nikhithakarennagari/Git"},{"title":"nikhithakarennagari/1311","url":"https://github.com/nikhithakarennagari/1311"},{"title":"devsoft123/fast-cnn","url":"https://github.com/devsoft123/fast-cnn"},{"title":"GAOwy123/py-faster-rcnn","url":"https://github.com/GAOwy123/py-faster-rcnn"},{"title":"yj-littlesky/py-faster-rcnn","url":"https://github.com/yj-littlesky/py-faster-rcnn"},{"title":"elonashatri/pitch_mask_rcnn","url":"https://github.com/elonashatri/pitch_mask_rcnn"},{"title":"Qinhj07/ATOMCode","url":"https://github.com/Qinhj07/ATOMCode"},{"title":"ls5122/mask-rcnn","url":"https://github.com/ls5122/mask-rcnn"},{"title":"sunhui1234/haha","url":"https://github.com/sunhui1234/haha"},{"title":"xjnpark/ds","url":"https://github.com/xjnpark/ds"},{"title":"Biantian/MscProject","url":"https://github.com/Biantian/MscProject"},{"title":"TejasBajania/Mtech_thesis_project","url":"https://github.com/TejasBajania/Mtech_thesis_project"},{"title":"saehan-choi/pixellib_auto_labelling","url":"https://github.com/saehan-choi/pixellib_auto_labelling"},{"title":"yangyucheng000/maskrcnn_mobilenetv1","url":"https://github.com/yangyucheng000/maskrcnn_mobilenetv1"},{"title":"alililia/gpu_maskrcnn_mobilenetv1","url":"https://github.com/alililia/gpu_maskrcnn_mobilenetv1"},{"title":"Makunda/DeepLearningASL","url":"https://github.com/Makunda/DeepLearningASL"},{"title":"muyistarsky/MaskRCNN","url":"https://github.com/muyistarsky/MaskRCNN"},{"title":"2023-MindSpore-1/ms-code-207","url":"https://github.com/2023-MindSpore-1/ms-code-207"},{"title":"2023-MindSpore-1/ms-code-208","url":"https://github.com/2023-MindSpore-1/ms-code-208"},{"title":"UPCLJ/py-faster-rcnn","url":"https://github.com/UPCLJ/py-faster-rcnn"},{"title":"qilei123/pyfasterrcnn","url":"https://github.com/qilei123/pyfasterrcnn"},{"title":"MindSpore-paper-code-3/code3","url":"https://github.com/MindSpore-paper-code-3/code3/tree/main/faster_rcnn_dcn"},{"title":"leonardhan1979/fasterRCNN","url":"https://github.com/leonardhan1979/fasterRCNN"},{"title":"harrybolingot/mymaskrcnn","url":"https://github.com/harrybolingot/mymaskrcnn"}],"syntology":{"n":140,"n_ran":42,"n_unverified":98,"n_pointer_only":23}}},{"leaderboard":"/sota/keypoint-detection-on-apollocar3d","slug":"keypoint-detection-on-apollocar3d","dataset":"ApolloCar3D","dataset_url":"/dataset/apollocar3d","rows_in_archive":1,"metrics":["A3DP"],"first_row_in_archive_order":{"model":"GSNet","paper_title":"GSNet: Joint Vehicle Pose and Shape Reconstruction with Geometrical and Scene-aware Supervision","paper_url":"/paper/gsnet-joint-vehicle-pose-and-shape","paper_date":"2020-07-26","arxiv_id":"2007.13124","code_links":[{"title":"lkeab/gsnet","url":"https://github.com/lkeab/gsnet"}],"syntology":null}}],"datasets":[{"url":"/dataset/coco","name":"COCO (Common Objects in Context)","full_name":"Common Objects in Context","num_papers_in_archive":11922},{"url":"/dataset/mpii","name":"MPII","full_name":"MPII Human Pose","num_papers_in_archive":495},{"url":"/dataset/pascal3d-2","name":"PASCAL3D+","full_name":"","num_papers_in_archive":237},{"url":"/dataset/ochuman","name":"OCHuman","full_name":null,"num_papers_in_archive":66},{"url":"/dataset/keypointnet","name":"KeypointNet","full_name":null,"num_papers_in_archive":24},{"url":"/dataset/apollocar3d","name":"ApolloCar3D","full_name":"","num_papers_in_archive":17},{"url":"/dataset/grit","name":"GRIT","full_name":"General Robust Image Task Benchmark","num_papers_in_archive":16},{"url":"/dataset/awa-pose","name":"AwA Pose","full_name":"","num_papers_in_archive":4},{"url":"/dataset/fish-keypoints-detection","name":"Fish Keypoints Detection","full_name":"","num_papers_in_archive":1},{"url":"/dataset/radiogalaxynet","name":"RadioGalaxyNET","full_name":"","num_papers_in_archive":1},{"url":"/dataset/radiogalaxynet-dataset","name":"RadioGalaxyNET Dataset","full_name":"","num_papers_in_archive":1},{"url":"/dataset/tampar","name":"TAMPAR","full_name":"","num_papers_in_archive":1},{"url":"/dataset/vicos-towel-dataset","name":"ViCoS Towel Dataset","full_name":"","num_papers_in_archive":1}],"subtasks":[],"parent_tasks":[{"url":"/task/pose-estimation","name":"Pose Estimation"}],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":30,"of":180,"tagged_in_all":339,"items":[{"url":"/paper/mask-r-cnn","title":"Mask R-CNN","date":"2017-03-20","arxiv_id":"1703.06870","repositories_listed":179,"syntology":{"n":140,"n_ran":42,"n_unverified":98,"n_pointer_only":23}},{"url":"/paper/objects-as-points","title":"Objects as Points","date":"2019-04-16","arxiv_id":"1904.07850","repositories_listed":76,"syntology":{"n":130,"n_ran":10,"n_unverified":120,"n_pointer_only":0}},{"url":"/paper/realtime-multi-person-2d-pose-estimation","title":"Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields","date":"2016-11-24","arxiv_id":"1611.08050","repositories_listed":61,"syntology":{"n":23,"n_ran":4,"n_unverified":19,"n_pointer_only":4}},{"url":"/paper/openpose-realtime-multi-person-2d-pose","title":"OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields","date":"2018-12-18","arxiv_id":"1812.08008","repositories_listed":51,"syntology":{"n":16,"n_ran":3,"n_unverified":13,"n_pointer_only":2}},{"url":"/paper/deep-high-resolution-representation-learning","title":"Deep High-Resolution Representation Learning for Human Pose Estimation","date":"2019-02-25","arxiv_id":"1902.09212","repositories_listed":39,"syntology":{"n":25,"n_ran":8,"n_unverified":17,"n_pointer_only":0}},{"url":"/paper/hand-keypoint-detection-in-single-images","title":"Hand Keypoint Detection in Single Images using Multiview Bootstrapping","date":"2017-04-25","arxiv_id":"1704.07809","repositories_listed":39,"syntology":null},{"url":"/paper/non-local-neural-networks","title":"Non-local Neural Networks","date":"2017-11-21","arxiv_id":"1711.07971","repositories_listed":32,"syntology":{"n":4,"n_ran":3,"n_unverified":1,"n_pointer_only":4}},{"url":"/paper/simple-baselines-for-human-pose-estimation","title":"Simple Baselines for Human Pose Estimation and Tracking","date":"2018-04-17","arxiv_id":"1804.06208","repositories_listed":27,"syntology":null},{"url":"/paper/deepercut-a-deeper-stronger-and-faster-multi","title":"DeeperCut: A Deeper, Stronger, and Faster Multi-Person Pose Estimation Model","date":"2016-05-10","arxiv_id":"1605.03170","repositories_listed":16,"syntology":null},{"url":"/paper/arttrack-articulated-multi-person-tracking-in","title":"ArtTrack: Articulated Multi-person Tracking in the Wild","date":"2016-12-05","arxiv_id":"1612.01465","repositories_listed":14,"syntology":{"n":29,"n_ran":1,"n_unverified":28,"n_pointer_only":2}},{"url":"/paper/rmpe-regional-multi-person-pose-estimation","title":"RMPE: Regional Multi-person Pose Estimation","date":"2016-12-01","arxiv_id":"1612.00137","repositories_listed":14,"syntology":null},{"url":"/paper/simple-pose-rethinking-and-improving-a-bottom","title":"Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose Estimation","date":"2019-11-24","arxiv_id":"1911.10529","repositories_listed":8,"syntology":null},{"url":"/paper/rethinking-on-multi-stage-networks-for-human","title":"Rethinking on Multi-Stage Networks for Human Pose Estimation","date":"2019-01-01","arxiv_id":"1901.00148","repositories_listed":7,"syntology":null},{"url":"/paper/pose2seg-detection-free-human-instance","title":"Pose2Seg: Detection Free Human Instance Segmentation","date":"2018-03-28","arxiv_id":"1803.10683","repositories_listed":7,"syntology":{"n":16,"n_ran":0,"n_unverified":16,"n_pointer_only":0}},{"url":"/paper/vitpose-simple-vision-transformer-baselines","title":"ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation","date":"2022-04-26","arxiv_id":"2204.12484","repositories_listed":6,"syntology":{"n":31,"n_ran":18,"n_unverified":13,"n_pointer_only":6}},{"url":"/paper/polarized-self-attention-towards-high-quality-1","title":"Polarized Self-Attention: Towards High-quality Pixel-wise Regression","date":"2021-07-02","arxiv_id":"2107.00782","repositories_listed":6,"syntology":null},{"url":"/paper/openpifpaf-composite-fields-for-semantic","title":"OpenPifPaf: Composite Fields for Semantic Keypoint Detection and Spatio-Temporal Association","date":"2021-03-03","arxiv_id":"2103.02440","repositories_listed":6,"syntology":null},{"url":"/paper/distribution-aware-coordinate-representation","title":"Distribution-Aware Coordinate Representation for Human Pose Estimation","date":"2019-10-14","arxiv_id":"1910.06278","repositories_listed":6,"syntology":{"n":9,"n_ran":5,"n_unverified":4,"n_pointer_only":0}},{"url":"/paper/rotate-to-attend-convolutional-triplet","title":"Rotate to Attend: Convolutional Triplet Attention Module","date":"2020-10-06","arxiv_id":"2010.03045","repositories_listed":5,"syntology":null},{"url":"/paper/dynamic-convolution-attention-over","title":"Dynamic Convolution: Attention over Convolution Kernels","date":"2019-12-07","arxiv_id":"1912.03458","repositories_listed":5,"syntology":{"n":4,"n_ran":1,"n_unverified":3,"n_pointer_only":0}},{"url":"/paper/cascaded-pyramid-network-for-multi-person","title":"Cascaded Pyramid Network for Multi-Person Pose Estimation","date":"2017-11-20","arxiv_id":"1711.07319","repositories_listed":5,"syntology":null},{"url":"/paper/associative-embedding-end-to-end-learning-for","title":"Associative Embedding: End-to-End Learning for Joint Detection and Grouping","date":"2016-11-16","arxiv_id":"1611.05424","repositories_listed":5,"syntology":{"n":8,"n_ran":0,"n_unverified":8,"n_pointer_only":0}},{"url":"/paper/regionvit-regional-to-local-attention-for","title":"RegionViT: Regional-to-Local Attention for Vision Transformers","date":"2021-06-04","arxiv_id":"2106.02689","repositories_listed":4,"syntology":{"n":25,"n_ran":13,"n_unverified":12,"n_pointer_only":0}},{"url":"/paper/learning-delicate-local-representations-for","title":"Learning Delicate Local Representations for Multi-Person Pose Estimation","date":"2020-03-09","arxiv_id":"2003.04030","repositories_listed":4,"syntology":{"n":3,"n_ran":2,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/slimmable-neural-networks","title":"Slimmable Neural Networks","date":"2018-12-21","arxiv_id":"1812.08928","repositories_listed":4,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":2}},{"url":"/paper/crowdpose-efficient-crowded-scenes-pose","title":"CrowdPose: Efficient Crowded Scenes Pose Estimation and A New Benchmark","date":"2018-12-02","arxiv_id":"1812.00324","repositories_listed":4,"syntology":null},{"url":"/paper/multiposenet-fast-multi-person-pose","title":"MultiPoseNet: Fast Multi-Person Pose Estimation using Pose Residual Network","date":"2018-07-11","arxiv_id":"1807.04067","repositories_listed":4,"syntology":null},{"url":"/paper/data-distillation-towards-omni-supervised","title":"Data Distillation: Towards Omni-Supervised Learning","date":"2017-12-12","arxiv_id":"1712.04440","repositories_listed":4,"syntology":null},{"url":"/paper/explicit-box-detection-unifies-end-to-end","title":"Explicit Box Detection Unifies End-to-End Multi-Person Pose Estimation","date":"2023-02-03","arxiv_id":"2302.01593","repositories_listed":3,"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":1}},{"url":"/paper/houghnet-integrating-near-and-long-range-1","title":"HoughNet: Integrating near and long-range evidence for visual detection","date":"2021-04-14","arxiv_id":"2104.06773","repositories_listed":3,"syntology":null}],"syntology_records":16,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}