{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/panoptic-studio-a-massively-multiview-system","title":"Panoptic Studio: A Massively Multiview System for Social Interaction Capture","arxiv_id":"1612.03153","date":"2016-12-09","proceeding":null,"authors":["Hanbyul Joo","Tomas Simon","Xulong Li","Hao liu","Lei Tan","Lin Gui","Sean Banerjee","Timothy Godisart","Bart Nabbe","Iain Matthews","Takeo Kanade","Shohei Nobuhara","Yaser Sheikh"],"abstract":"We present an approach to capture the 3D motion of a group of people engaged\nin a social interaction. The core challenges in capturing social interactions\nare: (1) occlusion is functional and frequent; (2) subtle motion needs to be\nmeasured over a space large enough to host a social group; (3) human appearance\nand configuration variation is immense; and (4) attaching markers to the body\nmay prime the nature of interactions. The Panoptic Studio is a system organized\naround the thesis that social interactions should be measured through the\nintegration of perceptual analyses over a large variety of view points. We\npresent a modularized system designed around this principle, consisting of\nintegrated structural, hardware, and software innovations. The system takes, as\ninput, 480 synchronized video streams of multiple people engaged in social\nactivities, and produces, as output, the labeled time-varying 3D structure of\nanatomical landmarks on individuals in the space. Our algorithm is designed to\nfuse the \"weak\" perceptual processes in the large number of views by\nprogressively generating skeletal proposals from low-level appearance cues, and\na framework for temporal refinement is also presented by associating body parts\nto reconstructed dense 3D trajectory stream. Our system and method are the\nfirst in reconstructing full body motion of more than five people engaged in\nsocial interactions without using markers. We also empirically demonstrate the\nimpact of the number of views in achieving this goal.","url_abs":"http://arxiv.org/abs/1612.03153v1","url_pdf":"http://arxiv.org/pdf/1612.03153v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"panoptic-studio-a-massively-multiview-system","repo_url":"https://github.com/CMU-Perceptual-Computing-Lab/panoptic-toolbox","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":0,"framework":"none","reach":{"status":"ok"}},{"paper_slug":"panoptic-studio-a-massively-multiview-system","repo_url":"https://gitlab.com/Percipiote/skelda","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":null}],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=1612.03153","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}