{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/human-skeletons-and-change-detection-for","title":"Human skeletons and change detection for efficient violence detection in surveillance videos","arxiv_id":null,"date":"2023-05-20","proceeding":"Computer Vision and Image Understanding 2023 5","authors":["Guillermo Garcia-Cobo","Juan C. San Miguel"],"abstract":"In our constantly monitored world, surveillance cameras play a crucial role in curbing crime and violence in public spaces by serving as a deterrent. To enhance their effectiveness, there is a growing need for automated tools that can detect crimes in real time. In this paper, we propose a novel deep learning architecture that accurately and efficiently detects violent crimes in surveillance videos. We rely on what we believe are the most essential pieces of information to detect violence, namely: human bodies and their interaction. To this end, we employ human pose extractors and change detectors as the input of our proposal. Subsequently, we combine them using a novel method, which relies on additions instead of multiplications to guarantee the transmission of information even when one of the inputs provides a zero-valued signal; outperforming other combination alternatives of the literature. Finally, to account for both spatial and temporal information, we use a convolutional alternative of the standard LSTM, the ConvLSTM. The experiments performed on several benchmark datasets demonstrate the efficacy and efficiency of our proposal, achieving state-of-the-art results with much fewer trainable parameters. We release the code to replicate the proposed architecture at https://github.com/atmguille/Violence-Detection-With-Human-Skeletons","url_abs":"https://www.sciencedirect.com/science/article/pii/S1077314223001194","url_pdf":"https://www.sciencedirect.com/science/article/pii/S1077314223001194","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"human-skeletons-and-change-detection-for","repo_url":"https://github.com/atmguille/Violence-Detection-With-Human-Skeletons","is_official":0,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[{"task_slug":"activity-recognition","task_name":"Activity Recognition"},{"task_slug":"change-detection","task_name":"Change Detection"},{"task_slug":"violence-and-weaponized-violence-detection","task_name":"Violence and Weaponized Violence Detection"}],"methods":[{"method_slug":"convlstm","method_name":"ConvLSTM"},{"method_slug":"lstm","method_name":"LSTM"},{"method_slug":"openpose","method_name":"OpenPose"}],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/activity-recognition-on-rwf-2000","task":"Activity Recognition","dataset":"RWF-2000","model":"Human Skeletons + Change Detection","rank_in_archive_order":3,"of":6,"metrics":{"Accuracy":"90.25"},"uses_additional_data":false}],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}