{"url":"/method/rpn","slug":"rpn","name":"RPN","full_name":"Region Proposal Network","full_name_withheld":false,"description_markdown":"A **Region Proposal Network**, or **RPN**, is a fully convolutional network that simultaneously predicts object bounds and objectness scores at each position. The RPN is trained end-to-end to generate high-quality region proposals. RPN and algorithms like [Fast R-CNN](https://paperswithcode.com/method/fast-r-cnn) can be merged into a single network by sharing their convolutional features - using the recently popular terminology of neural networks with attention mechanisms, the RPN component tells the unified network where to look.\r\n\r\nRPNs are designed to efficiently predict region proposals with a wide range of scales and aspect ratios. RPNs use anchor boxes that serve as references at multiple scales and aspect ratios. The scheme can be thought of as a pyramid of regression references, which avoids enumerating images or filters of multiple scales or aspect ratios.","description_state":"present","introduced_year":null,"introduced_by":{"title":"Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks","paper":"/paper/faster-r-cnn-towards-real-time-object","first_author":"Shaoqing Ren","n_authors":4,"url_abs":null,"archive_paper_url":"https://paperswithcode.com/paper/faster-r-cnn-towards-real-time-object"},"source":{"url":"http://arxiv.org/abs/1506.01497v3","title":"Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks","url_on_a_paper_host":true},"code_snippet_url":null,"code_snippet_url_on_a_code_host":false,"categories":[{"area":"Computer Vision","area_id":"computer-vision","collection":"Region Proposal","url":"/methods/category/region-proposal","pwc_aliases":[]}],"n_papers_tagged":1045,"archive_num_papers":1045,"papers_newest_first":[{"paper":null,"title":"Pattern-Based Phase-Separation of Tracer and Dispersed Phase Particles in Two-Phase Defocusing Particle Tracking Velocimetry","date":"2025-06-22","arxiv_id":"2506.18157","n_code_links":0,"syntology":null},{"paper":null,"title":"Prmpt2Adpt: Prompt-Based Zero-Shot Domain Adaptation for Resource-Constrained Environments","date":"2025-06-20","arxiv_id":"2506.16994","n_code_links":0,"syntology":null},{"paper":null,"title":"A novel visual data-based diagnostic approach for estimation of regime transition in pool boiling","date":"2025-06-12","arxiv_id":"2506.10832","n_code_links":0,"syntology":null},{"paper":null,"title":"Bringing SAM to new heights: Leveraging elevation data for tree crown segmentation from drone imagery","date":"2025-06-05","arxiv_id":"2506.04970","n_code_links":0,"syntology":null},{"paper":null,"title":"Hierarchical Text Classification Using Contrastive Learning Informed Path Guided Hierarchy","date":"2025-06-04","arxiv_id":"2506.04381","n_code_links":0,"syntology":null},{"paper":"/paper/3d-gaussian-splat-vulnerabilities","title":"3D Gaussian Splat Vulnerabilities","date":"2025-05-30","arxiv_id":"2506.00280","n_code_links":1,"syntology":null},{"paper":null,"title":"Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting","date":"2025-05-22","arxiv_id":"2505.16513","n_code_links":0,"syntology":null},{"paper":null,"title":"AppleGrowthVision: A large-scale stereo dataset for phenological analysis, fruit detection, and 3D reconstruction in apple orchards","date":"2025-05-20","arxiv_id":"2505.14029","n_code_links":0,"syntology":null},{"paper":null,"title":"SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision","date":"2025-05-16","arxiv_id":"2505.11439","n_code_links":0,"syntology":null},{"paper":null,"title":"Object detection in adverse weather conditions for autonomous vehicles using Instruct Pix2Pix","date":"2025-05-13","arxiv_id":"2505.08228","n_code_links":0,"syntology":null},{"paper":"/paper/kg-htc-integrating-knowledge-graphs-into-llms","title":"KG-HTC: Integrating Knowledge Graphs into LLMs for Effective Zero-shot Hierarchical Text Classification","date":"2025-05-08","arxiv_id":"2505.05583","n_code_links":1,"syntology":null},{"paper":null,"title":"A Robust Deep Networks based Multi-Object MultiCamera Tracking System for City Scale Traffic","date":"2025-05-01","arxiv_id":"2505.00534","n_code_links":0,"syntology":null},{"paper":null,"title":"Transcending Dimensions using Generative AI: Real-Time 3D Model Generation in Augmented Reality","date":"2025-04-27","arxiv_id":"2504.21033","n_code_links":0,"syntology":null},{"paper":null,"title":"Real-time Seafloor Segmentation and Mapping","date":"2025-04-14","arxiv_id":"2504.10750","n_code_links":0,"syntology":null},{"paper":null,"title":"FMNV: A Dataset of Media-Published News Videos for Fake News Detection","date":"2025-04-10","arxiv_id":"2504.07687","n_code_links":0,"syntology":null},{"paper":null,"title":"RipVIS: Rip Currents Video Instance Segmentation Benchmark for Beach Monitoring and Safety","date":"2025-04-01","arxiv_id":"2504.01128","n_code_links":0,"syntology":null},{"paper":"/paper/ai-assisted-colonoscopy-polyp-detection-and","title":"AI-Assisted Colonoscopy: Polyp Detection and Segmentation using Foundation Models","date":"2025-03-31","arxiv_id":"2503.24138","n_code_links":1,"syntology":null},{"paper":null,"title":"BBoxCut: A Targeted Data Augmentation Technique for Enhancing Wheat Head Detection Under Occlusions","date":"2025-03-31","arxiv_id":"2503.24032","n_code_links":0,"syntology":null},{"paper":"/paper/a-gan-enhanced-deep-learning-framework-for","title":"A GAN-Enhanced Deep Learning Framework for Rooftop Detection from Historical Aerial Imagery","date":"2025-03-29","arxiv_id":"2503.23200","n_code_links":1,"syntology":null},{"paper":null,"title":"Autonomous AI for Multi-Pathology Detection in Chest X-Rays: A Multi-Site Study in the Indian Healthcare System","date":"2025-03-28","arxiv_id":"2504.00022","n_code_links":0,"syntology":null},{"paper":null,"title":"Assessing SAM for Tree Crown Instance Segmentation from Drone Imagery","date":"2025-03-26","arxiv_id":"2503.20199","n_code_links":0,"syntology":null},{"paper":null,"title":"Exploring Few-Shot Object Detection on Blood Smear Images: A Case Study of Leukocytes and Schistocytes","date":"2025-03-21","arxiv_id":"2503.17107","n_code_links":0,"syntology":null},{"paper":null,"title":"YOLO-LLTS: Real-Time Low-Light Traffic Sign Detection via Prior-Guided Enhancement and Multi-Branch Feature Interaction","date":"2025-03-18","arxiv_id":"2503.13883","n_code_links":0,"syntology":null},{"paper":null,"title":"Securing Virtual Reality Experiences: Unveiling and Tackling Cybersickness Attacks with Explainable AI","date":"2025-03-17","arxiv_id":"2503.13419","n_code_links":0,"syntology":null},{"paper":"/paper/overlock-an-overview-first-look-closely-next","title":"OverLoCK: An Overview-first-Look-Closely-next ConvNet with Context-Mixing Dynamic Kernels","date":"2025-02-27","arxiv_id":"2502.20087","n_code_links":1,"syntology":{"ran":0,"of":3,"unverified":3,"pointer_only":0}},{"paper":"/paper/walnutdata-a-uav-remote-sensing-dataset-of","title":"WalnutData: A UAV Remote Sensing Dataset of Green Walnuts and Model Evaluation","date":"2025-02-27","arxiv_id":"2502.20092","n_code_links":1,"syntology":null},{"paper":null,"title":"Automatic Vehicle Detection using DETR: A Transformer-Based Approach for Navigating Treacherous Roads","date":"2025-02-25","arxiv_id":"2502.17843","n_code_links":0,"syntology":null},{"paper":null,"title":"Autonomous Vision-Guided Resection of Central Airway Obstruction","date":"2025-02-25","arxiv_id":"2502.18586","n_code_links":0,"syntology":null},{"paper":null,"title":"Hybrid Answer Set Programming: Foundations and Applications","date":"2025-02-13","arxiv_id":"2502.09235","n_code_links":0,"syntology":null},{"paper":"/paper/sasvi-segment-any-surgical-video","title":"SASVi - Segment Any Surgical Video","date":"2025-02-12","arxiv_id":"2502.09653","n_code_links":1,"syntology":null}],"papers_shown":30,"tasks":[{"task":"/task/object-detection","name":"Object Detection","papers":501},{"task":"/task/object-detection-1","name":"object-detection","papers":458},{"task":"/task/object","name":"Object","papers":282},{"task":"/task/semantic-segmentation","name":"Semantic Segmentation","papers":235},{"task":"/task/instance-segmentation","name":"Instance Segmentation","papers":201},{"task":"/task/segmentation","name":"Segmentation","papers":160},{"task":"/task/region-proposal","name":"Region Proposal","papers":91},{"task":"/task/image-classification","name":"Image Classification","papers":56},{"task":"/task/transfer-learning","name":"Transfer Learning","papers":42},{"task":"/task/classification","name":"General Classification","papers":39},{"task":"/task/data-augmentation","name":"Data Augmentation","papers":37},{"task":"/task/image-classification","name":"image-classification","papers":36},{"task":"/task/classification-1","name":"Classification","papers":31},{"task":"/task/deep-learning","name":"Deep Learning","papers":31},{"task":"/task/regression-1","name":"regression","papers":29},{"task":"/task/autonomous-driving","name":"Autonomous Driving","papers":27},{"task":null,"name":"GPU","papers":26},{"task":"/task/decoder","name":"Decoder","papers":24},{"task":"/task/pose-estimation","name":"Pose Estimation","papers":24},{"task":"/task/object-recognition","name":"Object Recognition","papers":22}],"tasks_shown":20,"n_tasks":505,"usage_by_year":[{"year":"2015","papers":5},{"year":"2016","papers":13},{"year":"2017","papers":41},{"year":"2018","papers":98},{"year":"2019","papers":152},{"year":"2020","papers":177},{"year":"2021","papers":189},{"year":"2022","papers":124},{"year":"2023","papers":107},{"year":"2024","papers":98},{"year":"2025","papers":41}],"row_source":"methods_table","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/rpn"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}