skip to main content
research-article

Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

Published: 01 June 2017 Publication History

Abstract

State-of-the-art object detection networks depend on region proposal algorithms to hypothesize object locations. Advances like SPPnet [1] and Fast R-CNN [2] have reduced the running time of these detection networks, exposing region proposal computation as a bottleneck. In this work, we introduce a Region Proposal Network (RPN) that shares full-image convolutional features with the detection network, thus enabling nearly cost-free region proposals. An RPN is a fully convolutional network that simultaneously predicts object bounds and objectness scores at each position. The RPN is trained end-to-end to generate high-quality region proposals, which are used by Fast R-CNN for detection. We further merge RPN and Fast R-CNN into a single network by sharing their convolutional features—using the recently popular terminology of neural networks with ’attention’ mechanisms, the RPN component tells the unified network where to look. For the very deep VGG-16 model [3], our detection system has a frame rate of 5 fps ( including all steps ) on a GPU, while achieving state-of-the-art object detection accuracy on PASCAL VOC 2007, 2012, and MS COCO datasets with only 300 proposals per image. In ILSVRC and COCO 2015 competitions, Faster R-CNN and RPN are the foundations of the 1st-place winning entries in several tracks. Code has been made publicly available.

Cited By

View all
  • (2024)Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and PerceptionProceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems10.5555/3635637.3663065(2011-2019)Online publication date: 6-May-2024
  • (2024)YOLO-DCNetInternational Journal on Semantic Web & Information Systems10.4018/IJSWIS.33900020:1(1-23)Online publication date: 27-Feb-2024
  • (2024)Delivery Garbage Behavior Detection Based on Deep LearningInternational Journal of Information Technologies and Systems Approach10.4018/IJITSA.34363217:1(1-15)Online publication date: 31-Jan-2024
  • Show More Cited By

Recommendations

Comments

Information & Contributors

Information

Published In

cover image IEEE Transactions on Pattern Analysis and Machine Intelligence
IEEE Transactions on Pattern Analysis and Machine Intelligence  Volume 39, Issue 6
June 2017
240 pages

Publisher

IEEE Computer Society

United States

Publication History

Published: 01 June 2017

Qualifiers

  • Research-article

Contributors

Other Metrics

Bibliometrics & Citations

Bibliometrics

Article Metrics

  • Downloads (Last 12 months)0
  • Downloads (Last 6 weeks)0
Reflects downloads up to 24 Aug 2024

Other Metrics

Citations

Cited By

View all
  • (2024)Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and PerceptionProceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems10.5555/3635637.3663065(2011-2019)Online publication date: 6-May-2024
  • (2024)YOLO-DCNetInternational Journal on Semantic Web & Information Systems10.4018/IJSWIS.33900020:1(1-23)Online publication date: 27-Feb-2024
  • (2024)Delivery Garbage Behavior Detection Based on Deep LearningInternational Journal of Information Technologies and Systems Approach10.4018/IJITSA.34363217:1(1-15)Online publication date: 31-Jan-2024
  • (2024)Predicting diabetic macular edema in retina fundus images based on optimized deep residual network techniques on medical internet of thingsJournal of Intelligent & Fuzzy Systems: Applications in Engineering and Technology10.3233/JIFS-23464946:1(105-117)Online publication date: 1-Jan-2024
  • (2024)HDTNetJournal of Intelligent & Fuzzy Systems: Applications in Engineering and Technology10.3233/JIFS-23015046:1(1531-1541)Online publication date: 1-Jan-2024
  • (2024)Number detection of cylindrical objects based on improved Yolov5s algorithmIntelligent Decision Technologies10.3233/IDT-23054518:1(441-456)Online publication date: 1-Jan-2024
  • (2024)Cross-modality semantic guidance for multi-label image classificationIntelligent Data Analysis10.3233/IDA-23023928:3(633-646)Online publication date: 1-Jan-2024
  • (2024)Multi-layer features template update object tracking algorithm based on SiamFC++Journal on Image and Video Processing10.1186/s13640-023-00616-x2024:1Online publication date: 4-Jan-2024
  • (2024)Integrating Street Views, Satellite Imageries and Remote Sensing Data Into Economics and the Social SciencesSocial Science Computer Review10.1177/0894439323117860442:1(326-351)Online publication date: 1-Feb-2024
  • (2024)A cross-domain challenge with panoptic segmentation in agricultureInternational Journal of Robotics Research10.1177/0278364924122744843:8(1151-1174)Online publication date: 1-Jul-2024
  • Show More Cited By

View Options

View options

Get Access

Login options

Media

Figures

Other

Tables

Share

Share

Share this Publication link

Share on social media

-