AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Full Length Article | Open Access

Pre-locate net for object detection in high-resolution images

Yunhao ZHANGTing-Bing XUZhenzhong WEI( )
Key Laboratory of Precision Opto-mechatronics Technology, Ministry of Education, and the School of Instrumentation and Optoelectronic Engineering, Beihang University, Beijing 100083, China

Peer review under responsibility of Editorial Committee of CJA.

Show Author Information

Abstract

Small-object detection has long been a challenge. High-megapixel cameras are used to solve this problem in industries. However, current detectors are inefficient for high-resolution images. In this work, we propose a new module called Pre-Locate Net, which is a plug-and-play structure that can be combined with most popular detectors. We inspire the use of classification ideas to obtain candidate regions in images, greatly reducing the amount of calculation, and thus achieving rapid detection in high-resolution images. Pre-Locate Net mainly includes two parts, candidate region classification and behavior classification. Candidate region classification is used to obtain a candidate region, and behavior classification is used to estimate the scale of an object. Different follow-up processing is adopted according to different scales to balance the variance of the network input. Different from the popular candidate region generation method, we abandon the idea of regression of a bounding box and adopt the concept of classification, so as to realize the prediction of a candidate region in the shallow network. We build a high-resolution dataset of aircraft and landing gears covering complex scenes to verify the effectiveness of our method. Compared to state-of-the-art detectors (e.g., Guided Anchoring, Libra-RCNN, and FASF), our method achieves the best mAP of 94.5 on 1920 × 1080 images at 16.7 FPS.

References

【1】
【1】
 
 
Chinese Journal of Aeronautics
Pages 313-325

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
ZHANG Y, XU T-B, WEI Z. Pre-locate net for object detection in high-resolution images. Chinese Journal of Aeronautics, 2022, 35(10): 313-325. https://doi.org/10.1016/j.cja.2021.10.022

415

Views

7

Crossref

4

Web of Science

4

Scopus

0

CSCD

Received: 18 April 2021
Revised: 15 June 2021
Accepted: 10 July 2021
Published: 24 November 2021
© 2021 Chinese Society of Aeronautics and Astronautics.

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).