AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.1 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

DV-YOLO: a deep learning framework for small-object detection in UAV-based remote sensing imagery with applications to smart logistics

Ahmed A. Alsheikhy1Mohammad Barr2Sahbi Boubaker1( )Yahia Said3( )
Department of Computer and Network Engineering, College of Computer Science and Engineering, University of Jeddah, Jeddah 21959, Saudi Arabia
Department of Electrical Engineering, College of Engineering, Northern Border University, Arar 91431, Saudi Arabia
Center for Scientific Research and Entrepreneurship, Northern Border University, Arar 73213, Saudi Arabia
Show Author Information

Abstract

Unmanned aerial vehicles (UAVs) are being increasingly adopted as flexible remote sensing platforms for smart logistics applications, including warehouse inventory, last-mile delivery supervision, traffic flow analysis, port operations, and infrastructure inspection. Despite their advantages, reliable object detection in UAV-based remote sensing imagery remains challenging due to small object sizes, dense object distributions, arbitrary orientations, and complex backgrounds commonly encountered in logistics environments. Although recent YOLO-based detectors have shown promising performance, their effectiveness is often limited in high-resolution aerial scenes and under practical computational constraints imposed by UAV platforms. To address these challenges, this paper proposes DV-YOLO, an enhanced deep learning framework tailored for object detection in UAV-based remote sensing imagery for logistics-oriented applications. The proposed model extends YOLOv9 through a deeper and wider backbone architecture coupled with optimized feature fusion strategies that jointly exploit spatial and semantic representations. A novel cross-path fusion network at deep feature map (CPFNDFM) is introduced to improve the detection of small and densely distributed logistics-related objects such as vehicles, containers, and infrastructure elements. In addition, a lightweight connection aggregation (CA) module, inspired by VoVNet and ShuffleNetV2, is integrated to enhance feature reuse while maintaining computational efficiency suitable for real-time UAV deployment. Furthermore, a challenging benchmark dataset, termed harder vision drone, is constructed by combining and refining samples from VisDrone and DOTA to better reflect real-world UAV remote sensing scenarios in logistics environments. Extensive experimental evaluations conducted on VisDrone 2021, DOTA v2, and the proposed dataset demonstrate that DV-YOLO consistently outperforms state-of-the-art detectors, achieving up to 3.5% improvement in mean average precision (mAP) compared with YOLOv9. These results highlight the potential of the proposed framework to support robust, accurate, and efficient aerial perception for smart logistics and UAV-based remote sensing applications.

CLC number: 68T07, 68U10

References

【1】
【1】
 
 
AIMS Mathematics
Pages 12043-12063

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Alsheikhy AA, Barr M, Boubaker S, et al. DV-YOLO: a deep learning framework for small-object detection in UAV-based remote sensing imagery with applications to smart logistics. AIMS Mathematics, 2026, 11(4): 12043-12063. https://doi.org/10.3934/math.2026494

554

Views

27

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 31 January 2026
Revised: 07 March 2026
Accepted: 17 March 2026
Published: 29 April 2026
©2026 the Author(s), licensee AIMS Press.

This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0)