计算机科学
保险丝(电气)
无人机
目标检测
特征(语言学)
人工智能
对象(语法)
计算
比例(比率)
实时计算
计算机视觉
数据挖掘
模式识别(心理学)
工程类
地理
哲学
电气工程
生物
地图学
遗传学
语言学
算法
作者
Yaning Kong,Xiangfeng Shang,Shijie Jia
出处
期刊:Sensors
[Multidisciplinary Digital Publishing Institute]
日期:2024-08-24
卷期号:24 (17): 5496-5496
被引量:96
摘要
Performing low-latency, high-precision object detection on unmanned aerial vehicles (UAVs) equipped with vision sensors holds significant importance. However, the current limitations of embedded UAV devices present challenges in balancing accuracy and speed, particularly in the analysis of high-precision remote sensing images. This challenge is particularly pronounced in scenarios involving numerous small objects, intricate backgrounds, and occluded overlaps. To address these issues, we introduce the Drone-DETR model, which is based on RT-DETR. To overcome the difficulties associated with detecting small objects and reducing redundant computations arising from complex backgrounds in ultra-wide-angle images, we propose the Effective Small Object Detection Network (ESDNet). This network preserves detailed information about small objects, reduces redundant computations, and adopts a lightweight architecture. Furthermore, we introduce the Enhanced Dual-Path Feature Fusion Attention Module (EDF-FAM) within the neck network. This module is specifically designed to enhance the network’s ability to handle multi-scale objects. We employ a dynamic competitive learning strategy to enhance the model’s capability to efficiently fuse multi-scale features. Additionally, we incorporate the P2 shallow feature layer from the ESDNet into the neck network to enhance the model’s ability to fuse small-object features, thereby enhancing the accuracy of small object detection. Experimental results indicate that the Drone-DETR model achieves an mAP50 of 53.9% with only 28.7 million parameters on the VisDrone2019 dataset, representing an 8.1% enhancement over RT-DETR-R18.
科研通智能强力驱动
Strongly Powered by AbleSci AI