Fusion of multi-scale attention for aerial images small-target detection model based on PARE-YOLO
Huiying Zhang, Xiao Pan, Feifan Yao, Qinghua Zhang, Yifei Gong
Jilin University of Chemical Technology
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
In view of the complex environments and varying object scales in drone-captured imagery, a novel PARE-YOLO algorithm based on YOLOv8 for small object detection is proposed. This model enhances feature extraction and fusion across multiple scales through a restructured neck network. Additionally, it incorporates a lightweight detection head that is optimized for small objects, thereby significantly improving detection performance in cluttered and intricate backgrounds. To further enhance the extraction of small object features, the conventional C2f is replaced with a novel architecture. Moreover, the EMA-GIoU loss function is proposed to mitigate class imbalance and enhance robustness, particularly in scenarios characterized by skewed class distributions. Evaluation on the VisDrone2019 dataset indicates that PARE-YOLO achieves a 5.9% improvement in mean Average Precision (mAP) at a threshold of 0.5, compared to the original YOLOv8 model. In addition, the PARE-YOLO model exhibits significant robustness, achieving a mean Average Precision (mAP) at a threshold of 0.5 values on the HIT-UAV dataset that are 0.8%, 0.5%, and 1.2% higher than those of YOLOv8, YOLOv10, and RT-DETR. These results underscore the effectiveness of PARE-YOLO in addressing the challenges inherent in aerial scenarios. The code will be available online (https://github.com/Sunnyxiao69/PARE-YOLO).
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Infrared Target Detection Methodologies
Advanced Image Fusion Techniques · Advanced Neural Network Applications
参考文献 39
此处列出前 3 条
引用本文 41
按被引量排序,此处列出前 3 条