YOLOv8 to YOLO11 Performance Benchmark and Comprehensive Architectural Comparative Review
Priyanto Hidayatullah, Nurjannah Syakrani, Muhammad Rizqi Sholahuddin, Trisna Gelar, Refdinal Tubagus
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
In the domain of deep learning-driven computer vision, YOLO is revolutionary. However, not all YOLO models are accompanied by academic articles and architectural diagrams. It complicates the comprehension of the model's operation. Moreover, the existing review papers fail to examine each model comprehensively. This work aims to provide a thorough comparative analysis of the architectures from YOLOv8 to YOLO11, allowing readers to swiftly understand the operational mechanisms and differences among the models. We analyzed the architecture of each YOLO version by reviewing relevant scholarly articles, official documentation, and examining the source code. In particular, we discovered that YOLOv8 through YOLO11 differ in novelty while sharing similarities in the anchor-free and Non-Maximum Suppression (NMS) aspects, except YOLOv10 (NMS-free). Each also has drawbacks, such as differing levels of complexity in the way features are connected (v8), architectural structure and training (v9), training methods or dual assignments (v10), inference, and code implementation (v11). While each version improves architecture, some blocks remain unchanged. This study helps readers understand different YOLO version architectures and inspires how to improve their performance. It also provides readers with a comprehensive architecture diagram and detailed descriptions of each block, serving as a reference for both academic and practical applications. In terms of performance, a benchmark using the Roboflow 100 dataset reveals that YOLOv9 achieves superior accuracy; however, it is eight times slower owing to its NMS mechanism. YOLOv10 is the fastest but least accurate, whereas YOLOv8 and YOLO11 provide a balanced compromise between speed and accuracy.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Building Energy and Comfort Optimization
参考文献 0
引用本文 47
按被引量排序,此处列出前 3 条