Small Object Detection with YOLO: A Performance Analysis Across Model Versions and Hardware

๐Ÿ“… 2025-04-14
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This work addresses the degradation of small-object detection performance in YOLO variants (v5โ€“v11) across heterogeneous hardware platforms (CPU/GPU) and inference backends (ONNX Runtime, OpenVINO, TensorRT), specifically for objects occupying 1%โ€“5% of image area. We conduct a systematic benchmark evaluating accuracyโ€“latency trade-offs under realistic deployment conditions. This is the first cross-generational, horizontal sensitivity analysis of five YOLO versions to object scale, establishing a four-dimensional benchmark spanning hardware, model architecture, accuracy (mAP@0.5), and latency. Results show that YOLOv8/v10 achieve the best balance between CPU inference speed and small-object recall; YOLOv11 improves GPU mAP@0.5 by 3.2% but exhibits >40% miss rate for objects <2.5% image area. The study delivers a reproducible, cross-platform model selection decision map, providing empirical guidance for deploying small-object detectors in edge and cloud environments.

Technology Category

Machine Learning: Hardware-aware MLComputer Vision: Object Detection & CategorizationSearch and Optimization: Evaluation and Analysis

Application Category

Systems and Infrastructure for Web, Mobile and WoT: Web performance, measurement, and characterizationSearch and Retrieval-Augmented AI: Web evaluation methodologies and metricsGraph Algorithms and Modeling for the Web: Foundation models and LLMs for Web-related graphs
๐Ÿ“ Abstract
This paper provides an extensive evaluation of YOLO object detection models (v5, v8, v9, v10, v11) by com- paring their performance across various hardware platforms and optimization libraries. Our study investigates inference speed and detection accuracy on Intel and AMD CPUs using popular libraries such as ONNX and OpenVINO, as well as on GPUs through TensorRT and other GPU-optimized frameworks. Furthermore, we analyze the sensitivity of these YOLO models to object size within the image, examining performance when detecting objects that occupy 1%, 2.5%, and 5% of the total area of the image. By identifying the trade-offs in efficiency, accuracy, and object size adaptability, this paper offers insights for optimal model selection based on specific hardware constraints and detection requirements, aiding practitioners in deploying YOLO models effectively for real-world applications.
Problem

Research questions and friction points this paper is trying to address.

Evaluating YOLO model performance across versions and hardware platforms
Analyzing inference speed and accuracy on CPUs and GPUs with optimization libraries
Assessing YOLO sensitivity to object size for optimal model selection
Innovation

Methods, ideas, or system contributions that make the work stand out.

Evaluates YOLO models across hardware platforms
Analyzes performance with ONNX, OpenVINO, TensorRT
Assesses sensitivity to small object detection
๐Ÿ”Ž Similar Papers
2021-07-01IEEE International Conference on Application-Specific Systems, Architectures, and ProcessorsCitations: 13
๐Ÿ’ผ Related Jobs
No related jobs found.
University of Management and Technology
M
Muhammad Fasih Tariq
School of Systems and Technology, University of Management and Technology
M
Muhammad Azeem Javed
School of Systems and Technology, University of Management and Technology