YOLO Adaptive Developments in Complex Natural Environments for Tiny Object Detection

Jikun Zhong,Qing Cheng,Xingchen Hu,Zhong Liu

DOI: https://doi.org/10.3390/electronics13132525

IF: 2.9

2024-06-27

Electronics

Abstract:Detection of tiny object in complex environments is a matter of urgency, not only because of the high real-world demand, but also the high deployment and real-time requirements. Although many current single-stage algorithms have good detection performance under low computing power requirements, there are still significant challenges such as distinguishing the background from object features and extracting small-scale target features in complex natural environments. To address this, we first created real datasets based on natural environments and improved dataset diversity using a combination of copy–paste enhancement and multiple image enhancement techniques. As for the choice of network, we chose YOLOV5s due to its nature of fewer parameters and easier deployment in the same class of models. Most improvement strategies to boost detection performance claim to improve the performance of privilege extraction and recognition. However, we prefer to consider the combination of realistic deployment feasibility and detection performance. Therefore, based on the hottest improvement methods of YOLOV5s, we try to make adaptive improvements in three aspects, namely attention mechanism, head network, and backbone network. The experimental results proved that the decoupled head and Slimneck based improvements achieved, respectively, 0.872 and 0.849, 0.538 and 0.479, 87.5% and 89.8% on the mAP0.5, mAP0.5:0.95, and Precision metrics, surpassing the results of the baseline model on these three metrics: 0.705, 0.405 and 83.6%. This result suggests that the adaptively improved model can better meet routine testing needs without significantly increasing the number of parameters. These models perform well on our custom dataset and are also effective on images that are difficult to detect by naked eye. Meanwhile, we find that YOLOV8s, which also has the decoupled head improvement, has the results of 0.743, 0.461, and 87.17% on these three metrics. It proves that under our dataset, it is possible to achieve more advanced results with lower number of model parameters just by adding decoupled head. And according to the results, we also discuss and analyze some improvements that are not adapted to our dataset, which also provides ideas for researchers in similar scenarios: in the booming development of object detection, choosing the suitable model and adapting to combine with other technologies would help to provide solutions to real-world problems.

engineering, electrical & electronic,physics, applied,computer science, information systems

What problem does this paper attempt to address?

The paper aims to address the problem of small object detection in complex natural environments. Specifically, although existing single-stage detection algorithms have good detection performance with low computational resource requirements, there are still significant challenges in distinguishing background and target features and extracting small-scale target features in complex natural environments. To tackle these challenges, the authors first created a real dataset based on natural environments and enhanced the diversity of the dataset through a combination of copy-paste augmentation and various other image enhancement techniques. The paper chose YOLOV5s as the base model because it has fewer parameters and is easy to deploy. The authors attempted to adaptively improve the model in three aspects: attention mechanism, head network, and backbone network. Experimental results show that the model, after decoupled head and Slimneck improvements, achieved 0.872, 0.849, and 87.5% in mAP0.5, mAP0.5:0.95, and precision metrics, respectively, outperforming the baseline model's performance (0.705, 0.405, and 83.6%, respectively). Additionally, the authors discussed some improvement methods that are not suitable for their dataset, providing insights for research in similar scenarios. In summary, the main contributions of the paper include: 1. Creating a natural scene dataset for human detection that includes mountain, forest, and plain scenes. 2. Proposing a framework to enhance the detector's ability to extract weak targets in complex environments, which can be integrated with other popular detectors. 3. Conducting a comparative analysis of nine YOLOV5-based improvement methods, guiding researchers in choosing appropriate strategies for similar work.

YOLO Adaptive Developments in Complex Natural Environments for Tiny Object Detection

Improved small-object detection using YOLOv8: A comparative study

A novel algorithm for small object detection based on YOLOv4

SP-YOLOv8s: An Improved YOLOv8s Model for Remote Sensing Image Tiny Object Detection

YOLOv10: Real-Time End-to-End Object Detection

Multi-scene small object detection with modified YOLOv4

YOLO-TLA: An Efficient and Lightweight Small Object Detection Model based on YOLOv5

An improved YOLOv8 algorithm for small object detection in autonomous driving

LAYN: Lightweight Multi-Scale Attention YOLOv8 Network for Small Object Detection

An improved lightweight object detection algorithm for YOLOv5

YOLO-SDH: improved YOLOv5 using scaled decoupled head for object detection

YOLO-Z: Improving small object detection in YOLOv5 for autonomous vehicles

Small Target-YOLOv5: Enhancing the Algorithm for Small Object Detection in Drone Aerial Imagery Based on YOLOv5

Enhancing YOLOv8's Performance in Complex Traffic Scenarios: Optimization Design for Handling Long-Distance Dependencies and Complex Feature Relationships

YOLOv6: A Single-Stage Object Detection Framework for Industrial Applications

MS-YOLO: integration-based multi-subnets neural network for object detection in aerial images

Object Detection for Remote Sensing Based on the Enhanced YOLOv8 With WBiFPN

The phosphorus distribution in blood and the calcium and phosphorus excretion during hypervitaminosis D.

Accurate and real-time object detection in crowded indoor spaces based on the fusion of DBSCAN algorithm and improved YOLOv4-tiny network

A YOLO-NL object detector for real-time detection

AIE-YOLO: Auxiliary Information Enhanced YOLO for Small Object Detection