Abstract:Target detection is the basis for automatic target recognition system of infrared imaging guidance to complete subsequent tasks such as recognition and tracking. Existing systems have not the autonomous learning ability of target feature, and it will be powerless once the task environment exceeds the pre-planned condition. The single-stage target detection based on deep learning has the ability of autonomous learning and high computational efficiency, which is an effective way to solve the problem of infrared imaging guidance target detection in complex environment. SSD (Single Shot MultiBox Detector) is a classical single-stage detection model, however, the convolution layer with strong semantic information in SSD has low resolution, which is not conducive to small target detection. In addition, the location loss of SSD does not consider the impact of target scale change. Therefore, this paper puts forward two improvement ideas in view of SSD: (1) Starting from FPN (Feature Pyramid Network), feature channel’s importance is distinguished through efficient channel attention mechanism, the contribution of each feature layer to the fusion output is described based on the learnable weight, and the feature weighted fusion of bidirectional multi-scale is realized between the feature layer which has low resolution and strong semantics and the feature layer which has high resolution and weak semantics. (2) Starting from IoU (Intersection over Union) and considering non overlapping parts and geometric relationship between the predicted box and the ground-truth box, the location loss of SSD that remains invariable to the target scale change is constructed to improve the sensitivity of the detection model to the locating error of small target. The experimental results show that, for 300 × 300 input, the presented method achieves 84.7% mAP (mean Average Precision) on VOC2007 test and for 512 × 512 input, it reaches 86.6%. On the self-built infrared aircraft data set, the proposed method achieves 81.1% mAP and can detect more small targets. Without affecting detection speed, the presented method on experimental results outperforms some comparable state-of-the-art models such as YOLOv3 (You Only Look Once), DSSD (Deconvolutional Single Shot Multibox Detector), RSSD (Rainbow Single Shot Multibox Detector) and FSSD (Fusion Single Shot Multibox Detector).

Multi-Objective Detection of Traffic Scenes Based on Improved SSD

Real-Time Detection Network SI-SSD for Weak Targets in Complex Traffic Scenarios

Lightweight Real-Time Object Detection via Enhanced Global Perception and Intra-Layer Interaction for Complex Traffic Scenarios

High-precision real-time autonomous driving target detection based on YOLOv8

Traffic Sign Detection Method Based on Improved SSD

An improved SSD method for infrared target detection based on convolutional neural network

Multiclass objects detection algorithm using DarkNet-53 and DenseNet for intelligent vehicles

Attention Mechanism and Detection Box Information Based Real-time Multi-Object Vehicle Detection

Multitarget Detection in Depth-Perception Traffic Scenarios

An Intelligent Traffic Analysis and Prediction System Using Deep Learning Technique

Multi-Object Detection in Security Screening Scene Based on Convolutional Neural Network

A Multi-Scale Traffic Object Detection Algorithm for Road Scenes Based on Improved YOLOv5

Video Face Detection Based on Improved SSD Model and Target Tracking Algorithm

Generalized Haar Filter based Deep Networks for Real-Time Object Detection in Traffic Scene

Front Vehicle Detection Algorithm for Smart Car Based on Improved SSD Model

Real-Time Target Detection Method Based on Lightweight Convolutional Neural Network

A study on a target detection model for autonomous driving tasks

Improving real-time object detection in Internet-of-Things smart city traffic with YOLOv8-DSAF method

Efficient Automatic Driving Instance Segmentation Method Based on Detection

Real-Time Vehicle Object Detection Method Based on Multi-Scale Feature Fusion