Abstract:Pedestrian detection plays a critical role in computer vision as it contributes to ensuring traffic safety. Existing methods that rely solely on RGB images suffer from performance degradation under low-light conditions due to the lack of useful information. To address this issue, recent multispectral detection approaches have combined thermal images to provide complementary information and have obtained enhanced performances. Nevertheless, few approaches focus on the negative effects of false positives (FPs) caused by noisy fused feature maps. Different from them, we comprehensively analyze the impacts of FPs on detection performance and find that enhancing feature contrast can significantly reduce these FPs. In this article, we propose a novel target-aware fusion strategy for multispectral pedestrian detection, named TFDet. The target-aware fusion strategy employs a fusion-refinement paradigm. In the fusion phase, we reveal the parallel-and cross-channel similarities in RGB and thermal features and learn an adaptive receptive field to collect useful information from both features. In the refinement phase, we use a segmentation branch to discriminate the pedestrian features from the background features. We propose a correlation-maximum loss function to enhance the contrast between the pedestrian features and background features. As a result, our fusion strategy highlights pedestrian-related features and suppresses unrelated ones, generating more discriminative fused features. TFDet achieves state-of-the-art performance on two multispectral pedestrian benchmarks, KAIST and LLVIP, with absolute gains of 0.65% and 4.1% over the previous best approaches, respectively. TFDet can easily extend to multiclass object detection scenarios. It outperforms the previous best approaches on two multispectral object detection benchmarks, FLIR and M3FD, with absolute gains of 2.2% and 1.9%, respectively. Importantly, TFDet has comparable inference efficiency to the previous approaches and has remarkably good detection performance even under low-light conditions, which is a significant advancement for ensuring road safety. The code will be made publicly available at https://github.com/XueZ-phd/TFDet.git.

Fusion-attention network using dense scale-invariant feature transform flow image and point cloud for 3D pedestrian detection

Transformer fusion and histogram layer multispectral pedestrian detection network

Fused DNN: A deep neural network fusion approach to fast and robust pedestrian detection

Accurate and Real-Time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network

3D Vehicle Detection Using Multi-Level Fusion From Point Clouds and Images

3D Object Detection Based on Attention and Multi-Scale Feature Fusion

A Pedestrian Detection Algorithm Based on Score Fusion for Multi-LiDAR Systems

Cascade fusion of multi-modal and multi-source feature fusion by the attention for three-dimensional object detection

A Pedestrian Detection and Tracking Framework for Autonomous Cars: Efficient Fusion of Camera and LiDAR Data

Spatio-Contextual Deep Network Based Multimodal Pedestrian Detection For Autonomous Driving

TFDet: Target-Aware Fusion for RGB-T Pedestrian Detection

AEPF: Attention-Enabled Point Fusion for 3D Object Detection

A Fast RetinaNet Fusion Framework for Multi-Spectral Pedestrian Detection

SemanticVoxels: Sequential Fusion for 3D Pedestrian Detection using LiDAR Point Cloud and Semantic Segmentation

Key points and visible part fusion attention network for occluded pedestrian detection in traffic environments

Optimal Fusion-based Asymmetric Two-stream Networks for Multispectral Image Pedestrian Detection

CrossFusion net: Deep 3D object detection based on RGB images and point clouds in autonomous driving

3D object detection based on fusion of image and point cloud in autonomous driving traffic scenarios

S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection

Multiattention Mechanism 3D Object Detection Algorithm Based on RGB and LiDAR Fusion for Intelligent Driving