Affiliation:
1. School of Electronic Information Engineering, China West Normal University, Nanchong 637009, China
Abstract
Object detection is one of the core technologies for autonomous driving. Current road object detection mainly relies on visible light, which is prone to missed detections and false alarms in rainy, night-time, and foggy scenes. Multispectral object detection based on the fusion of RGB and infrared images can effectively address the challenges of complex and changing road scenes, improving the detection performance of current algorithms in complex scenarios. However, previous multispectral detection algorithms suffer from issues such as poor fusion of dual-mode information, poor detection performance for multi-scale objects, and inadequate utilization of semantic information. To address these challenges and enhance the detection performance in complex road scenes, this paper proposes a novel multispectral object detection algorithm called MRD-YOLO. In MRD-YOLO, we utilize interaction-based feature extraction to effectively fuse information and introduce the BIC-Fusion module with attention guidance to fuse different modal information. We also incorporate the SAConv module to improve the model’s detection performance for multi-scale objects and utilize the AIFI structure to enhance the utilization of semantic information. Finally, we conduct experiments on two major public datasets, FLIR_Aligned and M3FD. The experimental results demonstrate that compared to other algorithms, the proposed algorithm achieves superior detection performance in complex road scenes.
Funder
China West Normal University
Reference47 articles.
1. Navarro, P.J., Fernández, C., Borraz, R., and Alonso, D. (2017). A Machine Learning Approach to Pedestrian Detection for Autonomous Vehicles Using High-Definition 3D Range Data. Sensors, 17.
2. Deep reinforcement learning with visual attention for vehicle classification;Zhao;IEEE Trans. Cogn. Devel. Syst.,2017
3. Human behavior-based target tracking with an omni-directional thermal camera;Benli;IEEE Trans. Cogn. Devel. Syst.,2019
4. Bao, C., Cao, J., Hao, Q., Cheng, Y., Ning, Y., and Zhao, T. (2023). Dual-YOLO Architecture from Infrared and Visible Images for Object Detection. Sensors, 23.
5. Two-scale image fusion of visible and infrared images using saliency detection;Bavirisetti;Infrared Phys. Technol.,2016
Cited by
2 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献