Improved Mask R-CNN Multi-Target Detection and Segmentation for Autonomous Driving in Complex Scenes-Reference-Cited by-同舟云学术

Improved Mask R-CNN Multi-Target Detection and Segmentation for Autonomous Driving in Complex Scenes

Published:2023-04-10 Issue:8 Volume:23 Page:3853
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Fang Shuqi¹^ORCID,Zhang Bin¹,Hu Jingyu¹

Affiliation:

1. School of Electronic and Automation, Guilin University of Electronic Technology, Guilin 541004, China

Abstract

Vision-based target detection and segmentation has been an important research content for environment perception in autonomous driving, but the mainstream target detection and segmentation algorithms have the problems of low detection accuracy and poor mask segmentation quality for multi-target detection and segmentation in complex traffic scenes. To address this problem, this paper improved the Mask R-CNN by replacing the backbone network ResNet with the ResNeXt network with group convolution to further improve the feature extraction capability of the model. Furthermore, a bottom-up path enhancement strategy was added to the Feature Pyramid Network (FPN) to achieve feature fusion, while an efficient channel attention module (ECA) was added to the backbone feature extraction network to optimize the high-level low resolution semantic information graph. Finally, the bounding box regression loss function smooth L1 loss was replaced by CIoU loss to speed up the model convergence and minimize the error. The experimental results showed that the improved Mask R-CNN algorithm achieved 62.62% mAP for target detection and 57.58% mAP for segmentation accuracy on the publicly available CityScapes autonomous driving dataset, which were 4.73% and 3.96%% better than the original Mask R-CNN algorithm, respectively. The migration experiments showed that it has good detection and segmentation effects in each traffic scenario of the publicly available BDD autonomous driving dataset.

Funder

National Natural Science Foundation of China

Guangxi Natural Science Foundation

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/8/3853/pdf

Reference36 articles.

1. A survey of deep learning techniques for autonomous driving;Grigorescu;J. Field Robot.,2022

2. Computer vision for autonomous vehicles: Problems, datasets and state of the art;Janai;Found. Trends® Comput. Graph. Vis.,2020

3. A survey of instance segmentation research based on deep learning;Su;CAAI Trans. Intell. Syst.,2022

4. Joseph, R., Santosh, D., Ross, G., and Ali, F. (2016, January 27–30). You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.

5. Liu, W., Anguelov, D., Erhan, D., Szegedy, C., Reed, S., Fu, C.-F., and Berg, A.C. (2016, January 11–14). Ssd: Single shot multibox detector. Proceedings of the Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands. Part I.

Cited by 19 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Enhanced human motion detection with hybrid RDA-WOA-based RNN and multiple hypothesis tracking for occlusion handling;Image and Vision Computing;2024-10

2. Deep learning for automated boundary detection and segmentation in organ donation photography;Innovative Surgical Sciences;2024-08-20

3. UAV Inspections of Power Transmission Networks with AI Technology: A Case Study of Lesvos Island in Greece;Energies;2024-07-18

4. A semi-supervised mixture model of visual language multitask for vehicle recognition;Applied Soft Computing;2024-07

5. Detection of Straw Coverage under Conservation Tillage Based on an Improved Mask Regional Convolutional Neural Network (Mask R-CNN);Agronomy;2024-06-28