Optimal Configuration of Multi-Task Learning for Autonomous Driving-Reference-Cited by-同舟云学术

Optimal Configuration of Multi-Task Learning for Autonomous Driving

Published:2023-12-09 Issue:24 Volume:23 Page:9729
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Jun Woomin¹²,Son Minjun¹²,Yoo Jisang¹²,Lee Sungjin¹²^ORCID

Affiliation:

1. Electronic Engineering, Dong Seoul University, Seongnam 13117, Republic of Korea

2. Autonomous Driving Lab., MODULABS, Seoul 06252, Republic of Korea

Abstract

For autonomous driving, it is imperative to perform various high-computation image recognition tasks with high accuracy, utilizing diverse sensors to perceive the surrounding environment. Specifically, cameras are used to perform lane detection, object detection, and segmentation, and, in the absence of lidar, tasks extend to inferring 3D information through depth estimation, 3D object detection, 3D reconstruction, and SLAM. However, accurately processing all these image recognition operations in real-time for autonomous driving under constrained hardware conditions is practically unfeasible. In this study, considering the characteristics of image recognition tasks performed by these sensors and the given hardware conditions, we investigated MTL (multi-task learning), which enables parallel execution of various image recognition tasks to maximize their processing speed, accuracy, and memory efficiency. Particularly, this study analyzes the combinations of image recognition tasks for autonomous driving and proposes the MDO (multi-task decision and optimization) algorithm, consisting of three steps, as a means for optimization. In the initial step, a MTS (multi-task set) is selected to minimize overall latency while meeting minimum accuracy requirements. Subsequently, additional training of the shared backbone and individual subnets is conducted to enhance accuracy with the predefined MTS. Finally, both the shared backbone and each subnet undergo compression while maintaining the already secured accuracy and latency performance. The experimental results indicate that integrated accuracy performance is critically important in the configuration and optimization of MTL, and this integrated accuracy is determined by the ITC (inter-task correlation). The MDO algorithm was designed to consider these characteristics and construct multi-task sets with tasks that exhibit high ITC. Furthermore, the implementation of the proposed MDO algorithm, coupled with additional SSL (semi-supervised learning) based training, resulted in a significant performance enhancement. This advancement manifested as approximately a 12% increase in object detection mAP performance, a 15% improvement in lane detection accuracy, and a 27% reduction in latency, surpassing the results of previous three-task learning techniques like YOLOP and HybridNet.

Funder

National Research Foundation of Korea

Ministry of Education and Brain Impact

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/24/9729/pdf

Reference64 articles.

1. A Survey of Deep Learning Techniques for Autonomous Driving;Grigorescu;J. Field Robot.,2020

2. Deep Learning in Robotics: Survey on Model Structures and Training Strategies;Galambos;IEEE Trans. Syst. Man Cybern.,2021

3. Rethinking Real-Time Lane Detection Technology for Autonomous Driving;Kwak;J. Korean Inst. Commun. Inf. Sci.,2023

4. Efficient Training Methodology in an Image Classification Network;Bae;J. Korean Inst. Commun. Inf. Sci.,2021

5. Lee, H., Lee, N., and Lee, S. (2022). A Method of Deep Learning Model Optimization for Image Classification on Edge Device. Sensors, 22.

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Multi-Object Trajectory Prediction Based on Lane Information and Generative Adversarial Network;Sensors;2024-02-17