Recognizing Trained and Untrained Obstacles around a Port Transfer Crane Using an Image Segmentation Model and Coordinate Mapping between the Ground and Image-Reference-Cited by-同舟云学术

Recognizing Trained and Untrained Obstacles around a Port Transfer Crane Using an Image Segmentation Model and Coordinate Mapping between the Ground and Image

Published:2023-06-27 Issue:13 Volume:23 Page:5982
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Yu Eunseop¹^ORCID,Ryu Bohyun¹^ORCID

Affiliation:

1. Department of Plant Engineering Center, Institute for Advanced Engineering, Yongin-si 175-28, Republic of Korea

Abstract

Container yard congestion can become a bottleneck in port logistics and result in accidents. Therefore, transfer cranes, which were previously operated manually, are being automated to increase their work efficiency. Moreover, LiDAR is used for recognizing obstacles. However, LiDAR cannot distinguish obstacle types; thus, cranes must move slowly in the risk area, regardless of the obstacle, which reduces their work efficiency. In this study, a novel method for recognizing the position and class of trained and untrained obstacles around a crane using cameras installed on the crane was proposed. First, a semantic segmentation model, which was trained on images of obstacles and the ground, recognizes the obstacles in the camera images. Then, an image filter extracts the obstacle boundaries from the segmented image. Finally, the coordinate mapping table converts the obstacle boundaries in the image coordinate system to the real-world coordinate system. Estimating the distance of a truck with our method resulted in 32 cm error at a distance of 5 m and in 125 cm error at a distance of 30 m. The error of the proposed method is large compared with that of LiDAR; however, it is acceptable because vehicles in ports move at low speeds, and the error decreases as obstacles move closer.

Funder

Ministry of Trade, Industry and Energy

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/13/5982/pdf

Reference27 articles.

1. Zhou, Y., and Tuzel, O. (2018, January 18–23). VoxelNet: End-to-end learning for point cloud based 3D object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.

2. He, Q., Wang, Z., Zeng, H., Zeng, Y., and Liu, Y. (March, January 22). SVGA-Net: Sparse voxel-graph attention network for 3D object detection from point clouds. Proceedings of the AAAI Conference on Artificial Intelligence, Online.

3. Ren, S., He, K., Girshick, R., and Sun, J. (2015, January 7–12). Faster R-CNN: Towards real-time object detection with region proposal networks. Proceedings of the Advances in Neural Information Processing Systems 28, Montreal, QC, Canada.

4. SSD: Single shot multibox detector;Leibe;Computer Vision–ECCV 2016. Proceeding of the ECCV 2016, Amsterdam, The Netherlands, 11–14 October 2016. Lecture Notes in Computer Science,2016

5. Redmon, J., Divvala, S., Girshick, R., and Farhadi, A. (2016, January 27–30). You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Global Time-Varying Path Planning Method Based on Tunable Bezier Curves;Applied Sciences;2023-12-18

2. The Mushroom Sorting System Based on the Analysis of Stereo Vision;2023 12th International Conference on Awareness Science and Technology (iCAST);2023-11-09