Monocular Depth Estimation Using Res-UNet with an Attention Model-Reference-Cited by-同舟云学术

Monocular Depth Estimation Using Res-UNet with an Attention Model

Published:2023-05-22 Issue:10 Volume:13 Page:6319
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Jan Abdullah¹,Seo Suyoung¹^ORCID

Affiliation:

1. Department of Civil Engineering, School of Architectural, Civil, Environmental, and Energy Engineering, Kyungpook National University, Daegu 41566, Republic of Korea

Abstract

Depth maps are single image metrics that carry the information of a scene in three-dimensional axes. Accurate depth maps can recreate the 3D structure of a scene, which helps in understanding the full geometry of the objects within the scene. Depth maps can be generated from a single image or multiple images. Single-image depth mapping is also known as monocular depth mapping. Depth maps are ill-posed problems that are complex and require extensive calibration. Therefore, recent methods use deep learning to develop depth maps. We propose a new method in monocular depth estimation to develop a high-quality depth map. Our approach is based on a convolutional neural network in which we used Res-UNet with a spatial attention model to develop depth maps. The addition of an attention mechanism increases the capability of feature extraction and enhances the boundaries features. It does not add any extra parameters to the network. With our proposed model, we demonstrate that a simple CNN model aided with an attention mechanism can create high-quality depth maps with a small iteration and training time. Our model performs very well compared to the existing state-of-the-art methods on the benchmark NYU-depth v2 dataset. Our model is flexible and can be applied to any depth mapping or multi-segmentation tasks.

Funder

Ministry of Education

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/13/10/6319/pdf

Reference54 articles.

1. Hartley, R., and Zisserman, A. (2004). Multiple View Geometry in Computer Vision, Cambridge University Press. [2nd ed.].

2. Mobile Robot Localization by Tracking Geometric Beacons;Leonard;IEEE Trans. Robot. Autom.,1991

3. Leonard, J.J., Durrant-Whyte, H.F., and Pj, O. (2002). Directed Sonar Sensing for Mobile Robot Navigation, Springer Science & Business Media.

4. Indoor Segmentation and Support Inference from RGBD Images;Fitzgibbon;Computer Vision—ECCV 2012,2012

5. Eigen, D., Puhrsch, C., and Fergus, R. (2014). Depth Map Prediction from a Single Image Using a Multi-Scale Deep Network. arXiv.

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Hybrid Visual Odometry Algorithm Using a Downward-Facing Monocular Camera;Applied Sciences;2024-09-02

2. Structural Crack Detection Using Deep Learning: An In-depth Review;KOREAN J REMOTE SENS;2023