A multiscale dilated convolution and mixed-order attention-based deep neural network for monocular depth prediction

Huihui Xu,Fei Li

doi:10.1007/s42452-022-05235-1

Huihui Xu, Fei Li

Open Access

PDF Available

https://doi.org/10.1007/s42452-022-05235-1

Copy DOI

Export

Save

Cite

Journal: SN Applied Sciences	Publication Date: Dec 14, 2022
Citations: 2	License type: open-access

Affiliation: Shandong Jianzhu University

Abstract
Full-Text PDF
Similar Papers

Abstract

Listen

Recovering precise depth information from different scenes has become a popular subject in the semantic segmentation and virtual reality fields. This study presents a multiscale dilated convolution and mixed-order attention-based deep neural network for monocular depth recovery. Specifically, we design a multilevel feature enhancement scheme to enhance and fuse high-resolution and low-resolution features on the basis of mixed-order attention. Moreover, a multiscale dilated convolution module that combines four different dilated convolutions is explored for deriving multiscale information and increasing the receptive field. Recent studies have shown that the design of loss terms is crucial to depth prediction. Therefore, an efficient loss function that combines the ℓ1 loss, gradient loss, and classification loss is also designed to promote rich details. Experiments on three public datasets show that the presented approach achieves better performance than state-of-the-art depth prediction methods.

Full Text