Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

Chen, Xiaotian; Chen, Xuejin; Zha, Zheng-Jun

Computer Science > Computer Vision and Pattern Recognition

arXiv:1907.06023 (cs)

[Submitted on 13 Jul 2019]

Title:Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

Authors:Xiaotian Chen, Xuejin Chen, Zheng-Jun Zha

View PDF

Abstract:Monocular depth estimation is an essential task for scene understanding. The underlying structure of objects and stuff in a complex scene is critical to recovering accurate and visually-pleasing depth maps. Global structure conveys scene layouts, while local structure reflects shape details. Recently developed approaches based on convolutional neural networks (CNNs) significantly improve the performance of depth estimation. However, few of them take into account multi-scale structures in complex scenes. In this paper, we propose a Structure-Aware Residual Pyramid Network (SARPN) to exploit multi-scale structures for accurate depth prediction. We propose a Residual Pyramid Decoder (RPD) which expresses global scene structure in upper levels to represent layouts, and local structure in lower levels to present shape details. At each level, we propose Residual Refinement Modules (RRM) that predict residual maps to progressively add finer structures on the coarser structure predicted at the upper level. In order to fully exploit multi-scale image features, an Adaptive Dense Feature Fusion (ADFF) module, which adaptively fuses effective features from all scales for inferring structures of each scale, is introduced. Experiment results on the challenging NYU-Depth v2 dataset demonstrate that our proposed approach achieves state-of-the-art performance in both qualitative and quantitative evaluation. The code is available at this https URL.

Comments:	7pages, 7figures, Accepted by the 28th International Joint Conference on Artificial Intelligence (IJCAI-2019)
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1907.06023 [cs.CV]
	(or arXiv:1907.06023v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1907.06023

Submission history

From: Xiaotian Chen [view email]
[v1] Sat, 13 Jul 2019 07:31:24 UTC (6,575 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators