Unsupervised Learning of Geometry From Videos With Edge-Aware Depth-Normal Consistency

Zhenheng Yang; Peng Wang; Wei Xu; Liang Zhao; Ramakant Nevatia

doi:10.1609/aaai.v32i1.12257

Authors

Zhenheng Yang University of Southern California
Peng Wang Baidu USA
Wei Xu Baidu USA
Liang Zhao Baidu USA
Ramakant Nevatia University of Southern California

DOI:

https://doi.org/10.1609/aaai.v32i1.12257

Keywords:

unsupervised learning, depth, 3D geometry

Abstract

Learning to reconstruct depths from a single image by watching unlabeled videos via deep convolutional network (DCN) is attracting significant attention in recent years, e.g. (Zhou et al. 2017). In this paper, we propose to use surface normal representation for unsupervised depth estimation framework. Our estimated depths are constrained to be compatible with predicted normals, yielding more robust geometry results. Specifically, we formulate an edge-aware depth-normal consistency term, and solve it by constructing a depth-to-normal layer and a normal-to-depth layer inside of the DCN. The depth-to-normal layer takes estimated depths as input, and computes normal directions using cross production based on neighboring pixels. Then given the estimated normals, the normal-to-depth layer outputs a regularized depth map through local planar smoothness. Both layers are computed with awareness of edges inside the image to help address the issue of depth/normal discontinuity and preserve sharp edges. Finally, to train the network, we apply the photometric error and gradient smoothness to supervise both depth and normal predictions. We conducted experiments on both outdoor (KITTI) and indoor (NYUv2) datasets, and showed that our algorithm vastly outperforms state-of-the-art, which demonstrates the benefits of our approach.

Unsupervised Learning of Geometry From Videos With Edge-Aware Depth-Normal Consistency

Authors

DOI:

Keywords:

Abstract

Downloads

Published

How to Cite

Issue

Section

Information

Subscription