Deep panoramic depth prediction and completion for indoor scenes

N/ACitations
Citations of this article
8Readers
Mendeley users who have this article in their library.

Abstract

We introduce a novel end-to-end deep-learning solution for rapidly estimating a dense spherical depth map of an indoor environment. Our input is a single equirectangular image registered with a sparse depth map, as provided by a variety of common capture setups. Depth is inferred by an efficient and lightweight single-branch network, which employs a dynamic gating system to process together dense visual data and sparse geometric data. We exploit the characteristics of typical man-made environments to efficiently compress multi-resolution features and find short- and long-range relations among scene parts. Furthermore, we introduce a new augmentation strategy to make the model robust to different types of sparsity, including those generated by various structured light sensors and LiDAR setups. The experimental results demonstrate that our method provides interactive performance and outperforms state-of-the-art solutions in computational efficiency, adaptivity to variable depth sparsity patterns, and prediction accuracy for challenging indoor data, even when trained solely on synthetic data without any fine tuning.

Cite

CITATION STYLE

APA

Pintore, G., Almansa, E., Sanchez, A., Vassena, G., & Gobbetti, E. (2024). Deep panoramic depth prediction and completion for indoor scenes. Computational Visual Media, 10(5), 903–922. https://doi.org/10.1007/s41095-023-0358-0

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free