Weakly supervised learning of indoor geometry by dual warping

08/10/2018
by   Pulak Purkait, et al.
4

A major element of depth perception and 3D understanding is the ability to predict the 3D layout of a scene and its contained objects for a novel pose. Indoor environments are particularly suitable for novel view prediction, since the set of objects in such environments is relatively restricted. In this work we address the task of 3D prediction especially for indoor scenes by leveraging only weak supervision. In the literature 3D scene prediction is usually solved via a 3D voxel grid. However, such methods are limited to estimating rather coarse 3D voxel grids, since predicting entire voxel spaces has large computational costs. Hence, our method operates in image-space rather than in voxel space, and the task of 3D estimation essentially becomes a depth image completion problem. We propose a novel approach to easily generate training data containing depth maps with realistic occlusions, and subsequently train a network for completing those occluded regions. Using multiple publicly available dataset song2017semantic,Silberman:ECCV12 we benchmark our method against existing approaches and are able to obtain superior performance. We further demonstrate the flexibility of our method by presenting results for new view synthesis of RGB-D images.

READ FULL TEXT

page 3

page 6

page 7

page 9

page 10

page 11

research
11/19/2018

Indoor GeoNet: Weakly Supervised Hybrid Learning for Depth and Pose Estimation

Humans naturally perceive a 3D scene in front of them through accumulati...
research
03/13/2019

Putting Humans in a Scene: Learning Affordance in 3D Indoor Environments

Affordance modeling plays an important role in visual understanding. In ...
research
06/26/2023

Self-supervised novel 2D view synthesis of large-scale scenes with efficient multi-scale voxel carving

The task of generating novel views of real scenes is increasingly import...
research
04/03/2017

Hierarchical Surface Prediction for 3D Object Reconstruction

Recently, Convolutional Neural Networks have shown promising results for...
research
12/12/2021

BIPS: Bi-modal Indoor Panorama Synthesis via Residual Depth-aided Adversarial Learning

Providing omnidirectional depth along with RGB information is important ...
research
02/23/2023

VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion

Humans can easily imagine the complete 3D geometry of occluded objects a...
research
03/15/2022

From 2D to 3D: Re-thinking Benchmarking of Monocular Depth Prediction

There have been numerous recently proposed methods for monocular depth p...

Please sign up or login with your details

Forgot password? Click here to reset