Inferring Semantic Information with 3D Neural Scene Representations

03/28/2020
by   Amit Kohli, et al.
0

Biological vision infers multi-modal 3D representations that support reasoning about scene properties such as materials, appearance, affordance, and semantics in 3D. These rich representations enable us humans, for example, to acquire new skills, such as the learning of a new semantic class, with extremely limited supervision. Motivated by this ability of biological vision, we demonstrate that 3D-structure-aware representation learning leads to multi-modal representations that enable 3D semantic segmentation with extremely limited, 2D-only supervision. Building on emerging neural scene representations, which have been developed for modeling the shape and appearance of 3D scenes supervised exclusively by posed 2D images, we are first to demonstrate a representation that jointly encodes shape, appearance, and semantics in a 3D-structure-aware manner. Surprisingly, we find that only a few tens of labeled 2D segmentation masks are required to achieve dense 3D semantic segmentation using a semi-supervised learning strategy. We explore two novel applications for our semantically aware neural scene representation: 3D novel view and semantic label synthesis given only a single input RGB image or 2D label mask, as well as 3D interpolation of appearance and semantics.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/04/2019

Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene Representations

The advent of deep learning has given rise to neural scene representatio...
research
12/27/2018

S4-Net: Geometry-Consistent Semi-Supervised Semantic Segmentation

We show that it is possible to learn semantic segmentation from very lim...
research
11/25/2021

NeSF: Neural Semantic Fields for Generalizable Semantic Segmentation of 3D Scenes

We present NeSF, a method for producing 3D semantic fields from posed RG...
research
03/12/2021

Juggling With Representations: On the Information Transfer Between Imagery, Point Clouds, and Meshes for Multi-Modal Semantics

The automatic semantic segmentation of the huge amount of acquired remot...
research
09/09/2020

Unsupervised Part Discovery by Unsupervised Disentanglement

We address the problem of discovering part segmentations of articulated ...
research
12/28/2020

DeepSurfels: Learning Online Appearance Fusion

We present DeepSurfels, a novel hybrid scene representation for geometry...
research
10/21/2020

Semantics-Guided Representation Learning with Applications to Visual Synthesis

Learning interpretable and interpolatable latent representations has bee...

Please sign up or login with your details

Forgot password? Click here to reset