Semantic MapNet: Building Allocentric SemanticMaps and Representations from Egocentric Views

10/02/2020
by   Vincent Cartillier, et al.
1

We study the task of semantic mapping - specifically, an embodied agent (a robot or an egocentric AI assistant) is given a tour of a new environment and asked to build an allocentric top-down semantic map ("what is where?") from egocentric observations of an RGB-D camera with known pose (via localization sensors). Towards this goal, we present SemanticMapNet (SMNet), which consists of: (1) an Egocentric Visual Encoder that encodes each egocentric RGB-D frame, (2) a Feature Projector that projects egocentric features to appropriate locations on a floor-plan, (3) a Spatial Memory Tensor of size floor-plan length x width x feature-dims that learns to accumulate projected egocentric features, and (4) a Map Decoder that uses the memory tensor to produce semantic top-down maps. SMNet combines the strengths of (known) projective camera geometry and neural representation learning. On the task of semantic mapping in the Matterport3D dataset, SMNet significantly outperforms competitive baselines by 4.01-16.81 metrics. Moreover, we show how to use the neural episodic memories and spatio-semantic allocentric representations build by SMNet for subsequent tasks in the same space - navigating to objects seen during the tour("Find chair") or answering questions about the space ("How many chairs did you see in the house?").

READ FULL TEXT

page 1

page 3

page 5

page 7

page 12

research
05/03/2022

Episodic Memory Question Answering

Egocentric augmented reality devices such as wearable glasses passively ...
research
09/17/2022

Topological Semantic Graph Memory for Image-Goal Navigation

A novel framework is proposed to incrementally collect landmark-based gr...
research
09/23/2022

Automatic Sign Reading and Localization for Semantic Mapping with an Office Robot

Semantic mapping is the task of providing a robot with a map of its envi...
research
03/01/2019

Volumetric Instance-Aware Semantic Mapping and 3D Object Discovery

To autonomously navigate and plan interactions in real-world environment...
research
11/18/2019

Simultaneous Mapping and Target Driven Navigation

This work presents a modular architecture for simultaneous mapping and t...
research
11/13/2020

Online Object-Oriented Semantic Mapping and Map Updating with Modular Representations

Creating and maintaining an accurate representation of the environment i...
research
11/26/2019

Semantic Interior Mapology: A Toolbox For Indoor Scene Description From Architectural Floor Plans

We introduce the Semantic Interior Mapology (SIM) toolbox for the conver...

Please sign up or login with your details

Forgot password? Click here to reset