HUMANISE: Language-conditioned Human Motion Generation in 3D Scenes

10/18/2022
by   Zan Wang, et al.
30

Learning to generate diverse scene-aware and goal-oriented human motions in 3D scenes remains challenging due to the mediocre characteristics of the existing datasets on Human-Scene Interaction (HSI); they only have limited scale/quality and lack semantics. To fill in the gap, we propose a large-scale and semantic-rich synthetic HSI dataset, denoted as HUMANISE, by aligning the captured human motion sequences with various 3D indoor scenes. We automatically annotate the aligned motions with language descriptions that depict the action and the unique interacting objects in the scene; e.g., sit on the armchair near the desk. HUMANISE thus enables a new generation task, language-conditioned human motion generation in 3D scenes. The proposed task is challenging as it requires joint modeling of the 3D scene, human motion, and natural language. To tackle this task, we present a novel scene-and-language conditioned generative model that can produce 3D human motions of the desirable action interacting with the specified objects. Our experiments demonstrate that our model generates diverse and semantically consistent human motions in 3D scenes.

READ FULL TEXT

page 1

page 4

page 6

page 8

page 9

page 10

page 14

page 16

research
12/08/2022

MIME: Human-Aware 3D Scene Generation

Generating realistic 3D worlds occupied by moving humans has many applic...
research
05/31/2021

Scene-aware Generative Network for Human Motion Synthesis

We revisit human motion synthesis, a task useful in various real world a...
research
04/04/2023

Generating Continual Human Motion in Diverse 3D Scenes

We introduce a method to synthesize animator guided human motion across ...
research
05/18/2017

Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks

Linking human whole-body motion and natural language is of great interes...
research
08/09/2023

Neural Field Movement Primitives for Joint Modelling of Scenes and Motions

This paper presents a novel Learning from Demonstration (LfD) method tha...
research
05/25/2022

Towards Diverse and Natural Scene-aware 3D Human Motion Synthesis

The ability to synthesize long-term human motion sequences in real-world...
research
03/23/2023

Task-Oriented Human-Object Interactions Generation with Implicit Neural Representations

Digital human motion synthesis is a vibrant research field with applicat...

Please sign up or login with your details

Forgot password? Click here to reset