CHORE: Contact, Human and Object REconstruction from a single RGB image

04/05/2022
by   Xianghui Xie, et al.
0

While most works in computer vision and learning have focused on perceiving 3D humans from single images in isolation, in this work we focus on capturing 3D humans interacting with objects. The problem is extremely challenging due to heavy occlusions between human and object, diverse interaction types and depth ambiguity. In this paper, we introduce CHORE, a novel method that learns to jointly reconstruct human and object from a single image. CHORE takes inspiration from recent advances in implicit surface learning and classical model-based fitting. We compute a neural reconstruction of human and object represented implicitly with two unsigned distance fields, and additionally predict a correspondence field to a parametric body as well as an object pose field. This allows us to robustly fit a parametric body model and a 3D object template, while reasoning about interactions. Furthermore, prior pixel-aligned implicit learning methods use synthetic data and make assumptions that are not met in real data. We propose a simple yet effective depth-aware scaling that allows more efficient shape learning on real data. Our experiments show that our joint reconstruction learned with the proposed strategy significantly outperforms the SOTA. Our code and models will be released to foster future research in this direction.

READ FULL TEXT

page 2

page 4

page 8

page 10

page 11

page 12

research
07/08/2020

PaMIR: Parametric Model-Conditioned Implicit Representation for Image-based Human Reconstruction

Modeling 3D humans accurately and robustly from a single image is very c...
research
03/29/2023

Visibility Aware Human-Object Interaction Tracking from Single RGB Camera

Capturing the interactions between humans and their environment in 3D is...
research
09/14/2023

HandNeRF: Learning to Reconstruct Hand-Object Interaction Scene from a Single RGB Image

This paper presents a method to learn hand-object interaction prior for ...
research
07/23/2023

LIST: Learning Implicitly from Spatial Transformers for Single-View 3D Reconstruction

Accurate reconstruction of both the geometric and topological details of...
research
11/15/2022

IntegratedPIFu: Integrated Pixel Aligned Implicit Function for Single-view Human Reconstruction

We propose IntegratedPIFu, a new pixel aligned implicit model that build...
research
09/17/2019

Is That a Chair? Imagining Affordances Using Simulations of an Articulated Human Body

For robots to exhibit a high level of intelligence in the real world, th...
research
09/12/2022

Articulated 3D Human-Object Interactions from RGB Videos: An Empirical Analysis of Approaches and Challenges

Human-object interactions with articulated objects are common in everyda...

Please sign up or login with your details

Forgot password? Click here to reset