Coarse-to-Fine Volumetric Prediction for Single-Image 3D Human Pose

11/23/2016
by   Georgios Pavlakos, et al.
0

This paper addresses the challenge of 3D human pose estimation from a single color image. Despite the general success of the end-to-end learning paradigm, top performing approaches employ a two-step solution consisting of a Convolutional Network (ConvNet) for 2D joint localization and a subsequent optimization step to recover 3D pose. In this paper, we identify the representation of 3D pose as a critical issue with current ConvNet approaches and make two important contributions towards validating the value of end-to-end learning for this task. First, we propose a fine discretization of the 3D space around the subject and train a ConvNet to predict per voxel likelihoods for each joint. This creates a natural representation for 3D pose and greatly improves performance over the direct regression of joint coordinates. Second, to further improve upon initial estimates, we employ a coarse-to-fine prediction scheme. This step addresses the large dimensionality increase and enables iterative refinement and repeated processing of the image features. The proposed approach outperforms all state-of-the-art methods on standard benchmarks achieving a relative error reduction greater than 30 Additionally, we investigate using our volumetric representation in a related architecture which is suboptimal compared to our end-to-end approach, but is of practical interest, since it enables training when no image with corresponding 3D groundtruth is available, and allows us to present compelling results for in-the-wild images.

READ FULL TEXT

page 3

page 8

research
12/29/2018

Skeleton Transformer Networks: 3D Human Pose and Skinned Mesh from Single RGB Image

In this paper, we present Skeleton Transformer Networks (SkeletonNet), a...
research
10/26/2019

HEMlets Pose: Learning Part-Centric Heatmap Triplets for Accurate 3D Human Pose Estimation

Estimating 3D human pose from a single image is a challenging task. This...
research
03/10/2020

HEMlets PoSh: Learning Part-Centric Heatmap Triplets for 3D Human Pose and Shape Estimation

Estimating 3D human pose from a single image is a challenging task. This...
research
11/28/2016

3D Human Pose Estimation from a Single Image via Distance Matrix Regression

This paper addresses the problem of 3D human pose estimation from a sing...
research
03/22/2018

Deep Pose Consensus Networks

In this paper, we address the problem of estimating a 3D human pose from...
research
06/15/2017

Holistic Planimetric prediction to Local Volumetric prediction for 3D Human Pose Estimation

We propose a novel approach to 3D human pose estimation from a single de...
research
01/28/2018

Joint Voxel and Coordinate Regression for Accurate 3D Facial Landmark Localization

3D face shape is more expressive and viewpoint-consistent than its 2D co...

Please sign up or login with your details

Forgot password? Click here to reset