MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation

11/24/2021
by   Wenhao Li, et al.
2

Estimating 3D human poses from monocular videos is a challenging task due to depth ambiguity and self-occlusion. Most existing works attempt to solve both issues by exploiting spatial and temporal relationships. However, those works ignore the fact that it is an inverse problem where multiple feasible solutions (i.e., hypotheses) exist. To relieve this limitation, we propose a Multi-Hypothesis Transformer (MHFormer) that learns spatio-temporal representations of multiple plausible pose hypotheses. In order to effectively model multi-hypothesis dependencies and build strong relationships across hypothesis features, the task is decomposed into three stages: (i) Generate multiple initial hypothesis representations; (ii) Model self-hypothesis communication, merge multiple hypotheses into a single converged representation and then partition it into several diverged hypotheses; (iii) Learn cross-hypothesis communication and aggregate the multi-hypothesis features to synthesize the final 3D pose. Through the above processes, the final representation is enhanced and the synthesized pose is much more accurate. Extensive experiments show that MHFormer achieves state-of-the-art results on two challenging datasets: Human3.6M and MPI-INF-3DHP. Without bells and whistles, its performance surpasses the previous best result by a large margin of 3 https://github.com/Vegetebird/MHFormer.

READ FULL TEXT

page 7

page 8

page 13

page 14

page 15

research
04/11/2019

Generating Multiple Hypotheses for 3D Human Pose Estimation with Mixture Density Network

3D human pose estimation from a monocular image or 2D joints is an ill-p...
research
03/21/2023

Diffusion-Based 3D Human Pose Estimation with Multi-Hypothesis Aggregation

In this paper, a novel Diffusion-based 3D Pose estimation (D3DP) method ...
research
06/29/2023

MPM: A Unified 2D-3D Human Pose Representation via Masked Pose Modeling

Estimating 3D human poses only from a 2D human pose sequence is thorough...
research
07/29/2021

Probabilistic Monocular 3D Human Pose Estimation with Normalizing Flows

3D human pose estimation from monocular images is a highly ill-posed pro...
research
11/29/2022

DiffPose: Multi-hypothesis Human Pose Estimation using Diffusion models

Traditionally, monocular 3D human pose estimation employs a machine lear...
research
02/08/2017

Generating Multiple Diverse Hypotheses for Human 3D Pose Consistent with 2D Joint Detections

We propose a method to generate multiple diverse and valid human pose hy...
research
12/07/2016

Global Hypothesis Generation for 6D Object Pose Estimation

This paper addresses the task of estimating the 6D pose of a known 3D ob...

Please sign up or login with your details

Forgot password? Click here to reset