Multi-initialization Optimization Network for Accurate 3D Human Pose and Shape Estimation

12/24/2021
by   Zhiwei Liu, et al.
0

3D human pose and shape recovery from a monocular RGB image is a challenging task. Existing learning based methods highly depend on weak supervision signals, e.g. 2D and 3D joint location, due to the lack of in-the-wild paired 3D supervision. However, considering the 2D-to-3D ambiguities existed in these weak supervision labels, the network is easy to get stuck in local optima when trained with such labels. In this paper, we reduce the ambituity by optimizing multiple initializations. Specifically, we propose a three-stage framework named Multi-Initialization Optimization Network (MION). In the first stage, we strategically select different coarse 3D reconstruction candidates which are compatible with the 2D keypoints of input sample. Each coarse reconstruction can be regarded as an initialization leads to one optimization branch. In the second stage, we design a mesh refinement transformer (MRT) to respectively refine each coarse reconstruction result via a self-attention mechanism. Finally, a Consistency Estimation Network (CEN) is proposed to find the best result from mutiple candidates by evaluating if the visual evidence in RGB image matches a given 3D reconstruction. Experiments demonstrate that our Multi-Initialization Optimization Network outperforms existing 3D mesh based methods on multiple public benchmarks.

READ FULL TEXT

page 2

page 4

research
01/31/2023

A Modular Multi-stage Lightweight Graph Transformer Network for Human Pose and Shape Estimation from 2D Human Pose

In this research, we address the challenge faced by existing deep learni...
research
12/17/2020

End-to-End Human Pose and Mesh Reconstruction with Transformers

We present a new method, called MEsh TRansfOrmer (METRO), to reconstruct...
research
04/29/2023

TAPE: Temporal Attention-based Probabilistic human pose and shape Estimation

Reconstructing 3D human pose and shape from monocular videos is a well-s...
research
02/15/2023

Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose Estimation

There has been a recent surge of interest in introducing transformers to...
research
03/15/2023

Mesh Strikes Back: Fast and Efficient Human Reconstruction from RGB videos

Human reconstruction and synthesis from monocular RGB videos is a challe...
research
02/18/2021

HandTailor: Towards High-Precision Monocular 3D Hand Recovery

3D hand pose estimation and shape recovery are challenging tasks in comp...
research
10/24/2022

Multi-Person 3D Pose and Shape Estimation via Inverse Kinematics and Refinement

Estimating 3D poses and shapes in the form of meshes from monocular RGB ...

Please sign up or login with your details

Forgot password? Click here to reset