On the role of depth predictions for 3D human pose estimation

Diaz-Arias, Alec; Messmore, Mitchell; Shin, Dmitriy; Baek, Stephen

Computer Science > Computer Vision and Pattern Recognition

arXiv:2103.02521 (cs)

[Submitted on 3 Mar 2021]

Title:On the role of depth predictions for 3D human pose estimation

Authors:Alec Diaz-Arias, Mitchell Messmore, Dmitriy Shin, Stephen Baek

View PDF

Abstract:Following the successful application of deep convolutional neural networks to 2d human pose estimation, the next logical problem to solve is 3d human pose estimation from monocular images. While previous solutions have shown some success, they do not fully utilize the depth information from the 2d inputs. With the goal of addressing this depth ambiguity, we build a system that takes 2d joint locations as input along with their estimated depth value and predicts their 3d positions in camera coordinates. Given the inherent noise and inaccuracy from estimating depth maps from monocular images, we perform an extensive statistical analysis showing that given this noise there is still a statistically significant correlation between the predicted depth values and the third coordinate of camera coordinates. We further explain how the state-of-the-art results we achieve on the H3.6M validation set are due to the additional input of depth. Notably, our results are produced on neural network that accepts a low dimensional input and be integrated into a real-time system. Furthermore, our system can be combined with an off-the-shelf 2d pose detector and a depth map predictor to perform 3d pose estimation in the wild.

Comments:	13 pages, 6 figures, and 8 tables
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2103.02521 [cs.CV]
	(or arXiv:2103.02521v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2103.02521

Submission history

From: Alec Diaz-Arias [view email]
[v1] Wed, 3 Mar 2021 16:51:38 UTC (5,101 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:On the role of depth predictions for 3D human pose estimation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:On the role of depth predictions for 3D human pose estimation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators