On the power of data augmentation for head pose estimation

2024-07-07Code Available1· sign in to hype

Michael Welter

Code Available — Be the first to reproduce this paper.

Code

github.com/opentrack/neuralnet-tracker-traincode
OfficialIn paperpytorch★ 36

Abstract

Deep learning has been impressively successful in the last decade in predicting human head poses from monocular images. However, for in-the-wild inputs the research community relies predominantly on a single training set, 300W-LP, of semisynthetic nature without many alternatives. This paper focuses on gradual extension and improvement of the data to explore the performance achievable with augmentation and synthesis strategies further. Modeling-wise a novel multitask head/loss design which includes uncertainty estimation is proposed. Overall, the thus obtained models are small, efficient, suitable for full 6 DoF pose estimation, and exhibit very competitive accuracy.

Tasks

Data Augmentation Face Alignment Head Pose Estimation Pose Estimation

On the power of data augmentation for head pose estimation

Code

Abstract

Tasks

Reproductions