Journal «Language & Science» UTMN.


№8 2019. 05.00.00 ТЕХНИЧЕСКИЕ НАУКИ


About the authors:

Bezhenar Aleksandr Vasilevich, Bachelor student, University of Tyumen,
Sizova Lyudmila Vladimirovna, Senior lecturer, University of Tyumen,


Monocular head pose estimation requires learning a model that computes the intrinsic Euler angles for pose (yaw, pitch, roll) from an input image of human face. Annotating ground truth head pose angles for images in the wild is difficult and requires ad-hoc fitting procedures. This highlights the need for approaches which can train on data captured in controlled environment and generalize on the images in the wild (with varying appearance and illumination of the face). The authors of the article propose to use a higher level representation to regress the head pose while using deep learning architectures. More specifically, they use the uncertainty maps in the form of 2D soft localization heatmap images over five facial key points, namely left ear, right ear, left eye, right eye and nose, and pass them through a convolutional neural network to regress the head-pose. The authors show head pose estimation results on two challenging benchmarks BIWI and AFLW.


