IEEE Signal Processing Letters
Joint Face Detection and Alignment Using Multitask Cascaded Convolutional Networks
✓ Free legal copy found
Preprint, hosted by Cornell University (arxiv.org)
This is the authors’ own version from before peer review, so it may differ from the published paper.
Read the free PDF →Other free copies
Abstract
Face detection and alignment in unconstrained environment are challenging due to various poses, illuminations, and occlusions. Recent studies show that deep learning approaches can achieve impressive performance on these two tasks. In this letter, we propose a deep cascaded multitask framework that exploits the inherent correlation between detection and alignment to boost up their performance. In particular, our framework leverages a cascaded architecture with three stages of carefully designed deep convolutional networks to predict face and landmark location in a coarse-to-fine manner. In addition, we propose a new online hard sample mining strategy that further improves the performance in practice. Our method achieves superior accuracy over the state-of-the-art techniques on the challenging face detection dataset and benchmark and WIDER FACE benchmarks for face detection, and annotated facial landmarks in the wild benchmark for face alignment, while keeps real-time performance.
DOI: 10.1109/lsp.2016.2603342 · Publisher: Institute of Electrical and Electronics Engineers (IEEE)