IEEE Signal Processing Letters

Joint Face Detection and Alignment Using Multitask Cascaded Convolutional Networks

Journal article · 2016 · Cited by 5,295

✓ Free legal copy found

Preprint, hosted by Cornell University (arxiv.org)

This is the authors’ own version from before peer review, so it may differ from the published paper.

Read the free PDF →

Other free copies

Abstract

Face detection and alignment in unconstrained environment are challenging due to various poses, illuminations, and occlusions. Recent studies show that deep learning approaches can achieve impressive performance on these two tasks. In this letter, we propose a deep cascaded multitask framework that exploits the inherent correlation between detection and alignment to boost up their performance. In particular, our framework leverages a cascaded architecture with three stages of carefully designed deep convolutional networks to predict face and landmark location in a coarse-to-fine manner. In addition, we propose a new online hard sample mining strategy that further improves the performance in practice. Our method achieves superior accuracy over the state-of-the-art techniques on the challenging face detection dataset and benchmark and WIDER FACE benchmarks for face detection, and annotated facial landmarks in the wild benchmark for face alignment, while keeps real-time performance.

DOI: 10.1109/lsp.2016.2603342 · Publisher: Institute of Electrical and Electronics Engineers (IEEE)

Guides

Find another paper