Robust facial landmark extraction scheme using multiple convolutional neural networks

Citations

WEB OF SCIENCE

13
Citations

SCOPUS

12

초록

Facial landmarks are a set of features that can be distinguished on the human face with the naked eye. Typical facial landmarks include eyes, eyebrows, nose, and mouth. Landmarks play an important role in human-related image analysis. For example, they can be used to determine whether there is a human being in the image, identify who the person is, or recognize the orientation of a face when taking a photograph. General techniques for detecting facial landmarks can be classified into two groups: One is based on traditional image processing techniques, such as Haar cascade classifiers and edge detection. The other is based on machine learning techniques in which landmarks can be detected by training neural network using facial features. However, such techniques have shown low accuracy, especially in some special conditions such as low luminance and overlapped faces. To overcome these problems, we proposed in our previous work a facial landmark extraction scheme using deep learning and semantic segmentation, and demonstrated that with even a small dataset, our scheme could achieve reasonable facial landmark extraction performance under such conditions. Nevertheless, for more extensive dataset, we found several exceptional cases where the scheme could not detect face landmarks precisely. Hence, in this paper, we revise our facial landmark extraction scheme using a deep learning model called Faster R-CNN and show how our scheme can improve the overall performance by handling such exceptional cases appropriately. Also, we show how to expand the training dataset by using image filters and image operations such as rotation for more robust landmark detection.

키워드

Convolutional neural networksFacial landmarkSemantic segmentationObject detectionFaster R-CNN
제목
Robust facial landmark extraction scheme using multiple convolutional neural networks
저자
Kim, HyungjoonPark, JisooKim, HyeonWooHwang, EenjunRho, Seungmin
DOI
10.1007/s11042-018-6482-7
발행일
2019-02
유형
Article
저널명
Multimedia Systems
78
3
페이지
3221 ~ 3238