论文交流 >  机器视觉

Face Detection with End-to-End Integration of a ConvNet and a 3D Model

利用端到端整合的ConvNet和3D Model实现人脸检测

1年前 822 1  点赞 (0)  收藏 (0)

研究领域: 机器视觉   arXiv2016

应用方向: 人脸识别

原理方法:

软件实现:

论文摘要:

This paper presents a method for face detection in the wild, which integrates a ConvNet and a 3D mean face model in an end-to-end multi-task discriminative learning framework. The 3D mean face model is predefined and fixed (e.g., we used the one provided in the AFLW dataset). The ConvNet consists of two components: (i) The face pro- posal component computes face bounding box proposals via estimating facial key-points and the 3D transformation (rotation and translation) parameters for each predicted key-point w.r.t. the 3D mean face model. (ii) The face verification component computes detection results by prun- ing and refining proposals based on facial key-points based configuration pooling. The proposed method addresses two issues in adapting state- of-the-art generic object detection ConvNets (e.g., faster R-CNN) for face detection: (i) One is to eliminate the heuristic design of prede- fined anchor boxes in the region proposals network (RPN) by exploit- ing a 3D mean face model. (ii) The other is to replace the generic RoI (Region-of-Interest) pooling layer with a configuration pooling layer to respect underlying object structures. The multi-task loss consists of three terms: the classification Softmax loss and the location smooth l1 -losses [14] of both the facial key-points and the face bounding boxes. In ex- periments, our ConvNet is trained on the AFLW dataset only and tested on the FDDB benchmark with fine-tuning and on the AFW benchmark without fine-tuning. The proposed method obtains very competitive state-of-the-art performance in the two benchmarks.

论文精要:

论文点评 

您可以在评论中对论文进行“摘要翻译” “标签备注” “精要点评” “疑难提问”,我们会及时更新到数据库中!

任何论文都是 "在特定领域内"、"基于某种学术原理"、"研究某个应用问题", 因此分 领域标签 / 应用标签 / 原理标签 / 补充标签

金元宝 10个月前

非常好

(0) 回复