计算机科学
强化学习
判别式
单眼
面子(社会学概念)
一般化
预处理器
特征(语言学)
人工智能
光学(聚焦)
特征提取
计算机视觉
模式识别(心理学)
社会科学
社会学
数学分析
数学
物理
语言学
哲学
光学
作者
Zhengwei Yang,Yange Wang,Lei Ma,Xiangzheng Li
标识
DOI:10.1007/978-3-031-53311-2_14
摘要
3D face reconstruction from monocular outdoor images has long been a challenging problem. Traditional attention network methods that directly regress parameters may suffer from inadequate learning of discriminative features. In this paper, we propose a method called collaborative reinforcement attention module (CRAM). CRAM comprises three major modules: the perception module (PM), the channel selection module (CSM), and the multi-level feature interaction module (MFIM). CRAM leverages contextual information to simultaneously focus on multiple prominent features in facial photos. It employs multi-level and multi-angle feature extraction and fusion techniques to adaptively learn the relationship between facial regions and key feature points. This results in enhanced accuracy in 3D face reconstruction and meticulous dense alignment. Furthermore, to enhance the model’s generalization performance, we introduce a regional noise injection and image composition module (RNICM) as a preprocessing step for sample data which help capture more local details and handle occluded faces, particularly under significant head rotations. Extensive experiments conducted on the AFLW2000-3D and AFLW datasets validate the effectiveness of the proposed approach.
科研通智能强力驱动
Strongly Powered by AbleSci AI