人工智能
计算机科学
姿势
增强现实
深度学习
软件部署
卷积神经网络
虚拟现实
计算机视觉
桥(图论)
构造(python库)
个性化
程序设计语言
万维网
医学
操作系统
内科学
作者
Ting Chou,Chih‐Hsing Chu,Shengjun Liu
摘要
Abstract Customization is an increasing trend in fashion product industry to reflect individual lifestyles. Previous studies have examined the idea of virtual footwear try-on in augmented reality (AR) using a depth camera. However, the depth camera restricts the deployment of this technology in practice. This research proposes to estimate the six degrees-of-freedom pose of a human foot from a color image using deep learning models to solve the problem. We construct a training dataset consisting of synthetic and real foot images that are automatically annotated. Three convolutional neural network models (deep object pose estimation (DOPE), DOPE2, and You Only Look Once (YOLO)-6D) are trained with the dataset to predict the foot pose in real-time. The model performances are evaluated using metrics for accuracy, computational efficiency, and training time. A prototyping system implementing the best model demonstrates the feasibility of virtual footwear try-on using a red–green–blue camera. Test results also indicate the necessity of real training data to bridge the reality gap in estimating the human foot pose.
科研通智能强力驱动
Strongly Powered by AbleSci AI