阿凡达
代码本
计算机科学
特征(语言学)
人机交互
人工智能
人体模型
语言学
哲学
作者
Chaoqun Gong,Yuqin Dai,Ronghui Li,Achun Bao,Jun Li,Jian Yang,Yachao Zhang,Xiu Li
出处
期刊:
日期:2024-03-18
卷期号:: 16-20
被引量:1
标识
DOI:10.1109/icassp48485.2024.10446237
摘要
Generating 3D human models directly from text helps reduce the cost and time of character modeling. However, achieving multi-attribute controllable and realistic 3D human avatar generation is still challenging due to feature coupling and the scarcity of realistic 3D human avatar datasets. To address these issues, we propose Text2Avatar, which can generate realistic-style 3D avatars based on the coupled text prompts. Text2Avatar leverages a discrete codebook as an intermediate feature to establish a connection between text and avatars, enabling the disentanglement of features. Furthermore, to alleviate the scarcity of realistic style 3D human avatar data, we utilize a pre-trained unconditional 3D human avatar generation model to obtain a large amount of 3D avatar pseudo data, which allows Text2Avatar to achieve realistic style generation. Experimental results demonstrate that our method can generate realistic 3D avatars from coupled textual data, which is challenging for other existing methods in this field.
科研通智能强力驱动
Strongly Powered by AbleSci AI