First International Symposium on 3D Data Processing Visualization and Transmission (3DPVT'02) Real-Time Speech-Driven 3D Face Animation Padova, Italy June 19-June 21 ISBN: 0-7695-1521-5
In this paper,we present an approach for real-time speech-driven 3D face animation using neural networks. We first analyze a 3D facial movement sequence of a talking subject and learn a quantitative representation of the facial deformations, called the 3D Motion Units (MUs). A 3D facial deformation can be approximated by a linear combination of the MUs weighted by the MU parameters (MUPs) — the visual features of the facial deformation. The facial movement sequence synchronizes with a audio track. The audio track is digitized and the audio features of each frame are calculated. A real-time audio-to-MUP mapping is constructed by training a set of neural networks using the calculated audio-visual features. The audio-visual features are divided into several groups based on the audio features. One neural network is trained per group to map the audio features to the corresponding MUPs. Given a new audio feature vector, we first classify it into one of the groups and select the corresponding neural network to map the audio feature vector to MUPs, which are used for face animation. The quantitative evaluation shows the effectiveness of the proposed approach.
Citation:
Pengyu Hong, Zhen Wen, Thomas S. Huang, Heung-Yeung Shum, "Real-Time Speech-Driven 3D Face Animation," 3dpvt, pp.713, First International Symposium on 3D Data Processing Visualization and Transmission (3DPVT'02), 2002 Usage of this product signifies your acceptance of the Terms of Use. | |||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||