loading...
 This Article 
   
 Share 
   
 Bibliographic References 
   
 Add to: 
 
Digg
Furl
Spurl
Blink
Simpy
Google
Del.icio.us
Y!MyWeb
 
 Search 
   
First International Symposium on 3D Data Processing Visualization and Transmission (3DPVT'02)
Real-Time Speech-Driven 3D Face Animation
Padova, Italy
June 19-June 21
ISBN: 0-7695-1521-5
Pengyu Hong, University of Illinois at Urban-Champaign
Zhen Wen, University of Illinois at Urban-Champaign
Thomas S. Huang, University of Illinois at Urban-Champaign
Heung-Yeung Shum, Microsoft Research Lab -Beijing
In this paper,we present an approach for real-time speech-driven 3D face animation using neural networks. We first analyze a 3D facial movement sequence of a talking subject and learn a quantitative representation of the facial deformations, called the 3D Motion Units (MUs). A 3D facial deformation can be approximated by a linear combination of the MUs weighted by the MU parameters (MUPs) — the visual features of the facial deformation. The facial movement sequence synchronizes with a audio track. The audio track is digitized and the audio features of each frame are calculated. A real-time audio-to-MUP mapping is constructed by training a set of neural networks using the calculated audio-visual features. The audio-visual features are divided into several groups based on the audio features. One neural network is trained per group to map the audio features to the corresponding MUPs. Given a new audio feature vector, we first classify it into one of the groups and select the corresponding neural network to map the audio feature vector to MUPs, which are used for face animation. The quantitative evaluation shows the effectiveness of the proposed approach.
Citation:
Pengyu Hong, Zhen Wen, Thomas S. Huang, Heung-Yeung Shum, "Real-Time Speech-Driven 3D Face Animation," 3dpvt, pp.713, First International Symposium on 3D Data Processing Visualization and Transmission (3DPVT'02), 2002
Usage of this product signifies your acceptance of the Terms of Use.