Text-driven automatic frame generation using MPEG-4 synthetic/natural hybrid coding for 2-D head-and-shoulder scene
Terence Chun-Ho Cheung, Lai-Man Po · 2002
In this paper, we propose a facial modeling technique based on the MPEG-4 synthetic/natural hybrid coding for automating frame sequence generation of a talking head. With the definition and animation parameters on a generic face object, the shape, textures and expressions of an adapted frontal face can generally be controlled and synchronized by the phonemes transcribed from plain text. By this developed facial modeling technique, it increases the intelligibility of an non-verbal facial communication for potential audiovisual lip-synch application on news reporting, lip-reading for the hearing-impaired or the deaf, virtual meeting through Internet, and story teller on demand (STOD).