The audio-video australian English speech data corpus AVOZES

J. Bruce Millar, Roland Goecke · 2004

This paper presents the Audio-Video Australian English Speech data corpus AVOZES. It contains recordings of 20 speakers ut-tering a variety of phrases. The corpus was designed for re-search on the statistical relationship of audio and video speech parameters with an audio-video (AV) automatic speech recog-nition (ASR) task in mind, but may be useful for other research tasks. AVOZES is the first published AV speaking-face data corpus for Australian English and is novel in its use of a stereo camera system for the video recordings and its modular design. 1.

Read the paper · More papers on PaperTik