The Vienna Prosodic Speech Corpus: Purpose, Content and Encoding
Friedrich Neubarth, Kai Alter, Hannes Pirker, Elisabeth Rieder, Harald Trost · 2000
This paper presents a corpus of spoken German especially designed for the investigation of prosodic properties of speech. After a short discussion of the content and set-up of the corpus, we describe in detail the additional linguistic information, introduced into the corpus by labelling and annotation. In this project, both qualitative and quantitative methods have been used for the acquisition of data. Our main concern is the development of a well-defined and transparent scheme for the structuring of this heterogeneous information. A second task is to incorporate all these data-generated by different tools with different data-formats- into a single data-base. 1 Motivation The corpus to be presented here is designed to serve multiple purposes in the research of prosodic properties of spoken language. Most of this research is performed within the SpeeDurCont project (Speech Duration in Context-to-Speech, cf. Alter et al., 1998, Pirker et al. 1996), which is an ongoing investigation of durational variation in German speech. The project goals are