The buckeye corpus of speech: updates and enhancements

Eric Fosler‐Lussier, Laura C. Dilley, Na’im R. Tyson, Mark A. Pitt · 2007

This paper describes recent progress in the development of the Buckeye Corpus of Speech, a phonetically labeled corpus of conversational American English speech, first described in [1]. With the publication of the second phase of transcription, the corpus has nearly doubled in size from the first release. We briefly give an overview of the corpus, report on additional stud-ies of inter-labeler agreement, and describe a new GUI designed to facilitate searching the annotated speech files. Index Terms: corpora, transcription, phonetics, search tool 1.

Read the paper · More papers on PaperTik