Speech and Spoken Document
Mari Ostendorf, Benoît Favre, Ralph Grishman, Dilek Zeynep Hakkani-Tür, Mary P. Harper, Dustin Hillard, Julia Hirschberg, Heng Ji, Jeremy G. Kahn, Yang Liu, Sameer Maskey, Evgeny Matusov, Hermann Ney, Andrew Rosenberg, Elizabeth E. Shriberg, Wen Wang, Chuck Wooters · 2010
rogress in both speech and language processing has spurred efforts to sup-port applications that rely on spoken—rather than written—language input.A key challenge in moving from text-based documents to such “spoken doc-uments” is that spoken language lacks explicit punctuation and formatting,which can be crucial for good performance. This article describes differentlevels of speech segmentation, approaches to automatically recovering segment bound-ary locations, and experimental results demonstrating impact on several language pro-cessing tasks. The results also show a need for optimizing segmentation for the endtask rather than independently.