Optimizing sentence segmentation for spoken language translation

Sharath Rao, Ian Richard Lane, Tanja Schultz · 2007

The conventional approach in text-based machine translation (MT) is to translate complete sentences, which are conveniently indicated by sentence boundary markers.However, since such boundary markers are not available for speech, new methods are required that define an optimal unit for translation.Our experimental results show that with a segment length optimized for a particular MT system, intrasentence segmentation can improve translation performance (measured in BLEU) by up to 11% for Arabic Broadcast Conversation (BC) and 6% for Arabic Broadcast News (BN).We show that acoustic segmentation that minimizes Word Error Rate (WER) may not give the best translation performance.We improve upon it by automatically resegmenting the ASR output in a way that is optimized for translation and argue that it might be necessary for different stages of a Spoken Language Translation (SLT) system to define their own optimal units.

Read the paper · More papers on PaperTik