Transcription System for Semi-Spontaneous Estonian Speech

Alum auml e Tanel · Frontiers in artificial intelligence and applications · 2012

This paper describes a speech-to-text system for semi-spontaneous Estonian speech. The system is trained on about 100 hours of manually transcribed speech and a 300M word text corpus. Compound words are split before building the language model and reconstructed from recognizer output using a hidden event N-gram model. We use a three pass transcription strategy with unsupervised speaker adaptation between individual passes. The system achieves a word error rate of 34.6% on conference speeches and 25.6% on radio talk shows.

Read the paper · More papers on PaperTik