Speech processing technologies and telecommunications applications at NTT
Noboru Sugamura, Tomohisa Hirokawa, Shigeki Sagayama, Sadaoki Furui · 2002
The paper describes major research and development in speech recognition and synthesis technologies at NTT from the telecommunications applications viewpoint. Technologies include speaker-dependent, speaker-independent word recognition based on DP matching, speaker-independent word spotting based on HMM, large vocabulary speaker-independent continuous speech recognition based on HMM-LR and high-quality Japanese text-to-speech synthesis. A commercial ANSER system that uses speech recognition and synthesis technologies is also introduced.>