Advances in Multi-speaker Conversational Speech Recognition and Understanding

Takaaki Hori, Shoko Araki, Tomohiro Nakatani, Atsushi Nakamura · NTT technical review · 2013

Opportunities have been increasing in recent years for ordinary people to use speech recognition technology.For example, we can easily operate smartphones using voice commands.However, attempts to construct a device that can recognize human conversation have produced unsatisfactory results in terms of accuracy and usability because current technology is not designed for this purpose.At NTT Communication Science Laboratories, our goal is to create a new technology for multi-speaker conversational speech recognition and understanding.In this article, we review the technology we have developed and present our meeting analysis system that can accurately recognize who spoke when, what, to whom, and how in meeting situations.

Read the paper · More papers on PaperTik