Speaker-independent phoneme recognition using large-scale neural networks

Satoru Nakamura, H. Sawai, Masashi Sugiyama · 1992

The authors describe a large-scale neural network architecture based on TDNN (time-delay neural networks) for speaker-independent phoneme recognition which represents an advance over speaker-dependent and multi-speaker phoneme recognition. Based on a preliminary study on speaker-independent phoneme recognition for voiced stops mod b,d,g mod , a large-scale network is constructed with about 330000 connections in a modular fashion. For speaker-independent all-consonant recognition, a multi-speaker training approach is implemented with several devices in the process of training. This network finally achieved favorable results for speaker-independent phoneme recognition.>

Read the paper · More papers on PaperTik