Neural network architectures for speaker independent phoneme recognition

Michelle Cutajar, Edward Gatt, Ivan Grech, Owen Casha, Joseph Micallef · International Symposium on Image and Signal Processing and Analysis · 2011

Two different neural network architectures were designed for speaker independent phoneme recognition systems. The first architecture consists of the Radial Basis Function (RBF), while in the second architecture a Self-Organising Maps (SOM) neural network replaces the RBF. The Discrete Wavelet Transform (DWT) is used for feature extraction in both systems. Both systems were tested on the TIMIT database. The highest recognition rates obtained are 36.3% and 46.7%, for the RBF and SOM architectures respectively for multi-speaker unlimited vocabulary speech.

Read the paper · More papers on PaperTik