A two stage hybrid embedded speech/audio coding structure

Sean A. Ramprashad · 2002

A two stage hybrid embedded speech/audio coding structure is proposed. The structure uses a speech coder as a core to provide the minimal bitrate and an acceptable performance on speech inputs. The second stage is a transform coder using a modified discrete cosine transform (MDCT) and perceptual coding principles. This stage is itself embedded both in complexity and bitrate, and provides various levels of enhancement of the core output, particularly for general audio signals like music. Informal A-B comparison tests show that the performance of the structure at 16 kb/s is between that of the GSM enhanced full rate coder at 12.2 kb/s, and the G.728 LD-CELP coder at 16 kb/s.

Read the paper · More papers on PaperTik