MPEG audio bit rate scaling on coded data domain
Yasuyuki Nakajima, Hiromasa Yanagihara, Akio Yoneyama, Masaru Sugano · 2002
Formerly, once the audio data is compressed, transcoding is used to scale the bit rate, where decoding and re-encoding are taking place. Therefore, data manipulation of coded data has been very complex and time consuming work. We describe three algorithms for bit rate scaling in the coded MPEG data domain. One is a bandwidth limitation method cutting higher frequency components until the target data rate is satisfied. The other two use a re-quantization process where a quantization step in each subband is modified. One of them reflects the psychoacoustic model from bit allocation information obtained in the bitstream in order to improve the bit rate scaling efficiency. The simulation results show that the re-quantization process provides a very high conversion efficiency and a nearly equal sound quality can be obtained as directly coding from PCM by reflecting the psychoacoustic model. It is also shown that a very fast scaling (factor of six) have been achieved when compared with the transcoding method.