Fast Bit Allocation Method for MP3/AAC Encoders

International Conference
2005-05-28 22:09
Authors : Kyoung Ho Bang, Keun-Sup Lee, Young-Cheol Park, Dae Hee Youn

Year : 2005

Publisher / Conference : 118th Convention of Audio Engineering Society

In MP3/AAC encoders, the quantization parameter called scalefactor controls the quantization noise and the bitrate. Tuning these encoders would require a characterization of the rate-distortion function per subband, which seems to be available only in a parametric manner. In this paper, a fast bit allocation method for MP3/AAC encoder is presented. The resulting encoder is able to produce an ISO/MPEG compliant bitstream which can guarantee better audio quality. More importantly, the number of computational steps is greatly reduced as compared to the method recommended by the ISO/MPEG committee because the efficient bit allocation algorithm significantly reduces the number of iterations required. It was found that the efficient bit allocation algorithm works best when the bit rate demanded by the psychoacoustic model in order to keep the quantization noise below the masking threshold is almost equal to the operational bit rate.
전체 322
322 International Conference Doyeon Kim, Hyewon Han, Hyeon-Kyeong Shin, Soo-Whan Chung, Hong-Goo Kang "Phase Continuity: Learning Derivatives of Phase Spectrum for Speech Enhancement" in ICASSP, 2022
321 International Conference Chanwoo Lee, Hyungseob Lim, Jihyun Lee, Inseon Jang, Hong-Goo Kang "Progressive Multi-Stage Neural Audio Coding with Guided References" in ICASSP, 2022
320 International Conference Jihyun Lee, Hyungseob Lim, Chanwoo Lee, Inseon Jang, Hong-Goo Kang "Adversarial Audio Synthesis Using a Harmonic-Percussive Discriminator" in ICASSP, 2022
319 International Conference Jinyoung Lee and Hong-Goo Kang "Stacked U-Net with High-level Feature Transfer for Parameter Efficient Speech Enhancement" in APSIPA ASC, 2021
318 International Conference Huu-Kim Nguyen, Kihyuk Jeong, Se-Yun Um, Min-Jae Hwang, Eunwoo Song, Hong-Goo Kang "LiteTTS: A Decoder-free Light-weight Text-to-wave Synthesis Based on Generative Adversarial Networks" in INTERSPEECH, 2021
317 International Conference Zainab Alhakeem, Yoohwan Kwon, Hong-Goo Kang "Disentangled Representations for Arabic Dialect Identification based on Supervised Clustering with Triplet Loss" in EUSIPCO, 2021
316 International Conference Miseul Kim, Minh-Tri Ho, Hong-Goo Kang "Self-supervised Complex Network for Machine Sound Anomaly Detection" in EUSIPCO, 2021
315 International Conference Kihyuk Jeong, Huu-Kim Nguyen, Hong-Goo Kang "A Fast and Lightweight Text-To-Speech Model with Spectrum and Waveform Alignment Algorithms" in EUSIPCO, 2021
314 International Conference Jiyoung Lee*, Soo-Whan Chung*, Sunok Kim, Hong-Goo Kang**, Kwanghoon Sohn** "Looking into Your Speech: Learning Cross-modal Affinity for Audio-visual Speech Separation" in CVPR, 2021
313 International Conference Zainab Alhakeem, Hong-Goo Kang "Confidence Learning from Noisy Labels for Arabic Dialect Identification" in ITC-CSCC, 2021