Speech Bandwidth Extension using Temporal Envelope Modeling

International Journal
2008-05-01 14:20
Authors : Kyung-Tae Kim, Min-Ki Lee, Hong-Goo Kang

Year : 2008

Publisher / Conference : IEEE Signal Processing Letters

Volume : 15

Page : 429-432

Speech bandwidth extension (SBE) assumes that high-frequency components of a speech signal, e.g., the frequency band of 4-7 kHz, can be estimated by parameters extracted from the narrowband signal (0-4 kHz). Therefore, it is very important to understand the characteristics of the highband signal as well as perceptual cues to represent the highband signal. This letter proposes a new SBE algorithm using a temporal envelope model. The temporal envelope model considers band-limited temporal envelopes as the perceptual cue of the 4-7 kHz band signal while it deemphasizes the importance of rapidly varying components. To implement the SBE with no additional bits, the proposed method adopts a Gaussian mixture model (GMM) to estimate the temporal envelope of the highband signal from that of the narrowband one. Simulation results confirm that the proposed SBE algorithm shows better perceptual quality than a conventional source-filter model-based approach.
전체 355
21 International Journal Dong-il Hyun, Donggeum Lee, Youngcheol Park, Dae Hee Youn, Jeongil Seo "Joint Channel Coding Based on Principal Component Analysis" in ETRI Journal, vol.32, issue 5, pp.831-834, 2010
20 International Journal Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang "Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information" in IEEE Transactions on Audio, Speech, and Language Processing, vol.18, issue 6, pp.1332-1340, 2010
19 International Journal Bong-Jin Lee, Chi-Sang Jung, Jeung-Yoon Choi, Hong-Goo Kang "On the Importance of Transition Regions for Automatic Speaker Recognition" in IEICE Transactions on Information and Systems, vol.E93-D, No.1, pp.197-200, 2010
18 International Journal Jae-Seong Lee, Chang-Joon Lee, Young-Cheol Park, Dae Hee Youn "Efficient FFT Algorithm for Psychoacoustic Model of the MPEG-4 AAC" in IEICE Transactions on Information and Systems, vol.E92-D, No.12, pp.2535-2539, 2009
17 International Journal Chang-Heon Lee, Hyen-O Oh, Hong-Goo Kang "On the Study of Noise Allocation for Speech Signal in Low Bit-Rate Audio Coding" in IEEE Signal Processing Letters, vol.16, issue 10, pp.849-852, 2009
16 International Journal Bong-Jin Lee, Jeung-Yoon Choi, Hong-Goo Kang "Phonetically optimized speaker modeling for robust speaker recognition" in The Journal of the Acoustical Society of America, vol.126, issue 3, 2009
15 International Journal Tacksung Choi, Sunkuk Moon, Young-Cheol Park, Dea Hee Youn, Seokpil Lee "A GMM-Based Feature Selection Algorithm for Multi-Class Classification" in IEICE Transactions on Information and Systems, vol.E92-D. No.8, pp.1584-1587, 2009
14 International Journal Jung-Won Lee, Jeung-Yoon Choi "Acoustic‐phonetic features for stop consonant place detection in clean and telephone speech" in The Journal of the Acoustical Society of America, vol.123, issue 5, 2008
13 International Journal Jungin Lee, Jeung-Yoon Choi "Detection of obstruent consonant landmark for knowledge based speech recognition" in The Journal of the Acoustical Society of America, vol.123, issue 8, 2008
12 International Journal Sukmyung Lee, Jeung-Yoon Choi "Vowel place detection for a knowledge‐based speech recognition system" in The Journal of the Acoustical Society of America, vol.123, issue 5, 2008