Papers

Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information

International Journal
2006~2010
작성자
이진영
작성일
2010-08-01 14:25
조회
3507
Authors : Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang

Year : 2010

Publisher / Conference : IEEE Transactions on Audio, Speech, and Language Processing

Volume : 18, issue 6

Page : 1332-1340

In this paper, an information theoretic approach to selecting feature frames for speaker recognition systems is proposed. A conventional approach in which the frame shift is fixed to around half of the frame length may not be the best choice, because the characteristics of the speech signal may rapidly change, especially at phonetic boundaries. Experimental results show that the recognition accuracy increases if the frame interval is directly controlled using phonetic information. By applying these results to the well-known fact that the recognition accuracy is directly correlated with the amount of mutual information, this paper suggests a novel feature frame selection method for speaker recognition. Specifically, feature frames are chosen to have minimum-redundancy within selected feature frames, but maximum-relevancy to speaker models. It is verified by experiments that the proposed method produces consistent improvement, especially in a speaker verification system. It is also robust against variations in acoustic environment.
전체 371
191 International Conference Jinkyu Lee, Soonho Baek, Hong-Goo Kang "Signal and feature domain enhancement approaches for robust speech recognition" in 8th International Conference on Information, communications and Signal Processing, 2011
190 International Journal Min-Seok Choi, Hong-Goo Kang "Transient noise reduction in speech signal with a modified long-term predictor" in EURASIP Journal on Advances in Signal Processing, vol.141, 2011
189 International Journal Myung-Suk Song, Cha Zhang, Dinei Florencio, Hong-Goo Kang "An Interactive 3-D Audio System With Loudspeakers" in IEEE Transactions on Multimedia, vol.13, issue 5, pp.844-855, 2011
188 International Conference Dong-il Hyun, Young-cheol Park, Seok-pil Lee, Dae Hee Youn "Enhanced Interchannel Correlation (ICC) Synthesis for Spatial Audio Coding" in AES 43th International Conference, 2011
187 International Journal Tacksung Choi, Young-Cheol Park, Dae Hee Youn, Seokpil Lee "Virtual Sound Rendering in a Stereophonic Loudspeaker Setup" in IEEE Transactions on Audio, Speech, and Language Processing, vol.19, issue 7, pp.1962-1974, 2011
186 International Conference Se-Woon Jeon, Young-cheol Park, Seok-Pil Lee, Dae Hee Youn "Virtual Source Panning using Multiple-Wise Vector Base in the Multispeaker Stereo Format" in EUSIPCO, pp.1337-1341, 2011
185 Domestic Conference 신호선,양재모,강홍구 "GSC 빔포머의 실시간 구현을 위한 고정 소수점 연산화" in 한국음향학회, 2011
184 Domestic Conference 신호선, 양재모, 강홍구 "GSC 빔포머의 실시간 구현을 위한 고정 소수점 연산화" in 제 28회 음성통신 및 신호처리 학술대회, vol.28, no.1, 2011
183 Domestic Conference 송은우, 강홍구 "Improved Single Channel speech Enhancement Algorithm Using Adaptive Postfiltering" in 한국방송공학회, 2011
182 Domestic Conference 이진규, 서현선, 강홍구 "A Single Channel Speech Enhancement for Automatic speech Recognition" in 한국방송공학회, 2011