Papers

Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information

International Journal

2006~2010

작성자

이진영

작성일

2010-08-01 14:25

조회

3158

Authors : Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang

Year : 2010

Publisher / Conference : IEEE Transactions on Audio, Speech, and Language Processing

Volume : 18, issue 6

Page : 1332-1340

In this paper, an information theoretic approach to selecting feature frames for speaker recognition systems is proposed. A conventional approach in which the frame shift is fixed to around half of the frame length may not be the best choice, because the characteristics of the speech signal may rapidly change, especially at phonetic boundaries. Experimental results show that the recognition accuracy increases if the frame interval is directly controlled using phonetic information. By applying these results to the well-known fact that the recognition accuracy is directly correlated with the amount of mutual information, this paper suggests a novel feature frame selection method for speaker recognition. Specifically, feature frames are chosen to have minimum-redundancy within selected feature frames, but maximum-relevancy to speaker models. It is verified by experiments that the proposed method produces consistent improvement, especially in a speaker verification system. It is also robust against variations in acoustic environment.

« On the Importance of Transition Regions for Automatic Speaker Recognition

Joint Channel Coding Based on Principal Component Analysis »

목록보기

전체 371

21	International Journal	Dong-il Hyun, Donggeum Lee, Youngcheol Park, Dae Hee Youn, Jeongil Seo "Joint Channel Coding Based on Principal Component Analysis" in ETRI Journal, vol.32, issue 5, pp.831-834, 2010
20	International Journal	Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang "Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information" in IEEE Transactions on Audio, Speech, and Language Processing, vol.18, issue 6, pp.1332-1340, 2010
19	International Journal	Bong-Jin Lee, Chi-Sang Jung, Jeung-Yoon Choi, Hong-Goo Kang "On the Importance of Transition Regions for Automatic Speaker Recognition" in IEICE Transactions on Information and Systems, vol.E93-D, No.1, pp.197-200, 2010
18	International Journal	Jae-Seong Lee, Chang-Joon Lee, Young-Cheol Park, Dae Hee Youn "Efficient FFT Algorithm for Psychoacoustic Model of the MPEG-4 AAC" in IEICE Transactions on Information and Systems, vol.E92-D, No.12, pp.2535-2539, 2009
17	International Journal	Chang-Heon Lee, Hyen-O Oh, Hong-Goo Kang "On the Study of Noise Allocation for Speech Signal in Low Bit-Rate Audio Coding" in IEEE Signal Processing Letters, vol.16, issue 10, pp.849-852, 2009
16	International Journal	Bong-Jin Lee, Jeung-Yoon Choi, Hong-Goo Kang "Phonetically optimized speaker modeling for robust speaker recognition" in The Journal of the Acoustical Society of America, vol.126, issue 3, 2009
15	International Journal	Tacksung Choi, Sunkuk Moon, Young-Cheol Park, Dea Hee Youn, Seokpil Lee "A GMM-Based Feature Selection Algorithm for Multi-Class Classification" in IEICE Transactions on Information and Systems, vol.E92-D. No.8, pp.1584-1587, 2009
14	International Journal	Jung-Won Lee, Jeung-Yoon Choi "Acoustic‐phonetic features for stop consonant place detection in clean and telephone speech" in The Journal of the Acoustical Society of America, vol.123, issue 5, 2008
13	International Journal	Jungin Lee, Jeung-Yoon Choi "Detection of obstruent consonant landmark for knowledge based speech recognition" in The Journal of the Acoustical Society of America, vol.123, issue 8, 2008
12	International Journal	Sukmyung Lee, Jeung-Yoon Choi "Vowel place detection for a knowledge‐based speech recognition system" in The Journal of the Acoustical Society of America, vol.123, issue 5, 2008

Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information

Previous

Sister Lab.

Yonsei University

Academic Website