Papers

Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information

International Journal

2006~2010

작성자

이진영

작성일

2010-08-01 14:25

조회

3780

Authors : Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang

Year : 2010

Publisher / Conference : IEEE Transactions on Audio, Speech, and Language Processing

Volume : 18, issue 6

Page : 1332-1340

In this paper, an information theoretic approach to selecting feature frames for speaker recognition systems is proposed. A conventional approach in which the frame shift is fixed to around half of the frame length may not be the best choice, because the characteristics of the speech signal may rapidly change, especially at phonetic boundaries. Experimental results show that the recognition accuracy increases if the frame interval is directly controlled using phonetic information. By applying these results to the well-known fact that the recognition accuracy is directly correlated with the amount of mutual information, this paper suggests a novel feature frame selection method for speaker recognition. Specifically, feature frames are chosen to have minimum-redundancy within selected feature frames, but maximum-relevancy to speaker models. It is verified by experiments that the proposed method produces consistent improvement, especially in a speaker verification system. It is also robust against variations in acoustic environment.

« Performance Analysis of a Class of Single Channel Speech Enhancement Algorithms for Automatic Speech Recognition

A Two-channel Noise Estimator for Speech Enhancement in Highly Non-stationary Environment »

목록보기

전체 372

50	International Journal	Min-Seok Choi, Hong-Goo Kang "Transient noise reduction in speech signal with a modified long-term predictor" in EURASIP Journal on Advances in Signal Processing, vol.141, 2011
49	International Journal	Myung-Suk Song, Cha Zhang, Dinei Florencio, Hong-Goo Kang "An Interactive 3-D Audio System With Loudspeakers" in IEEE Transactions on Multimedia, vol.13, issue 5, pp.844-855, 2011
48	International Journal	Chi-Sang Jung, Hyunson Seo, Hong-Goo Kang "Estimating Redundancy Information of Selected Features in Multi-dimensional Pattern Classification" in Pattern Recognition Letters, vol.32, issue 4, pp.590-596, 2011
47	International Conference	Myung-Suk Song, Cha Zhang, Dinei Florencio, Hong-Goo Kang "Enhancing loudspeaker-based 3D audio with room modeling" in MMSP, 2010
46	International Conference	Chi-Sang Jung, Kyu J. Han, Hyunson Seo, Shrikanth S. Narayanan, Hong-Goo Kang "A Variable Frame Length and Rate Algorithm Based on the Spectral Kurtosis Measure for Speaker Verification" in INTERPSEECH, pp.2754-2757, 2010
45	International Journal	Min-Seok Choi, Hong-Goo Kang "A Two-channel Noise Estimator for Speech Enhancement in Highly Non-stationary Environment" in IEEE Transactions on Audio, Speech, and Language Processing, vol.19, issue 4, pp.905-915, 2011
44	International Journal	Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang "Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information" in IEEE Transactions on Audio, Speech, and Language Processing, vol.18, issue 6, pp.1332-1340, 2010
43	Domestic Journal	Myung-Suk Song, Chang-Heon Lee, Seok-Pil Lee, Hong-Goo Kang "Performance Analysis of a Class of Single Channel Speech Enhancement Algorithms for Automatic Speech Recognition" in 한국음향학회지, vol.29, 제 2호, pp.86-99, 2010
42	International Conference	Myung-Suk Song , Cha Zhang, Dinei Florencio, Hong-Goo Kang "Personal 3D audio system with loudspeakers" in ICME, 2010
41	International Conference	Ho Seon Shin, Min-Seok Choi, Taesu Kim, Hong-Goo Kang "Binaural loudness based speech reinforcement with a closed-form solution" in ICASSP, 2010

Selecting Feature Frames for Automatic Speaker Recognition Using Mutual Information

Previous

Sister Lab.

Yonsei University

Academic Website