Papers

Mean normalization of power function based cepstral coefficients for robust speech recognition in noisy environment

International Conference

2011~2015

작성자

한혜원

작성일

2014-05-01 00:42

조회

3638

Authors : Soonho Baek, Hong-Goo Kang

Year : 2014

Publisher / Conference : ICASSP

This paper presents the effect of mean normalization to various types of cepstral coefficients for robust speech recognition in noisy environments. Although the cepstral mean normalization (CMN) technique was originally designed to compensate channel distortion, it has also been proved that the CMN also improves recognition accuracy in additive noisy environment. However, no one has yet considered the interaction of CMN with spectral mapping functions required for extracting cepstral features. This paper investigates the impact of CMN to the speech recognition system depending on the types of spectral mapping function by mathematically analyzing the amount of spectral distortion between clean and noisy conditions. The analytic result is also confirmed by comparing the type of recognition error patterns in automatic speech recognition experiment with Aurora 2 database. Experimental results show that the performance improvement by adopting CMN becomes significant if the logarithmic function is replaced with the appropriate setting of fractional power mapping function. Especially, the deletion errors are dramatically reduced.

« Online Speech Dereverberation Algorithm Based on Adaptive Multichannel Linear Prediction

Selection of spectral compressive operator for vector Taylor series-based model adaptation in noisy environments »

목록보기

전체 372

252	International Conference	Eunwoo Song, Hong-Goo Kang "Multi-class learning algorithm for deep neural network-based statistical parametric speech synthesis" in EUSIPCO, 2016
251	Domestic Conference	Min-jae Hwang, JeeSok Lee, Misuk Lee, and Hong-Goo Kang "사전 분석법을 통한 스프레드 스펙트럼 기반 오디오 워터마킹 알고리즘의 성능 향상" in 한국음향학회 제 33회 음성통신 및 신호처리 학술대회, 2016
250	Domestic Conference	문현기, 박영철, 윤대희 "위상 일치와 가변 지수 감쇄 가중치 방법이 적용된 가상 저음 시스템" in 2016 년 한국방송·미디어공학회 하계학술대회, 2016
249	Domestic Conference	문현기, 박영철, 윤대희 "주파수 종속 믹싱 타임 추정 기법 분석" in 한국음향학회 춘계학술대회, 2016
248	Domestic Conference	서지호, 박영철, 윤대희 "헤드폰/이어폰 환경에서의 제약 최적화 기반의 효율적인 피드백 능동 소음 제어 필터 설계 알고리즘" in 한국음향학회 춘계학술대회, 2016
247	Domestic Conference	박규태, 박영철, 윤대희 "가변 가중치 곡선을 적용한 가상저음시스템" in 한국음향학회 춘계학술대회, 2016
246	Domestic Conference	양해민, 변경근, 강홍구 "RTCP를 이용한 심층 신경망 기반 음질평가 점수 대역 분별 알고리즘" in 한국음향학회 춘계학술대회, 2016
245	Domestic Conference	김글빛, 이진규, 강홍구 "문장종속 화자검증 시스템을 위한 비음수 행렬 분해 기반 잡음 제거" in 한국음향학회 춘계학술대회, 2016
244	Domestic Conference	김진섭, 주영선, 강홍구(연세대학교), 장인선, 안충현(한국전자통신연구원) "음향 모델 성능 개선을 위한 피치 동기화 기반의 DNN-TTS 시스템" in 한국음향학회 춘계학술대회, 2016
243	International Conference	Hyeongi Moon, Gyutae Park, Yeong-cheol Park, Dae Hee Youn "A Phase-Matched Exponential Harmonic Weighting for Improved Sensation of Virtual Bass" in 140th Convention of Audio Engineering Society, pp.9544, 2016

Mean normalization of power function based cepstral coefficients for robust speech recognition in noisy environment

Previous

Sister Lab.

Yonsei University

Academic Website