A Priori SNR Estimation Using Air- and Bone-Conduction Microphones

International Journal
2015-11-01 22:11
Authors : Ho Seon Shin, Tim Fingscheidt, Hong-Goo Kang

Year : 2015

Publisher / Conference : IEEE/ACM Transactions on Audio, Speech, and Language Processing

Volume : 23, issue 11

Page : 2015-2025

This paper proposes an a priori signal-to-noise ratio (SNR) estimator using an air-conduction (AC) and a bone-conduction (BC) microphone. Among various ways of combining AC and BC microphones for speech enhancement, it is shown that the total enhancement performance can be maximized if the BC microphone is utilized for estimating the power spectral density (PSD) of the desired speech signal. Considering the fact that a small deviation in the speech PSD estimation process brings severe spectral distortion, this paper focuses on controlling weighting factors while estimating the a priori SNR with the decision-directed approach framework. The time-frequency varying weighting factor that is determined by taking a minimum mean square error criterion improves the capability of eliminating residual noise and minimizing speech distortion. Since the weighting factors are also adjusted by measuring the usefulness of the AC and BC microphones, the proposed approach is suitable for tracking the parameter even if the characteristic of environment changes rapidly. The simulation results confirm the superiority of the proposed algorithm to conventional algorithms in high noise environments.
전체 326
256 International Conference Haemin Yang, Kyungguen Byun, Youngsu Kwak, Hong-Goo Kang "Parametric-based non-intrusive speech quality assessment by deep neural network" in 21th International Conference on Digital Signal Processing (DSP), 2016
255 International Conference Jin-Seob Kim, Young-Sun Joo, Inseon Jang, ChungHyun Ahn, Jeongil Seo, Hong-Goo Kang "A pitch-synchronous speech analysis and synthesis method for DNN-SPSS system" in 21th International Conference on Digital Signal Processing (DSP), 2016
254 International Conference Eunwoo Song, Frank K. Soong, Hong-Goo Kang "Improved Time-Frequency Trajectory Excitation Vocoder for DNN-Based Speech Synthesis" in INTERSPEECH, 2016
253 Domestic Conference Hyeonjoo Kang, Young-sun Joo, Wonsuk Jun, Hong-goo Kang "다층신경망 기반 다중 화자 음성변환 시스템" in 한국음향학회 제 33회 음성통신 및 신호처리 학술대회, 2016
252 International Conference Eunwoo Song, Hong-Goo Kang "Multi-class learning algorithm for deep neural network-based statistical parametric speech synthesis" in EUSIPCO, 2016
251 Domestic Conference Min-jae Hwang, JeeSok Lee, Misuk Lee, and Hong-Goo Kang "사전 분석법을 통한 스프레드 스펙트럼 기반 오디오 워터마킹 알고리즘의 성능 향상" in 한국음향학회 제 33회 음성통신 및 신호처리 학술대회, 2016
250 Domestic Conference 문현기, 박영철, 윤대희 "위상 일치와 가변 지수 감쇄 가중치 방법이 적용된 가상 저음 시스템" in 2016 년 한국방송·미디어공학회 하계학술대회, 2016
249 Domestic Conference 문현기, 박영철, 윤대희 "주파수 종속 믹싱 타임 추정 기법 분석" in 한국음향학회 춘계학술대회, 2016
248 Domestic Conference 서지호, 박영철, 윤대희 "헤드폰/이어폰 환경에서의 제약 최적화 기반의 효율적인 피드백 능동 소음 제어 필터 설계 알고리즘" in 한국음향학회 춘계학술대회, 2016
247 Domestic Conference 박규태, 박영철, 윤대희 "가변 가중치 곡선을 적용한 가상저음시스템" in 한국음향학회 춘계학술대회, 2016