Papers

Deep bi-directional long short-term memory based speech enhancement for wind noise reduction

International Conference
2016~2020
작성자
한혜원
작성일
2017-03-01 16:30
조회
293
Authors : Jinkyu Lee, Keulbit Kim, Turaj Shabestary, Hong-Goo Kang

Year : 2017

Publisher / Conference : HSCMA

In this paper, we propose a new recurrent neural network (RNN)-based single-channel speech enhancement framework for off-line wind noise reduction. To adequately represent highly non-stationary characteristics of wind noise, we first adopt a deep bi-directional long short-term memory (DBLSTM) structure. However, its enhanced output becomes muffled due to the spectral over-smoothing effect. To overcome this problem, we propose a new structure of DBLSTM-based speech enhancement system that internally incorporates the speech and noise power estimation processes in the spectral filtering framework. Furthermore, we propose a structure with an additional internal constraint of minimizing log a priori SNR, which provides efficient learning mechanism. Experimental results show that the proposed method improves source-to-distortion ratio (SDR) by 6.9 dB and perceptual evaluation of speech quality (PESQ) by 0.24 in comparison to the conventional DBLSTM-based system.
전체 327
267 Domestic Conference 오상신, 정수환, 강홍구 "음성 인식 기반의 방송미디어 디바이스 제어 및 편집 시스템 구현" in 대한전자공학회 추계학술대회, 2017
266 International Journal Eunwoo Song, Frank K. Soong, Hong-Goo Kang "Effective Spectral and Excitation Modeling Techniques for LSTM-RNN-Based Speech Synthesis Systems" in IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol.25, issue 11, pp.2152-2161, 2017
265 International Conference Seung-chul Shin, Junhyung Moon, Saewon Kye, Kyoungwoo Lee, Yong Seung Lee, Hong-Goo Kang "Continuous bladder volume monitoring system for wearable applications" in EMBC, 2017
264 Domestic Conference 김정규, 박영철, 강홍구 "저사양 TV 사운드 설계환경을 위한 IIR 필터 기반 주파수 등화기" in 대한전자공학회 학술대회, 2017
263 International Conference Jinkyu Lee, Keulbit Kim, Turaj Shabestary, Hong-Goo Kang "Deep bi-directional long short-term memory based speech enhancement for wind noise reduction" in HSCMA, 2017
262 International Conference JeeSok Lee, Soo-Whan Chung, Min-Seok Choi, Hong-Goo Kang "A study on search grid points for data-driven 3-D beamsteering" in HSCMA, 2017
261 Domestic Journal Ji-ho Seo, Dae Hee Youn, Young-Cheol Park "A Method of Designing Low-power Feedback Active Noise Control Filter for Headphones/Earphones" in 한국통신학회논문지, vol.10, 제 1호, pp.57-65, 2017
260 Domestic Journal Hyeongi Moon, Young-cheol Park, Yong Ju Lee, Young-soo Whang "MPEG-H 3D Audio Decoder Structure and Complexity Analysis" in 한국통신학회논문지, vol.42, 제 2호, pp.432-443, 2017
259 International Conference Young-Sun Joo, Won-Suk Jun, Hong-Goo Kang "Efficient deep neural networks for speech synthesis using bottleneck features" in APSIPA, 2016
258 Domestic Journal 문현기, 박영철, 황영수 "위상 일치와 가변 지수 감쇠 가중치 부여 방법이 적용된 가상 저음 시스템" in 방송공학회논문지, vol.21, 제 6호, pp.889-898, 2016