Papers

Model Order Selection for Wind Noise Reduction in Non-negative Matrix Factorization

International Conference

2016~2020

작성자

한혜원

작성일

2019-06-01 16:47

조회

3918

Authors : Keulbit Kim, Jinkyu Lee, Jan Skoglund, Hong-Goo Kang

Year : 2019

Publisher / Conference : ITC-CSCC

In this paper, we propose a wind noise reduction method based on various types of non-negative matrix factorization (NMF) approaches. Since wind noise has highly non- stationary spectral characteristics that are difficult to remove using stochastic oriented methods, it is more effective to use template matching-based methods.

We first investigate whether the quality of enhanced output varies depending on the ratio of the size of speech models to that of noise models. We especially show that the optimal ratio is related to the signal-to-noise ratio (SNR) of the input signal. Based on the results of analysis, we propose an efficient algorithm for adaptively changing the size of speech and noise NMF models in each analysis frame. Since the proposed algorithm takes into account the trade- off relationship between speech distortion and noise reduction, its output quality becomes very natural. The experimental results also confirm the superiority of the proposed algorithm to conventional template matching based algorithms.

« Emotional Speech Synthesis Based on Style Embedded Tacotron2 Framework

Parameter enhancement for MELP speech codec in noisy communication environment »

목록보기

전체 372

104	International Conference	Min-Jae Hwang, Eunwoo Song, Ryuichi Yamamoto, Frank Soong, Hong-Goo Kang "Improving LPCNet-based Text-to-Speech with Linear Prediction-structured Mixture Density Network" in ICASSP, 2020
103	International Conference	Hyeonjoo Kang, Young-Sun Joo, Inseon Jang, Chunghyun Ahn, Hong-Goo Kang "A Study on Acoustic Parameter Selection Strategies to Improve Deep Learning-Based Speech Synthesis" in APSIPA, 2019
102	International Conference	Min-Jae Hwang, Hong-Goo Kang "Parameter enhancement for MELP speech codec in noisy communication environment" in INTERSPEECH, 2019
101	International Conference	Keulbit Kim, Jinkyu Lee, Jan Skoglund, Hong-Goo Kang "Model Order Selection for Wind Noise Reduction in Non-negative Matrix Factorization" in ITC-CSCC, 2019
100	International Conference	Ohsung Kwon, Inseon Jang, ChungHyun Ahn, Hong-Goo Kang "Emotional Speech Synthesis Based on Style Embedded Tacotron2 Framework" in ITC-CSCC, 2019
99	International Conference	Kyungguen Byun, Eunwoo Song, Jinseob Kim, Jae-Min Kim, Hong-Goo Kang "Excitation-by-SampleRNN Model for Text-to-Speech" in ITC-CSCC, 2019
98	International Conference	Yang Yuan, Soo-Whan Chung, Hong-Goo Kang "Gradient-based active learning query strategy for end-to-end speech recognition" in ICASSP, 2019
97	International Conference	Soo-Whan Chung, Joon Son Chung, Hong-Goo Kang "Perfect match: Improved cross-modal embeddings for audio-visual synchronisation" in ICASSP, 2019
96	International Conference	Hyewon Han, Kyungguen Byun, Hong-Goo Kang "A Deep Learning-based Stress Detection Algorithm with Speech Signal" in Workshop on Audio-Visual Scene Understanding for Immersive Multimedia (AVSU’18), 2018
95	International Conference	Min-Jae Hwang, Eunwoo Song, Jin-Seob Kim, Hong-Goo Kang "A Unified Framework for the Generation of Glottal Signals in Deep Learning-based Parametric Speech Synthesis Systems" in INTERSPEECH, 2018

Model Order Selection for Wind Noise Reduction in Non-negative Matrix Factorization

Previous

Sister Lab.

Yonsei University

Academic Website