Papers

딥러닝 기반 종단 간 다채널 음질 개선 알고리즘

Domestic Conference

2016~2020

작성자

이지현

작성일

2020-08-01 21:52

조회

2318

Authors : 이성현, 강홍구

Year : 2020

Publisher / Conference : 전자공학회 하계학술대회

Page : 968-970

In this paper, we propose a deep learning-based multi-channel speech enhancement algorithm. The proposed system consists of three sub-modules such as magnitude estimation, phase estimation, and spatial filtering modules. To minimize the distortion between the target speech and enhanced signal waveform, we adopt an end-to-end modeling architecture that considers time-domain reconstruction loss, magnitude and phase spectrum loss, and spatial information between microphones. The experimental results show that the proposed model shows much better performance than conventional algorithm in terms of noise reduction, intelligibility and speech quality.

삼성 전기 논문상 수상작

« 메타러닝을 이용한 SAR 영상 자동표적 인식

화자 및 발화 스타일 임베딩을 통한 다화자 음성합성 시스템 음질 향상 »

목록보기

전체 356

346	International Conference	WooSeok Ko, Seyun Um, Zhenyu Piao, Hong-goo Kang "Consideration of Varying Training Lengths for Short-Duration Speaker Verification" in APSIPA ASC, 2023
345	International Journal	Hyungchan Yoon, Changhwan Kim, Seyun Um, Hyun-Wook Yoon, Hong-Goo Kang "SC-CNN: Effective Speaker Conditioning Method for Zero-Shot Multi-Speaker Text-to-Speech Systems" in IEEE Signal Processing Letters, vol.30, pp.593-597, 2023
344	International Conference	Miseul Kim, Zhenyu Piao, Jihyun Lee, Hong-Goo Kang "BrainTalker: Low-Resource Brain-to-Speech Synthesis with Transfer Learning using Wav2Vec 2.0" in The IEEE-EMBS International Conference on Biomedical and Health Informatics (BHI), 2023
343	International Conference	Seyun Um, Jihyun Kim, Jihyun Lee, Hong-Goo Kang "Facetron: A Multi-speaker Face-to-Speech Model based on Cross-Modal Latent Representations" in EUSIPCO, 2023
342	International Conference	Hejung Yang, Hong-Goo Kang "Feature Normalization for Fine-tuning Self-Supervised Models in Speech Enhancement" in INTERSPEECH, 2023
341	International Conference	Jihyun Kim, Hong-Goo Kang "Contrastive Learning based Deep Latent Masking for Music Source Seperation" in INTERSPEECH, 2023
340	International Conference	Woo-Jin Chung, Doyeon Kim, Soo-Whan Chung, Hong-Goo Kang "MF-PAM: Accurate Pitch Estimation through Periodicity Analysis and Multi-level Feature Fusion" in INTERSPEECH, 2023
339	International Conference	Hyungchan Yoon, Seyun Um, Changhwan Kim, Hong-Goo Kang "Adversarial Learning of Intermediate Acoustic Feature for End-to-End Lightweight Text-to-Speech" in INTERSPEECH, 2023
338	International Conference	Hyungchan Yoon, Changhwan Kim, Eunwoo Song, Hyun-Wook Yoon, Hong-Goo Kang "Pruning Self-Attention for Zero-Shot Multi-Speaker Text-to-Speech" in INTERSPEECH, 2023
337	International Conference	Doyeon Kim, Soo-Whan Chung, Hyewon Han, Youna Ji, Hong-Goo Kang "HD-DEMUCS: General Speech Restoration with Heterogeneous Decoders" in INTERSPEECH, 2023

딥러닝 기반 종단 간 다채널 음질 개선 알고리즘

Previous

Sister Lab.

Yonsei University

Academic Website