Papers

Multi-class learning algorithm for deep neural network-based statistical parametric speech synthesis

International Conference
2016~2020
작성자
한혜원
작성일
2016-08-01 16:23
조회
1389
Authors : Eunwoo Song, Hong-Goo Kang

Year : 2016

Publisher / Conference : EUSIPCO

This paper proposes a multi-class learning (MCL) algorithm for a deep neural network (DNN)-based statistical parametric speech synthesis (SPSS) system. Although the DNN-based SPSS system improves the modeling accuracy of statistical parameters, its synthesized speech is often muffled because the training process only considers the global characteristics of the entire set of training data, but does not explicitly consider any local variations. We introduce a DNN-based context clustering algorithm that implicitly divides the training data into several classes, and train them via a shared hidden layer-based MCL algorithm. Since the proposed MCL method efficiently models both the universal and class-dependent characteristics of various phonetic information, it not only avoids the model over-fitting problem but also reduces the over-smoothing effect. Objective and subjective test results also verify that the proposed algorithm performs much better than the conventional method.
전체 355
8 Domestic Conference 문현기, 박영철, 윤대희 "위상 일치와 가변 지수 감쇄 가중치 방법이 적용된 가상 저음 시스템" in 2016 년 한국방송·미디어공학회 하계학술대회, 2016
7 Domestic Conference 문현기, 박영철, 윤대희 "주파수 종속 믹싱 타임 추정 기법 분석" in 한국음향학회 춘계학술대회, 2016
6 Domestic Conference 서지호, 박영철, 윤대희 "헤드폰/이어폰 환경에서의 제약 최적화 기반의 효율적인 피드백 능동 소음 제어 필터 설계 알고리즘" in 한국음향학회 춘계학술대회, 2016
5 Domestic Conference 박규태, 박영철, 윤대희 "가변 가중치 곡선을 적용한 가상저음시스템" in 한국음향학회 춘계학술대회, 2016
4 Domestic Conference 양해민, 변경근, 강홍구 "RTCP를 이용한 심층 신경망 기반 음질평가 점수 대역 분별 알고리즘" in 한국음향학회 춘계학술대회, 2016
3 Domestic Conference 김글빛, 이진규, 강홍구 "문장종속 화자검증 시스템을 위한 비음수 행렬 분해 기반 잡음 제거" in 한국음향학회 춘계학술대회, 2016
2 Domestic Conference 김진섭, 주영선, 강홍구(연세대학교), 장인선, 안충현(한국전자통신연구원) "음향 모델 성능 개선을 위한 피치 동기화 기반의 DNN-TTS 시스템" in 한국음향학회 춘계학술대회, 2016
1 International Conference Hyeongi Moon, Gyutae Park, Yeong-cheol Park, Dae Hee Youn "A Phase-Matched Exponential Harmonic Weighting for Improved Sensation of Virtual Bass" in 140th Convention of Audio Engineering Society, pp.9544, 2016