Papers

Improved time-frequency trajectory excitation modeling for a statistical parametric speech synthesis system

International Conference
2011~2015
작성자
한혜원
작성일
2015-04-01 00:45
조회
1591
Authors : Eunwoo Song, Young-Sun Joo, Hong-Goo Kang

Year : 2015

Publisher / Conference : ICASSP

This paper proposes an improved time-frequency trajectory excitation (TFTE) modeling method for a statistical parametric speech synthesis system. The proposed approach overcomes the dimensional variation problem of the training process caused by the inherent nature of the pitch-dependent analysis paradigm. By reducing the redundancies of the parameters using predicted average block coefficients (PABC), the proposed algorithm efficiently models excitation, even if its dimension is varied. Objective and subjective test results verify that the proposed algorithm provides not only robustness to the training process but also naturalness to the synthesized speech.
전체 355
4 Domestic Journal 박영철, 이태규, 윤대희 "MPEG-H 3D 오디오 바이노럴 렌더링 기술 표준화" in 대한전기학회, 전기의 세계, vol.64, 제 2호, pp.27-31, 2015
3 Domestic Journal 오현오, 이태규, 전세운, 윤대희, 박영철, 서정일, 이용주 "모바일 3D 사운드 : 바이노럴 오디오 기술 동향" in 방송공학회논문지, vol.19, 제 1호, pp.65-74, 2014
2 Domestic Conference 이태규, 이지석, 강홍구 "스펙트럼과 통계적 모델링을 통한 음질 개선" in 한국음향학회 추계학술발표대회, 2013
1 Domestic Conference 이태규, 백용현, 박영철, 윤대희 "스테레오-멀티채널 업믹스 시스템에서의 초기 반사음 생성 기법" in 한국방송공학회, 2013