Papers

Applying A Speaker-dependent Speech Compression Technique to Concatenative TTS Synthesizers

International Journal
2006~2010
작성자
이진영
작성일
2007-02-01 13:32
조회
1027
Authors : Chang-Heon Lee, Sung-Kyo Jung, Hong-Goo Kang

Year : 2007

Publisher / Conference : IEEE Transactions on Audio, Speech, and Language Processing

Volume : 15, 2

Page : 632-640

This paper proposes a new speaker-dependent coding algorithm to efficiently compress a large speech database for corpus-based concatenative text-to-speech (TTS) engines while maintaining high fidelity. To achieve a high compression ratio and meet the fundamental requirements of concatenative TTS synthesizers, such as partial segment decoding and random access capability, we adopt a nonpredictive analysis-by-synthesis scheme for speaker-dependent parameter estimation and quantization. The spectral coefficients are quantized by using a memoryless split vector quantization (VQ) approach that does not use frame correlation. Considering that excitation signals of a specific speaker show low intra-variation especially in the voiced regions, the conventional adaptive codebook for pitch prediction is replaced by a speaker-dependent pitch-pulse codebook trained by a corpus of single-speaker speech signals. To further improve the coding efficiency, the proposed coder flexibly combines nonpredictive and predictive type method considering the structure of the TTS system. By applying the proposed algorithm to a Korean TTS system, we could obtain comparable quality to the G.729 speech coder and satisfy all the requirements that TTS system needs. The results are verified by both objective and subjective quality measurements. In addition, the decoding complexity of the proposed coder is around 55% lower than that of G.729 annex A
전체 355
11 International Journal Kyung-Tae Kim, Min-Ki Lee, Hong-Goo Kang "Speech Bandwidth Extension using Temporal Envelope Modeling" in IEEE Signal Processing Letters, vol.15, pp.429-432, 2008
10 International Journal Junho Lee, Eunjung Song, Young-Cheol Park, Dae Hee Youn "Effective Bass Enhancement Using Second-Order Adaptive Notch Filter" in IEEE Transactions on Consumer Electronics, vol.54, issue 2, pp.663-668, 2008
9 International Journal Hee-Young Park, Ki-Man Kim, Hyun-Woo Kang, Dae Hee Youn, Chungyong Lee "A Simplified Subspace Fitting Method for Estimating Shape of a Towed Array" in IEEE Journal of Oceanic Engineering, vol.33, issue 2, pp.215-223, 2008
8 International Journal Tacksung Choi, Young-Cheol Park, Dae Hee Youn "Design of Time-Varying Reverberators for Low Memory Applications" in IEICE Transactions on Information and Systems, vol.E91-D, No.2, pp.379-382, 2008
7 International Journal Junho Lee, Young-Cheol Park, Dae Hee Youn "Robust pseudo affine projection algorithm with variable step-size" in Electronics Letters, vol.44, issue 3, pp.250-252, 2008
6 International Journal Chulhan Lee, Jeung-Yoon Choi, Kar-Ann Toh, Sangyoun Lee "Alignment-Free Cancelable Fingerprint Templates Based on Local Minutiae Information" in IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics), vol.37, issue 4, pp.980-992, 2007
5 International Journal Kyung Tae Kim, Jeung-Yoon Choi, Hong-Goo Kang "Perceptual relevance of the temporal envelope to the speech signal in the 4–7kHz band" in The Journal of the Acoustical Society of America, vol.122, issue 3, 2007
4 International Journal Chang-Heon Lee, Sung-Kyo Jung, Hong-Goo Kang "Applying A Speaker-dependent Speech Compression Technique to Concatenative TTS Synthesizers" in IEEE Transactions on Audio, Speech, and Language Processing, vol.15, 2, pp.632-640, 2007
3 International Journal Seungil Kim, Chungyong Lee, Hong-Goo Kang "Optimum beamformer in correlated source environments" in Journal of Acoustical Society of America, vol.120, issue 6, 2006
2 International Journal Kyoung Ho Bang, Young-Cheol Park, Jeongil Seo "Audio Transcoding for Audio Streams from a T-DTV Broadcasting Station to a T-DMB Receiver" in ETRI Journal, vol.28, issue5, pp.664-667, 2006