Skip to main navigation Skip to search Skip to main content

Perception-guided and phonetic clustering weight tuning based on diphone pairs for unit selection TTS

Research output: Conference paperContributionpeer-review

5 Citations (Scopus)

Abstract

The quality of corpus based text-to-speech systems depends on the accuracy of the unit selection process, which relies on the values of the weights of the cost function. This paper is focused on defining a new framework for the tuning of these weights. We propose a technique for taking into account the subjective perception of speech in the selection process by means of Interactive Genetic Algorithms. Moreover, we introduce a CART-based method for unit clustering. Both techniques are applied to weight tuning based on diphone pairs. The conducted experiments analyze the feasibility of both proposals separately.

Original languageEnglish
Pages1221-1224
Number of pages4
Publication statusPublished - 2004
Event8th International Conference on Spoken Language Processing, ICSLP 2004 - Jeju, Jeju Island, Korea, Republic of
Duration: 4 Oct 20048 Oct 2004

Conference

Conference8th International Conference on Spoken Language Processing, ICSLP 2004
Country/TerritoryKorea, Republic of
CityJeju, Jeju Island
Period4/10/048/10/04

Fingerprint

Dive into the research topics of 'Perception-guided and phonetic clustering weight tuning based on diphone pairs for unit selection TTS'. Together they form a unique fingerprint.

Cite this