Resum
Unit selection text-to-speech (TTS) conversion is an ongoing research for the speech synthesis community. This paper is focused on tuning the weights involved in the target concatenation cost metrics. We propose a method for automatically adjusting these weights simultaneously by means of diphone and triphone pairs. This method is based on techniques provided by the evolutionary computation community, taking advantage of their robustness in noisy domains. The experiments and their analyses demonstrate its good performance in this problem, thus, overcoming some constraints assumed by previous works leading to a new interesting framework for further investigations.
| Idioma original | Anglès |
|---|---|
| Pàgines | 1333-1336 |
| Nombre de pàgines | 4 |
| Estat de la publicació | Publicada - 2003 |
| Esdeveniment | 8th European Conference on Speech Communication and Technology, EUROSPEECH 2003 - Geneva, Switzerland Durada: 1 de set. 2003 → 4 de set. 2003 |
Conferència
| Conferència | 8th European Conference on Speech Communication and Technology, EUROSPEECH 2003 |
|---|---|
| País/Territori | Switzerland |
| Ciutat | Geneva |
| Període | 1/09/03 → 4/09/03 |
Fingerprint
Navegar pels temes de recerca de 'Evolutionary weight tuning based on diphone pairs for unit selection speech synthesis'. Junts formen un fingerprint únic.Com citar-ho
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver