Towards expressive speech synthesis in english on a robotic platform
Sigrid Roehling, Bruce A. MacDonald, Catherine Watson
- Year
- 2006
- Citations
- 12
Abstract
Affect influences speech, not only in the words we choose, but in the way we say them. This pa-per reviews the research on vocal correlates in the expression of affect and examines the ability of currently available major text-to-speech (TTS) systems to synthesize expressive speech for an emotional robot guide. Speech features discussed include pitch, duration, loudness, spectral structure, and voice quality. TTS systems are examined as to their ability to control the fea-tures needed for synthesizing expressive speech: pitch, duration, loudness, and voice quality. The OpenMARY system is recommended since it provides the highest amount of control over speech production as well as the ability to work with a sophisticated intonation model. Open-MARY is being actively developed, is supported on our current Linux platform, and provides timing information for talking heads such as our current robot face. 1.
Keywords
Related papers
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Fractional Differential Equations
Igor Podlubný
2025
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991