DI-UMONS : Dépôt institutionnel de l’université de Mons

Recherche transversale
(titres de publication, de périodique et noms de colloque inclus)
2013-08-31 - Colloque/Article dans les actes avec comité de lecture - Anglais - 6 page(s)

Astrinaki Maria , Moinet Alexis , Junichi Yamagishi, Richmond Korin, Ling Zen-Hua, King Simon, Dutoit Thierry , "MAGE - Reactive articulatory feature control of HMM-based parametric speech synthesis" in 8th Speech Synthesis Workshop (SSW8), 207-212, Barcelona, Spain, 2013

  • Codes CREF : Sciences de l'ingénieur (DI2000), Technologies de l'information et de la communication (TIC) (DI4730)
  • Unités de recherche UMONS : Théorie des circuits et traitement du signal (F105)
  • Instituts UMONS : Institut NUMEDIART pour les Technologies des Arts Numériques (Numédiart)
Texte intégral :

Abstract(s) :

(Anglais) In this paper, we present the integration of articulatory control into MAGE, a framework for realtime and interactive (reactive) parametric speech synthesis using hidden Markov models (HMMs). MAGE is based on the speech synthesis engine from HTS and uses acoustic features (spectrum and f0 ) to model and synthesize speech. In this work, we replace the standard acoustic models with models combining acoustic and articulatory features, such as tongue, lips and jaw positions. We then use feature-space-switched articulatory-to-acoustic regression matrices to enable us to control the spectral acoustic features by manipulating the articulatory features. Combining this synthesis model with M AGE allows us to interactively and intuitively modify phones synthesized in real time, for example transforming one phone into another, by controlling the configuration of the articulators in a visual display.