2013-08-31 - Colloque/Article dans les actes avec comité de lecture - Anglais

Astrinaki Maria , Moinet Alexis , Junichi Yamagishi, Richmond Korin, Ling Zen-Hua, King Simon, Dutoit Thierry , "MAGE - Reactive articulatory feature control of HMM-based parametric speech synthesis" in 8th Speech Synthesis Workshop (SSW8), 207-212, Barcelona, Spain, 2013

(Anglais) In this paper, we present the integration of articulatory control into MAGE, a framework for realtime and interactive (reactive) parametric speech synthesis using hidden Markov models (HMMs). MAGE is based on the speech synthesis engine from HTS and uses acoustic features (spectrum and f0 ) to model and synthesize speech. In this work, we replace the standard acoustic models with models combining acoustic and articulatory features, such as tongue, lips and jaw positions. We then use feature-space-switched articulatory-to-acoustic regression matrices to enable us to control the spectral acoustic features by manipulating the articulatory features. Combining this synthesis model with M AGE allows us to interactively and intuitively modify phones synthesized in real time, for example transforming one phone into another, by controlling the configuration of the articulators in a visual display.