Publication | Closed Access
Towards Personalized Speech Synthesis for Augmentative and Alternative Communication
42
Citations
79
References
2014
Year
Technology AdoptionCommunicationImproved PersonalizationSpeech RecognitionSurrogate TalkerConversation AnalysisAlternative CommunicationHealth SciencesSpeech PerceptionAugmentative And Alternative CommunicationSpeech SynthesisLinguisticsSpeech OutputText-to-speechSpeech CommunicationVoiceSpeech ProcessingArtsVoice TechnologySpeech InterfaceVoice Interaction
Text-to-speech options on augmentative and alternative communication (AAC) devices are limited. Often, several individuals in a group setting use the same synthetic voice. This lack of customization may limit technology adoption and social integration. This paper describes our efforts to generate personalized synthesis for users with profoundly limited speech motor control. Existing voice banking and voice conversion techniques rely on recordings of clearly articulated speech from the target talker, which cannot be obtained from this population. Our VocaliD approach extracts prosodic properties from the target talker's source function and applies these features to a surrogate talker's database, generating a synthetic voice with the vocal identity of the target talker and the clarity of the surrogate talker. Promising intelligibility results suggest areas of further development for improved personalization.
| Year | Citations | |
|---|---|---|
Page 1
Page 1