Artificial IntelligenceHybrid EventEngineeringArctic SubsetVoice EvaluationSpeech RecognitionComputational LinguisticsBlizzard ChallengeVoice RecognitionMary Tts ParticipationAcoustic AnalysisGame DesignMixed-timed CircuitsHealth SciencesEducational EntertainmentSpeech SynthesisSpeech OutputComputer ScienceText-to-speechSpeech CommunicationSpeech TechnologyUnit DefinitionVoiceSpeech AcousticsGlobal ChallengeSpeech ProcessingSpeech PerceptionVoice TechnologyLinguistics
This paper describes the second participation of the open source MARY TTS unit selection system in a Blizzard challenge. Compared to last year’s system, a number of welldefined changes have been made to the algorithm, concerning unit definition, prosody models, and signal modification. The results in this year’s challenge are considerably improved, confirming that the changes were worthwhile. The paper also reports on an approach to the selection of a subset of the utterances provided, in order to build a voice with good coverage not larger than the pre-defined “Arctic” subset. Results show that this small voice is perceived slightly better than the voice we built from the Arctic subset.
6
Praat: Doing Phonetics by Computer
Ear and Hearing · 2011 · 8K citations
Neurotology, Volume 32, Phonology +28
The CMU Arctic speech databases.
John Kominek, Alan W. Black · 2004 · 557 citations
Voice quality interpolation for emotional text-to-speech synthesis
Oytun Türk, Marc L. Schröder, Barış Bozkurt et al. · 2005 · 37 citations
Music, Engineering, Speech Corpus +18