Infoscience (Ecole Polytechnique Fédérale de Lausanne) · 2017 · 19 citations · 3 references
Open access
MusicFrenchSpeech CorpusSpoken FrenchPhonologyCorpus LinguisticsTts SystemsSpeech RecognitionLanguage DocumentationPhoneticsTen HoursSpeech InterfaceVoice RecognitionLanguage StudiesMachine TranslationHealth SciencesSpeech SynthesisSpeech OutputFrench Voice TalentText-to-speechSpeech CommunicationSpeech TechnologyVoiceSpeech ProcessingSpeech PerceptionLinguistics
We describe the design and recording of a high quality French speech corpus, aimed at building TTS systems, investigate multiple styles, and emphasis. The data was recorded by a French voice talent, and contains about ten hours of speech, including emphasised words in many different contexts. The database contains more than ten hours of speech and is freely available.
3
Europarl: A Parallel Corpus for Statistical Machine Translation
Philipp Koehn · 2005 · 3.1K citations