The Journal of the Acoustical Society of America · 2013 · 53 citations · 34 references
Endangered LanguagesEngineeringSpeech CorpusMultilingualismSpoken Language ProcessingPhonologyCorpus LinguisticsSpeech RecognitionNatural Language ProcessingApplied LinguisticsLanguage DocumentationData ScienceLanguage AdaptationLanguage TestingComputational LinguisticsLanguage AcquisitionLanguage EngineeringPhoneticsLanguage StudiesEndangered LanguageMachine TranslationNatural LanguageYoloxóchitl MixtecAutomatic AlignmentSpeech AnalysisSpeech TechnologyForced AlignmentUntrained AlignmentLanguage RecognitionEndangered Language DataLanguage MaintenanceSpeech ProcessingLinguistics
While efforts to document endangered languages have steadily increased, the phonetic analysis of endangered language data remains a challenge. The transcription of large documentation corpora is, by itself, a tremendous feat. Yet, the process of segmentation remains a bottleneck for research with data of this kind. This paper examines whether a speech processing tool, forced alignment, can facilitate the segmentation task for small data sets, even when the target language differs from the training language. The authors also examined whether a phone set with contextualization outperforms a more general one. The accuracy of two forced aligners trained on English (hmalign and p2fa) was assessed using corpus data from Yoloxóchitl Mixtec. Overall, agreement performance was relatively good, with accuracy at 70.9% within 30 ms for hmalign and 65.7% within 30 ms for p2fa. Segmental and tonal categories influenced accuracy as well. For instance, additional stop allophones in hmalign's phone set aided alignment accuracy. Agreement differences between aligners also corresponded closely with the types of data on which the aligners were trained. Overall, using existing alignment systems was found to have potential for making phonetic analysis of small corpora more efficient, with more allophonic phone sets providing better agreement than general ones.
34
Praat: Doing Phonetics by Computer
Ear and Hearing · 2011 · 8K citations
Neurotology, Volume 32, Phonology +28
The world's languages in crisis
Michael E. Krauss · Language · 1992 · 1.4K citations
The Discourse Basis of Ergativity
John W. Du Bois · Language · 1987 · 1.2K citations
Philosophy Of Language, Discourse Structure, Pragmatic Analysis +6
Darpa Timit Acoustic-Phonetic Continuous Speech Corpus CD-ROM {TIMIT} | NIST
John S. Garofolo, Lori Lamel, William M. Fisher et al. · 1993 · 679 citations