Class-based n -gram models of natural language
Computational Linguistics · 1992 · 2.9K citations · 11 references
Syntactic ParsingEngineeringSemanticsCorpus LinguisticsText MiningWord EmbeddingsNatural Language ProcessingInformation RetrievalData ScienceComputational LinguisticsLanguage EngineeringGrammarN-gram ModelsMachine TranslationNatural LanguageNlp TaskKnowledge DiscoveryTerminology ExtractionDistributional SemanticsUnderlying StatisticsPrevious WordsKeyword ExtractionArtsLinguistics
We address the problem of predicting a word from previous words in a sample of text. In particular, we discuss n-gram models based on classes of words. We also discuss several statistical algorithms for assigning words to classes based on the frequency of their co-occurrence with other words. We find that we are able to extract classes that have the flavor of either syntactically based groupings or semantically based groupings, depending on the nature of the underlying statistics.
11
Maximum Likelihood from Incomplete Data Via the <i>EM</i> Algorithm
A. P. Dempster, N. M. Laird, Donald B. Rubin · Journal of the Royal Statistical Society Series B (Statistical Methodology) · 1977
Statistical Signal ProcessingMixture DistributionEngineering+13
49.2K citations
An introduction to probability theory and its applications
Journal of the Franklin Institute · 1958
Discrete ProbabilityProbability TheoryProbabilistic Analysis+1
29.7K citations
Information Theory and Reliable Communication.
P. M. Lee, Robert T. Gallager · Journal of the Royal Statistical Society Series A (General) · 1970
5.5K citations
THE POPULATION FREQUENCIES OF SPECIES AND THE ESTIMATION OF POPULATION PARAMETERS
I. J. Good · Biometrika · 1953
3.2K citations
A statistical approach to machine translation
Peter F. Brown, John Cocke, Stephen A. Della Pietra et al. · Computational Linguistics · 1990
Natural Language ProcessingTranslation StudiesComputer-assisted Translation+12
1.7K citations