Concepedia

Class-based n -gram models of natural language

Peter F. Brown, P.V. deSouza, Robert L. Mercer, Vincent J. Della Pietra, Jenifer C. Lai

Computational Linguistics · 1992 · 2.9K citations · 11 references

Concepts

Abstract

We address the problem of predicting a word from previous words in a sample of text. In particular, we discuss n-gram models based on classes of words. We also discuss several statistical algorithms for assigning words to classes based on the frequency of their co-occurrence with other words. We find that we are able to extract classes that have the flavor of either syntactically based groupings or semantically based groupings, depending on the nature of the underlying statistics.

References

11

Maximum Likelihood from Incomplete Data Via the <i>EM</i> Algorithm

A. P. Dempster, N. M. Laird, Donald B. Rubin · Journal of the Royal Statistical Society Series B (Statistical Methodology) · 1977

+13

49.2K citations

Information Theory and Reliable Communication.

P. M. Lee, Robert T. Gallager · Journal of the Royal Statistical Society Series A (General) · 1970

+11

5.5K citations

A statistical approach to machine translation

Peter F. Brown, John Cocke, Stephen A. Della Pietra et al. · Computational Linguistics · 1990

+12

1.7K citations