2002 · 11 citations · 9 references
Bayesian network models are widely used for supervised prediction tasks such as classification. Usually the parameters of such models are determined using `unsupervised' methods such as likelihood maximization, as it has not been clear how to find the parameters maximizing the supervised likelihood or posterior globally. In this paper we show how this supervised learning problem can be solved efficiently for a large class of Bayesian network models, including the Naive Bayes (NB) and Tree-augmented NB (TAN) classifiers. We show that there exists an alternative parameterization of these models in which the supervised likelihood becomes concave. From this result it follows that there can be at most one maximum, easily found by local optimization methods.
9
Neural networks for pattern recognition
Choice Reviews Online · 1994 · 18.7K citations
Nir Friedman, Dan Geiger, Moisés Goldszmidt · Machine Learning · 1997 · 4.7K citations · Full text
Properties of Diagnostic Data Distributions
A. P. Dawid · Biometrics · 1976 · 179 citations