Concepedia

A study of cross-validation and bootstrap for accuracy estimation and model selection

Ron Kohavi

1995 · 10.7K citations · 11 references

Abstract

We review accuracy estimation methods and compare the two most common methods: crossvalidation and bootstrap. Recent experimental results on artificial data and theoretical results in restricted settings have shown that for selecting a good classifier from a set of classifiers (model selection), ten-fold cross-validation may be better than the more expensiveleaveone -out cross-validation. We report on a largescale experiment---over half a million runs of C4.5 and a Naive-Bayes algorithm---to estimate the effects of different parameters on these algorithms on real-world datasets. For crossvalidation, wevary the number of folds and whether the folds are stratified or not# for bootstrap, wevary the number of bootstrap samples. Our results indicate that for real-word datasets similar to ours, the best method to use for model selection is ten-fold stratified cross validation, even if computation power allows using more folds.

References

11

Classification and Regression Trees.

Alexander Gordon, Leo Breiman, Jerome H. Friedman et al. · Biometrics · 1984

+19

23.8K citations

Classification and Regression Trees.

John Van Ryzin, Leo Breiman, Jerome H. Friedman et al. · Journal of the American Statistical Association · 1986

+9

21K citations

Stacked generalization

David H. Wolpert · Neural Networks · 1992

7.1K citations

2.1K citations

An analysis of Bayesian classifiers

Pat Langley, and Wayne Iba, Kevin Thompson · 1992

1.1K citations