Scinovex
articleTop 1% cited

Nearest neighbor pattern classification

IEEE Transactions on Information Theory · 1967 · Vol. 13(1) · pp. 21–27
Thomas M. CoverPeter E. Hart

Abstract

The nearest neighbor decision rule assigns to an unclassified sample point the classification of the nearest of a set of previously classified points. This rule is independent of the underlying joint distribution on the sample points and their classifications, and hence the probability of error <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">R</tex> of such a rule must be at least as great as the Bayes probability of error <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">R^{\ast}</tex> --the minimum probability of error over all decision rules taking underlying probability structure into account. However, in a large sample analysis, we will show in the <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">M</tex> -category case that <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">R^{\ast} \leq R \leq R^{\ast}(2 --MR^{\ast}/(M-1))</tex> , where these bounds are the tightest possible, for all suitably smooth underlying distributions. Thus for any number of categories, the probability of error of the nearest neighbor rule is bounded above by twice the Bayes probability of error. In this sense, it may be said that half the classification information in an infinite sample set is contained in the nearest neighbor.

Statistical Methods and InferenceAdvanced Statistical Methods and ModelsRough Sets and Fuzzy LogicBayes' theoremk-nearest neighbors algorithmProbability distributionArtificial intelligenceMathematicsComputer scienceCombinatoricsPattern recognition (psychology)AlgorithmStatistics
Citations
15,642
FWCI
24.66
field-weighted impact
References
8
Percentile
100%
vs. same field & year
Citations per year
Cited by
Comparison and Evaluation of Methods for Liver Segmentation From CT Datasets
IEEE Transactions on Medical Imaging · 2009 · 1,115 citations
Data mining: concepts and techniques
Choice Reviews Online · 2012 · 28,852 citations
Graph-Theoretical Methods for Detecting and Describing Gestalt Clusters
IEEE Transactions on Computers · 1971 · 1,744 citations
A Direct Method of Nonparametric Measurement Selection
IEEE Transactions on Computers · 1971 · 838 citations
Biological shape and visual science (part I)
Journal of Theoretical Biology · 1973 · 1,162 citations
Window Size Impact in Human Activity Recognition
Sensors · 2014 · 619 citations
Consistent Nonparametric Regression
The Annals of Statistics · 1977 · 1,768 citations
Instance-based learning algorithms
Machine Learning · 1991 · 2,904 citations
References
Applications of Information Theory to Psychology
The Journal of Nervous and Mental Disease · 1959 · 769 citations
Related articles
Nearest neighbor pattern classification
IEEE Transactions on Information Theory · 1967 · 15,642 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.