The K-nearest neighbors algorithm takes the majority label among the training features closest to the query, with a stated tie rule. Its data-dependent risk is the conditional test error given the training sample, and denotes its expectation over that sample.
For one nearest neighbour, condition on a feature value and couple the coincident nearest feature as . The two labels are conditionally independent Bernoulli, so their mismatch probability isIntegration over gives
Articles by others on the same topic
There are currently no matching articles.