Minimax Rate Optimal Adaptive Nearest Neighbor Classification and Regression

10/22/2019
by   Puning Zhao, et al.
0

k Nearest Neighbor (kNN) method is a simple and popular statistical method for classification and regression. For both classification and regression problems, existing works have shown that, if the distribution of the feature vector has bounded support and the probability density function is bounded away from zero in its support, the convergence rate of the standard kNN method, in which k is the same for all test samples, is minimax optimal. On the contrary, if the distribution has unbounded support, we show that there is a gap between the convergence rate achieved by the standard kNN method and the minimax bound. To close this gap, we propose an adaptive kNN method, in which different k is selected for different samples. Our selection rule does not require precise knowledge of the underlying distribution of features. The new proposed method significantly outperforms the standard one. We characterize the convergence rate of the proposed adaptive method, and show that it matches the minimax lower bound.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/30/2020

Analysis of KNN Density Estimation

We analyze the ℓ_1 and ℓ_∞ convergence rates of k nearest neighbor densi...
research
08/03/2023

Minimax Optimal Q Learning with Nearest Neighbors

Q learning is a popular model free reinforcement learning method. Most o...
research
12/13/2022

Minimax Optimal Estimation of Stability Under Distribution Shift

The performance of decision policies and prediction models often deterio...
research
05/26/2023

Robust Nonparametric Regression under Poisoning Attack

This paper studies robust nonparametric regression, in which an adversar...
research
11/23/2017

The Nearest Neighbor Information Estimator is Adaptively Near Minimax Rate-Optimal

We analyze the Kozachenko--Leonenko (KL) nearest neighbor estimator for ...
research
05/09/2022

Mathematical Properties of Continuous Ranked Probability Score Forecasting

The theoretical advances on the properties of scoring rules over the pas...
research
02/26/2022

Enhanced Nearest Neighbor Classification for Crowdsourcing

In machine learning, crowdsourcing is an economical way to label a large...

Please sign up or login with your details

Forgot password? Click here to reset