Adaptive k-Nearest Neighbor Learning for Robust Modal Regression on Multimodal and Heavy-Tailed Data

Sutarman Sutarman, Netti Herawati, Adli Abdillah Nababan

Abstract


Modal regression has attracted increasing attention as an alternative to mean-based regression, particularly in settings characterized by heteroscedasticity, multimodal conditional distributions, and heavy-tailed noise. In such scenarios, estimators based on central tendency may yield predictions that fall in low-density regions of the response space. This paper proposes an adaptive k-nearest neighbor framework for modal regression that integrates entropy-guided neighborhood selection with nonparametric mode estimation, including MeanShift clustering and one-dimensional kernel density estimation. The proposed approach adjusts neighborhood size based on local uncertainty, allowing the regression model to adapt to variations in data density without relying on a globally fixed parameter. Extensive experiments on simulated datasets and real-world benchmarks demonstrate that adaptive modal regression methods generally reduce or stabilize prediction errors relative to fixed-k modal regression and classical kNN mean and median estimators, particularly under heteroscedastic and multimodal conditions, although the magnitude of improvement varies across scenarios. Statistical tests confirm significant differences in most experimental settings, with practical gains ranging from incremental to substantial depending on data complexity. In addition to accuracy, computational behavior is explicitly examined. The findings show a trade-off between computational cost and predictive robustness: entropy-guided adaptive modal regression requires additional runtime due to neighborhood adaptation and density estimation, but this overhead increases proportionally with sample size and remains manageable for medium-sized datasets. Based on these results, adaptive modal regression provides a useful and flexible alternative for regression tasks involving complex and heterogeneous data distributions where robustness is prioritized over minimal computation time.

Keywords


Adaptive K-Nearest Neighbors; Modal Regression; Entropy-Based Neighborhood Selection; Robust Nonparametric Regression; Multimodal Conditional Distributions; Heavy-Tailed Noise; Instance-Based Learning; Meanshift Clustering; Kernel Density Estimation; Loca

Full Text:

PDF

References


T. Hastie, R. Tibshirani, and J. Friedman, The Elements of Statistical Learning. in Springer Series in Statistics. New York, NY: Springer New York, 2009. doi: 10.1007/978-0-387-84858-7.

T. Cover and P. Hart, “Nearest neighbor pattern classification,” IEEE Trans. Inf. Theory, vol. 13, no. 1, pp. 21–27, Jan. 1967, doi: 10.1109/TIT.1967.1053964.

T. S. Breusch and A. R. Pagan, “A Simple Test for Heteroscedasticity and Random Coefficient Variation,” Econometrica, vol. 47, no. 5, p. 1287, Sept. 1979, doi: 10.2307/1911963.

Y.-C. Chen, C. R. Genovese, R. J. Tibshirani, and L. Wasserman, “Nonparametric modal regression,” Ann. Stat., vol. 44, no. 2, Art. no. 2, Apr. 2016, doi: 10.1214/15-AOS1373.

K. L. Lange, R. J. A. Little, and J. M. G. Taylor, “Robust Statistical Modeling Using the t Distribution,” J. Am. Stat. Assoc., vol. 84, no. 408, pp. 881–896, Dec. 1989, doi: 10.1080/01621459.1989.10478852.

R. Koenker and G. Bassett, “Regression Quantiles,” Econometrica, vol. 46, no. 1, p. 33, Jan. 1978, doi: 10.2307/1913643.

M. Lee, “Mode regression,” J. Econom., vol. 42, no. 3, pp. 337–349, Nov. 1989, doi: 10.1016/0304-4076(89)90057-2.

W. Yao and L. Li, “A New Regression Model: Modal Linear Regression,” Scand. J. Stat., vol. 41, no. 3, pp. 656–671, Sept. 2014, doi: 10.1111/sjos.12054.

Y. Chen, “Modal regression using kernel density estimation: A review,” WIREs Comput. Stat., vol. 10, no. 4, Art. no. 4, July 2018, doi: 10.1002/wics.1431.

D. W. Scott, Multivariate Density Estimation: Theory, Practice, and Visualization, 1st ed. in Wiley Series in Probability and Statistics. Wiley, 2015. doi: 10.1002/9781118575574.

D. Comaniciu and P. Meer, “Mean shift: a robust approach toward feature space analysis,” IEEE Trans. Pattern Anal. Mach. Intell., vol. 24, no. 5, pp. 603–619, May 2002, doi: 10.1109/34.1000236.

S. Bermejo and J. Cabestany, “Adaptive soft k-nearest-neighbour classifiers,” Pattern Recognit., vol. 33, no. 12, pp. 1999–2005, Dec. 2000, doi: 10.1016/S0031-3203(99)00186-7.

R. J. Samworth, “Optimal weighted nearest neighbour classifiers,” Ann. Stat., vol. 40, no. 5, Oct. 2012, doi: 10.1214/12-AOS1049.

A. Kraskov, H. Stögbauer, and P. Grassberger, “Estimating mutual information,” Phys. Rev. E, vol. 69, no. 6, p. 066138, June 2004, doi: 10.1103/PhysRevE.69.066138.

B. W. Silverman, Density Estimation for Statistics and Data Analysis, 1st ed. Routledge, 2018. doi: 10.1201/9781315140919.

R. P. Brent, Algorithms for minimization without derivatives. in Prentice-Hall series in automatic computation. Englewood Cliffs, N.J: Prentice-Hall, 1973.

P. Srisuradetchai and K. Suksrikran, “Random kernel k-nearest neighbors regression,” Front. Big Data, vol. 7, p. 1402384, July 2024, doi: 10.3389/fdata.2024.1402384.




DOI: https://doi.org/10.47738/jads.v7i2.1221

Refbacks

  • There are currently no refbacks.



Barcode

Journal of Applied Data Sciences

ISSN:2723-6471 (Online)
Publisher:Bright Publisher
Website:http://bright-journal.org/JADS
Email:taqwa@amikompurwokerto.ac.id (principal contact)
  support@bright-journal.org (technical issues)

This work is licensed under a Creative Commons Attribution-ShareAlike 4.0