An Optimized k-NN Approach for Classification on Imbalanced Datasets with Missing Data

  • Ezgi Can OzanEmail author
  • Ekaterina Riabchenko
  • Serkan Kiranyaz
  • Moncef Gabbouj
Conference paper
Part of the Lecture Notes in Computer Science book series (LNCS, volume 9897)


In this paper, we describe our solution for the machine learning prediction challenge in IDA 2016. For the given problem of 2-class classification on an imbalanced dataset with missing data, we first develop an imputation method based on k-NN to estimate the missing values. Then we define a tailored representation for the given problem as an optimization scheme, which consists of learned distance and voting weights for k-NN classification. The proposed solution performs better in terms of the given challenge metric compared to the traditional classification methods such as SVM, AdaBoost or Random Forests.


k-NN classifier Missing data Imbalanced datasets 


  1. 1.
    García-Laencina, P.J., Sancho-Gómez, J.-L., Figueiras-Vidal, A.R.: Pattern classification with missing data: a review. Neural Comput. Appl. 19(2), 263–282 (2009)CrossRefGoogle Scholar
  2. 2.
    Batista, G., Monard, M.C.: A study of k-nearest neighbour as an imputation method. Hybrid Intell. Syst. 87(48), 251–260 (2002)Google Scholar
  3. 3.
    Wu, X., Kumar, V., Ross, Q.J., Ghosh, J., Yang, Q., Motoda, H., McLachlan, G.J., Ng, A., Liu, B., Yu, P.S., Zhou, Z.-H., Steinbach, M., Hand, D.J., Steinberg, D.: Top 10 algorithms in data mining. Knowl. Inf. Syst. 14(1), 1–37 (2008)CrossRefGoogle Scholar
  4. 4.
    Dudani, S.A.: The distance-weighted k-nearest-neighbor rule. IEEE Trans. Syst. Man Cybern. SMC-6(4), 325–327 (1976)CrossRefGoogle Scholar
  5. 5.
    Pedregosa, F., Grisel, O., Weiss, R., Passos, A., Brucher, M.: Scikit-learn: machine learning in python. J. Mach. Learn. Res. 12(1), 2825–2830 (2011)MathSciNetzbMATHGoogle Scholar

Copyright information

© Springer International Publishing AG 2016

Authors and Affiliations

  • Ezgi Can Ozan
    • 1
    Email author
  • Ekaterina Riabchenko
    • 1
  • Serkan Kiranyaz
    • 2
  • Moncef Gabbouj
    • 1
  1. 1.Tampere University of TechnologyTampereFinland
  2. 2.Electrical Engineering Department, College of EngineeringQatar UniversityDohaQatar

Personalised recommendations