article Open access

Imputation Techniques in Machine Learning – A Survey

  • International Journal on Recent and Innovation Trends in Computing and Communication
Research footprint

At a glance

Citations
2
References
10
Comments
0
Paper overview

Abstract

Machine learning plays a pivotal role in data analysis and information extraction. However, one common challenge encountered in this process is dealing with missing values. Missing data can find its way into datasets for a variety of reasons. It can result from errors during data collection and management, intentional omissions, or even human errors. It's important to note that most machine learning models are not designed to handle missing values directly. Consequently, it becomes essential to perform data imputation before feeding the data into a machine learning model. Multiple techniques are available for imputing missing values, and the choice of technique should be made judiciously, considering various parameters. An inappropriate choice can disrupt the overall distribution of data values and subsequently impact the model's performance. In this paper, various imputation methods, including Mean, Median, K-nearest neighbors (KNN)-based imputation, Linear Regression, Miss Forest, and MICE are examined.

Record transparency

Publication details

DOI
10.17762/ijritcc.v11i10.8662
OpenAlex
W4389341925
Document type
article
Language
EN
Source
International Journal on Recent and Innovation Trends in Computing and Communication
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.