Bookbot

Nearest neighbor methods for the imputation of missing values in low and high-dimensional data

Parametry

  • 218 stron
  • 8 godzin czytania

Więcej o książce

Nowadays, due to the advancement and significantly rapid growth in the technology, the collection of high-dimensional data is no longer a tedious task. Regardless of considerable advances in technology over the last few decades, the analysis of high-dimensional data faces new challenges concerning interpretation and integration. One of the major problems in high-dimensional data is the occurrence of missing values. The problem is in particular hard to handle when the distributional forms of the variables are different or the variables are measured on different measurement scales (e. g. binary, multi-categorical, continuous, etc.). Whatever the reason, missing data may occur in all areas of applied research. The inadequate handling of missing values may lead to biased results and incorrect inference. The standard statistical techniques for analyzing the data require complete cases without any missing observations. The deletion of the cases with missing information to obtain complete data will not only cause the loss of important information but can also affect inferences. In this dissertation, different imputation techniques using nearest neighbors are developed to address the missing data issues in high-dimensional as well as low dimensional data structures.

Zakup książki

Nearest neighbor methods for the imputation of missing values in low and high-dimensional data, Shahla Faisal

Język
Rok wydania
2018
Jak tylko się pojawi, wyślemy Ci wiadomość e-mail.

Metody płatności

Nikt jeszcze nie ocenił.Oceń

Tytuł
Nearest neighbor methods for the imputation of missing values in low and high-dimensional data
Język
angielski
Wydawca
Cuvillier
Rok wydania
2018
Liczba stron
218
ISBN10
3736997418
ISBN13
9783736997417
Seria
Opis
Nowadays, due to the advancement and significantly rapid growth in the technology, the collection of high-dimensional data is no longer a tedious task. Regardless of considerable advances in technology over the last few decades, the analysis of high-dimensional data faces new challenges concerning interpretation and integration. One of the major problems in high-dimensional data is the occurrence of missing values. The problem is in particular hard to handle when the distributional forms of the variables are different or the variables are measured on different measurement scales (e. g. binary, multi-categorical, continuous, etc.). Whatever the reason, missing data may occur in all areas of applied research. The inadequate handling of missing values may lead to biased results and incorrect inference. The standard statistical techniques for analyzing the data require complete cases without any missing observations. The deletion of the cases with missing information to obtain complete data will not only cause the loss of important information but can also affect inferences. In this dissertation, different imputation techniques using nearest neighbors are developed to address the missing data issues in high-dimensional as well as low dimensional data structures.