
arXiv: 2204.04648
handle: 10498/28820 , 10486/709742
Missing values are common in many real-life datasets. However, most of the current machine learning methods can not handle missing values. This means that they should be imputed beforehand. Gaussian Processes (GPs) are non-parametric models with accurate uncertainty estimates that combined with sparse approximations and stochastic variational inference scale to large data sets. Sparse GPs can be used to compute a predictive distribution for missing data. Here, we present a hierarchical composition of sparse GPs that is used to predict missing values at each dimension using all the variables from the other dimensions. We call the approach missing GP (MGP). MGP can be trained simultaneously to impute all observed missing values. Specifically, it outputs a predictive distribution for each missing value that is then used in the imputation of other missing values. We evaluate MGP in one private clinical data set and four UCI datasets with a different percentage of missing values. We compare the performance of MGP with other state-of-the-art methods for imputing missing values, including variants based on sparse GPs and deep GPs. The results obtained show a significantly better performance of MGP.
Informática, FOS: Computer and information sciences, Computer Science - Machine Learning, Deep learning, Machine Learning (stat.ML), Statistics - Applications, Machine Learning (cs.LG), Methodology (stat.ME), Statistics - Machine Learning, Deep Gaussian processes, Missing values, Applications (stat.AP), Gaussian process, Variational inference, Statistics - Methodology
Informática, FOS: Computer and information sciences, Computer Science - Machine Learning, Deep learning, Machine Learning (stat.ML), Statistics - Applications, Machine Learning (cs.LG), Methodology (stat.ME), Statistics - Machine Learning, Deep Gaussian processes, Missing values, Applications (stat.AP), Gaussian process, Variational inference, Statistics - Methodology
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 30 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Top 10% | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
