REVIEW 1 cited by
Efficient EM Training of Gaussian Mixtures with Missing Data
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In data-mining applications, we are frequently faced with a large fraction of missing entries in the data matrix, which is problematic for most discriminant machine learning algorithms. A solution that we explore in this paper is the use of a generative model (a mixture of Gaussians) to compute the conditional expectation of the missing variables given the observed variables. Since training a Gaussian mixture with many different patterns of missing values can be computationally very expensive, we introduce a spanning-tree based algorithm that significantly speeds up training in these conditions. We also observe that good results can be obtained by using the generative model to fill-in the missing values for a separate discriminant learning algorithm.
Forward citations
Cited by 1 Pith paper
-
Mixture-based Multiple Imputation Model for Clinical Data with a Temporal Dimension
MixMI, a mixture of Gaussian-process and linear-regression imputers with individualized mixing weights, reports lower mean absolute scaled error than six benchmarks on all four datasets tested.
Discussion (0). Continue with ORCID to comment.