Review: A gentle introduction to imputation of missing values

Donders, A.R.T.; van der Heijden, Geert; Stijnen, Theo; Moons, Karel

doi:10.1016/j.jclinepi.2006.01.014

A.R.T. Donders, G.J.M.G. van der Heijden (Geert), Th. Stijnen (Theo) and K.G.M. Moons (Karel)

2006-10-01

Review: A gentle introduction to imputation of missing values

Journal of Clinical Epidemiology , Volume 59 - Issue 10 p. 1087- 1091

In most situations, simple techniques for handling missing data (such as complete case analysis, overall mean imputation, and the missing-indicator method) produce biased results, whereas imputation techniques yield valid results without complicating the analysis once the imputations are carried out. Imputation techniques are based on the idea that any subject in a study sample can be replaced by a new randomly chosen subject from the same source population. Imputation of missing data on a variable is replacing that missing by a value that is drawn from an estimate of the distribution of this variable. In single imputation, only one estimate is used. In multiple imputation, various estimates are used, reflecting the uncertainty in the estimation of this distribution. Under the general conditions of so-called missing at random and missing completely at random, both single and multiple imputations result in unbiased estimates of study associations. But single imputation results in too small estimated standard errors, whereas multiple imputation results in correctly estimated standard errors and confidence intervals. In this article we explain why all this is the case, and use a simple simulation study to demonstrate our explanations. We also explain and illustrate why two frequently used methods to handle missing data, i.e., overall mean imputation and the missing-indicator method, almost always result in biased estimates.

Additional Metadata
Keywords	Bias, Indicator method, Missing data, Multiple imputation, Precision, Single imputation
Persistent URL	doi.org/10.1016/j.jclinepi.2006.01.014, hdl.handle.net/1765/63193
Journal	Journal of Clinical Epidemiology
Organisation	Erasmus MC: University Medical Center Rotterdam
Citation APA Style AAA Style APA Style Cell Style Chicago Style Harvard Style IEEE Style MLA Style Nature Style Vancouver Style American-Institute-of-Physics Style Council-of-Science-Editors Style BibTex Format Endnote Format RIS Format CSL Format DOIs only Format	Donders, A. R. T., van der Heijden, G., Stijnen, T., & Moons, K. (2006). Review: A gentle introduction to imputation of missing values. Journal of Clinical Epidemiology, 59(10), 1087–1091. doi:10.1016/j.jclinepi.2006.01.014

Review: A gentle introduction to imputation of missing values

Publication

Publication

About

Review: A gentle introduction to imputation of missing values

Publication

Publication

Workflow

Workflow

Add Content