Applications of multiple imputation in medical studies: from AIDS to NHANES
- 1 February 1999
- journal article
- review article
- Published by SAGE Publications in Statistical Methods in Medical Research
- Vol. 8 (1) , 17-36
- https://doi.org/10.1177/096228029900800103
Abstract
Rubin's multiple imputation is a three-step method for handling complex missing data, or more generally, incomplete-data problems, which arise frequently in medical studies. At the first step, m (> 1) completed-data sets are created by imputing the unobserved data m times using m independent draws from an imputation model, which is constructed to reasonably approximate the true distributional relationship between the unobserved data and the available information, and thus reduce potentially very serious nonresponse bias due to systematic difference between the observed data and the unobserved ones. At the second step, m complete-data analyses are performed by treating each completed-data set as a real complete-data set, and thus standard complete-data procedures and software can be utilized directly. At the third step, the results from the m complete-data analyses are combined in a simple, appropriate way to obtain the so-called repeated-imputation inference, which properly takes into account the uncertainty in the imputed values. This paper reviews three applications of Rubin's method that are directly relevant for medical studies. The first is about estimating the reporting delay in acquired immune deficiency syndrome (AIDS) surveillance systems for the purpose of estimating survival time after AIDS diagnosis. The second focuses on the issue of missing data and noncompliance in randomized experiments, where a school choice experiment is used as an illustration. The third looks at handling nonresponse in United States National Health and Nutrition Examination Surveys (NHANES). The emphasis of our review is on the building of imputation models (i.e. the first step), which is the most fundamental aspect of the method.Keywords
This publication has 31 references indexed in Scilit:
- A Broader Template for Analyzing Broken Randomized ExperimentsSociological Methods & Research, 1998
- Ellipsoidally symmetric extensions of the general location model for mixed categorical and continuous dataBiometrika, 1998
- Bayesian inference for causal effects in randomized experiments with noncomplianceThe Annals of Statistics, 1997
- Multiple Imputation after 18+ YearsJournal of the American Statistical Association, 1996
- Performing likelihood ratio tests with multiply-imputed data setsBiometrika, 1992
- Multiple imputation in health‐are databases: An overview and some applicationsStatistics in Medicine, 1991
- Zidovudine in Asymptomatic Human Immunodeficiency Virus InfectionNew England Journal of Medicine, 1990
- Statistics and Causal InferenceJournal of the American Statistical Association, 1986
- Bayesianly Justifiable and Relevant Frequency Calculations for the Applied StatisticianThe Annals of Statistics, 1984
- Inference and missing dataBiometrika, 1976