Missing Data Analysis Using Multiple Imputation
Author(s) -
Yulei He
Publication year - 2010
Publication title -
circulation cardiovascular quality and outcomes
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 2.692
H-Index - 87
eISSN - 1941-7713
pISSN - 1941-7705
DOI - 10.1161/circoutcomes.109.875658
Subject(s) - missing data , imputation (statistics) , medicine , inference , data mining , data science , computer science , machine learning , artificial intelligence
Missing data are a pervasive problem in health investigations. We describe some background of missing data analysis and criticize ad hoc methods that are prone to serious problems. We then focus on multiple imputation, in which missing cases are first filled in by several sets of plausible values to create multiple completed datasets, then standard complete-data procedures are applied to each completed dataset, and finally the multiple sets of results are combined to yield a single inference. We introduce the basic concepts and general methodology and provide some guidance for application. For illustration, we use a study assessing the effect of cardiovascular diseases on hospice discussion for late stage lung cancer patients.
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom