Imputation of missing longitudinal data: a comparison of methods

被引:331
作者
Engels, JM
Diehr, P
机构
[1] Univ Washington, Dept Biostat, Seattle, WA 98195 USA
[2] Univ Washington, Dept Hlth Serv, Seattle, WA 98195 USA
关键词
missing data; imputation; longitudinal; depression; cohort;
D O I
10.1016/S0895-4356(03)00170-7
中图分类号
R19 [保健组织与事业(卫生事业管理)];
学科分类号
摘要
Background and Objective: Missing information is inevitable in longitudinal studies, and can result in biased estimates and a loss of power. One approach to this problem is to impute the missing data to yield a more complete data set. Our goal was to compare the performance of 14 methods of imputing missing data on depression, weight, cognitive functioning, and self-rated health in a longitudinal cohort of older adults. Methods: We identified situations where a person had a known value following one or more missing values, and treated the known value as a "missing value." This "missing value" was imputed using each method and compared to the observed value. Methods were compared on the root mean square error, mean absolute deviation, bias, and relative variance of the estimates. Results: Most imputation methods were biased toward estimating the "missing value" as too healthy, and most estimates had a variance that was too low. Imputed values based on a person's values before and after the "missing value" were superior to other methods, followed by imputations based on a person's values before the "missing value." Imputations that used no information specific to the person, such as using the sample mean, had the worst performance. Conclusions: We conclude that, in longitudinal studies where the overall trend is for worse health over time and where missing data can be assumed to be primarily related to worse health, missing data in a longitudinal sequence should be imputed from the available longitudinal data for that person. (C) 2003 Elsevier Inc. All rights reserved.
引用
收藏
页码:968 / 976
页数:9
相关论文
共 18 条
[1]  
[Anonymous], 1983, INCOMPLETE DATA SAMP
[2]  
Brick J M, 1996, Stat Methods Med Res, V5, P215, DOI 10.1177/096228029600500302
[3]  
BROOKS CA, 1978, 3 US DEP COMM
[4]  
Cox BG, 1985, METHODOLOGICAL ISSUE
[5]   Body mass index and mortality in nonsmoking older adults: The cardiovascular health study [J].
Diehr, P ;
Bild, DE ;
Harris, TB ;
Duxbury, A ;
Siscovick, D ;
Rossi, M .
AMERICAN JOURNAL OF PUBLIC HEALTH, 1998, 88 (04) :623-629
[6]  
Fairclough DL, 1998, STAT MED, V17, P667, DOI 10.1002/(SICI)1097-0258(19980315/15)17:5/7<667::AID-SIM813>3.3.CO
[7]  
2-Y
[8]   MINI-MENTAL STATE - PRACTICAL METHOD FOR GRADING COGNITIVE STATE OF PATIENTS FOR CLINICIAN [J].
FOLSTEIN, MF ;
FOLSTEIN, SE ;
MCHUGH, PR .
JOURNAL OF PSYCHIATRIC RESEARCH, 1975, 12 (03) :189-198
[9]  
Fried Linda P., 1991, Annals of Epidemiology, V1, P263
[10]  
Kalton G., 1983, Compensating for missing survey data