Multiple imputation in data that grow over time: A comparison of three strategies
Publication date
2022
Editors
Advisors
Supervisors
Document Type
Article
Metadata
Show full item recordCollections
License
cc_by_nc_nd
Abstract
Multiple imputation is a recommended technique to deal with missing data. We study the problem where the investigator has already created imputations before the arrival of the next wave of data. The newly arriving data contain missing values that need to be imputed. The standard method (RE-IMPUTE) is to combine the new and old data before imputation, and re-impute all missing values in the combined data. We study the properties of two methods that impute the missing data in the new part only, thus preserving the historic imputations. Method NEST multiply imputes the new data conditional on each filled-in old data (Formula presented.) times. Method APPEND is the special case of NEST with (Formula presented.) thus appending each filled-in data by single imputation. We found that NEST and APPEND have the same validity as RE-IMPUTE for monotone missing data-patterns. NEST and APPEND also work well when relations within waves are stronger than between waves and for moderate percentages of missing data. We do not recommend the use of NEST or APPEND when relations within time points are weak and when associations between time points are strong.
Keywords
Missing data, congeniality, multiple imputation, nested imputation, Statistics and Probability, Experimental and Cognitive Psychology, Arts and Humanities (miscellaneous)
Citation
Kavelaars, X M, van Ginkel, J R & van Buuren, S 2022, 'Multiple imputation in data that grow over time: A comparison of three strategies', Multivariate Behavioral Research, vol. 57, no. 2-3, pp. 513-523. https://doi.org/10.1080/00273171.2021.1912582