Multiple imputation in data that grow over time: A comparison of three strategies

Publication date

2022

Authors

Kavelaars, X.M.
van Ginkel, J.R.
van Buuren, S.ORCID 0000-0003-1098-2119ISNI 0000000032712898

Editors

Advisors

Supervisors

Document Type

Article
Open Access logo

License

cc_by_nc_nd

Abstract

Multiple imputation is a recommended technique to deal with missing data. We study the problem where the investigator has already created imputations before the arrival of the next wave of data. The newly arriving data contain missing values that need to be imputed. The standard method (RE-IMPUTE) is to combine the new and old data before imputation, and re-impute all missing values in the combined data. We study the properties of two methods that impute the missing data in the new part only, thus preserving the historic imputations. Method NEST multiply imputes the new data conditional on each filled-in old data (Formula presented.) times. Method APPEND is the special case of NEST with (Formula presented.) thus appending each filled-in data by single imputation. We found that NEST and APPEND have the same validity as RE-IMPUTE for monotone missing data-patterns. NEST and APPEND also work well when relations within waves are stronger than between waves and for moderate percentages of missing data. We do not recommend the use of NEST or APPEND when relations within time points are weak and when associations between time points are strong.

Keywords

Missing data, congeniality, multiple imputation, nested imputation, Statistics and Probability, Experimental and Cognitive Psychology, Arts and Humanities (miscellaneous)

Citation

Kavelaars, X M, van Ginkel, J R & van Buuren, S 2022, 'Multiple imputation in data that grow over time: A comparison of three strategies', Multivariate Behavioral Research, vol. 57, no. 2-3, pp. 513-523. https://doi.org/10.1080/00273171.2021.1912582