• Home
  • Search
  • Multiple imputation analysis of case–cohort studies
  • Open Access IconOpen Access
  • Cite Icon46
  • https://doi.org/10.1002/sim.4130Copy DOI Icon

Multiple imputation analysis of case–cohort studies

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The usual methods for analyzing case-cohort studies rely on sometimes not fully efficient weighted estimators. Multiple imputation might be a good alternative because it uses all the data available and approximates the maximum partial likelihood estimator. This method is based on the generation of several plausible complete data sets, taking into account uncertainty about missing values. When the imputation model is correctly defined, the multiple imputation estimator is asymptotically unbiased and its variance is correctly estimated. We show that a correct imputation model must be estimated from the fully observed data (cases and controls), using the case status among the explanatory variable. To validate the approach, we analyzed case-cohort studies first with completely simulated data and then with case-cohort data sampled from two real cohorts. The analyses of simulated data showed that, when the imputation model was correct, the multiple imputation estimator was unbiased and efficient. The observed gain in precision ranged from 8 to 37 per cent for phase-1 variables and from 5 to 19 per cent for the phase-2 variable. When the imputation model was misspecified, the multiple imputation estimator was still more efficient than the weighted estimators but it was also slightly biased. The analyses of case-cohort data sampled from complete cohorts showed that even when no strong predictor of the phase-2 variable was available, the multiple imputation was unbiased, as precised as the weighted estimator for the phase-2 variable and slightly more precise than the weighted estimators for the phase-1 variables. However, the multiple imputation estimator was found to be biased when, because of interaction terms, some coefficients of the imputation model had to be estimated from small samples. Multiple imputation is an efficient technique for analyzing case-cohort data. Practically, we suggest building the analysis model using only the case-cohort data and weighted estimators. Multiple imputation can eventually be used to reanalyze the data using the selected model in order to improve the precision of the results.

Similar Papers
  • PDF
  • Research Article
  • Citations21

Multiple imputation of missing data under missing at random: compatible imputation models are not sufficient to avoid bias if they are mis-specified

  • Jun 19, 2023
  • Journal of Clinical Epidemiology
  • Elinor Curnow +7
  • Research Article
  • Citations56

Selecting the model for multiple imputation of missing data: Just use an IC!

  • Feb 24, 2021
  • Statistics in Medicine
  • Firouzeh Noghrehchi +3
  • Research Article
  • Citations77

A Simplified Framework for Using Multiple Imputation in Social Work Research

  • Sep 01, 2008
  • Social Work Research
  • R A Rose +1
  • PDF
  • Research Article
  • Citations171

Outcome-sensitive multiple imputation: a simulation study

  • Jan 09, 2017
  • BMC Medical Research Methodology
  • Evangelos Kontopantelis +3
  • Research Article
  • Citations8

Accounting for bias due to outcome data missing not at random: comparison and illustration of two approaches to probabilistic bias analysis: a simulation study

  • Nov 13, 2024
  • BMC Medical Research Methodology
  • Emily Kawabata +13
  • Research Article
  • Citations112

Dealing with missing covariates in epidemiologic studies: a comparison between multiple imputation and a full Bayesian approach

  • Apr 04, 2016
  • Statistics in Medicine
  • Nicole S Erler +5
  • Research Article

Accommodating the Analysis Model in Multiple Imputation for the Weibull Mixture Cure Model: Performance Under Penalized Likelihood

  • Mar 01, 2026
  • Statistics in Medicine
  • Changchang Xu +3
  • Research Article
  • Citations28

Severe asthma in the US population and eligibility for mAb therapy

  • Dec 19, 2019
  • Journal of Allergy and Clinical Immunology
  • Ayobami Akenroye +2
  • Research Article
  • Citations127

9. Multiple Imputation of Incomplete Categorical Data Using Latent Class Analysis

  • Jul 01, 2008
  • Sociological Methodology
  • Jeroen K Vermunt +3
  • Research Article
  • Citations1

Simulation-Based Evaluation of Methods for Handling Nonwear Time in Accelerometer Studies of Physical Activity

  • Jul 12, 2022
  • Journal for the measurement of physical behaviour
  • Kristopher I Kapphahn +4
  • Research Article
  • Citations1

A Comparison of Multiple Imputation and Optimal Estimation for Missing and Uncertain Urban Air Toxics Data

  • Nov 01, 2006
  • Epidemiology
  • H Le +6
  • PDF
  • Research Article
  • Citations21

Should multiple imputation be stratified by exposure group when estimating causal effects via outcome regression in observational studies?

  • Feb 16, 2023
  • BMC Medical Research Methodology
  • Jiaxin Zhang +4
  • Research Article
  • Citations1005

The proportion of missing data should not be used to guide decisions on multiple imputation

  • Mar 13, 2019
  • Journal of Clinical Epidemiology
  • Paul Madley-Dowd +3
  • PDF
  • Research Article
  • Citations5

Multiple imputation strategies for a bounded outcome variable in a competing risks analysis.

  • Jan 19, 2021
  • Statistics in medicine
  • Elinor Curnow +5
  • Dissertation

Multiple imputation in high-dimensional data with variable selection

  • Aug 01, 2022
  • Qiushuang Li
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.