We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ME

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Statistics > Methodology

Title: Generalizing treatment effects with incomplete covariates: identifying assumptions and multiple imputation algorithms

Abstract: We focus on the problem of generalizing a causal effect estimated on a randomized controlled trial (RCT) to a target population described by a set of covariates from observational data. Available methods such as inverse propensity sampling weighting are not designed to handle missing values, which are however common in both data sources. In addition to coupling the assumptions for causal effect identifiability and for the missing values mechanism and to defining appropriate estimation strategies, one difficulty is to consider the specific structure of the data with two sources and treatment and outcome only available in the RCT. We propose three multiple imputation strategies to handle missing values when generalizing treatment effects, each handling the multi-source structure of the problem differently (separate imputation, joint imputation with fixed effect, joint imputation ignoring source information). As an alternative to multiple imputation, we also propose a direct estimation approach that treats incomplete covariates as semi-discrete variables. The multiple imputation strategies and the latter alternative rely on different sets of assumptions concerning the impact of missing values on identifiability. We discuss these assumptions and assess the methods through an extensive simulation study. This work is motivated by the analysis of a large registry of over 20,000 major trauma patients and an RCT studying the effect of tranexamic acid administration on mortality in major trauma patients admitted to ICU. The analysis illustrates how the missing values handling can impact the conclusion about the effect generalized from the RCT to the target population.
Comments: preprint, 38 pages, 14 figures
Subjects: Methodology (stat.ME); Applications (stat.AP)
MSC classes: 62P10, 93C41
ACM classes: G.3
Cite as: arXiv:2104.12639 [stat.ME]
  (or arXiv:2104.12639v4 [stat.ME] for this version)

Submission history

From: Imke Mayer [view email]
[v1] Mon, 26 Apr 2021 15:09:58 GMT (18447kb)
[v2] Mon, 6 Sep 2021 08:33:09 GMT (8694kb,D)
[v3] Sat, 26 Mar 2022 00:06:52 GMT (2302kb,D)
[v4] Fri, 24 Feb 2023 13:51:00 GMT (12872kb,D)

Link back to: arXiv, form interface, contact.