We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ML

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Statistics > Machine Learning

Title: Causal Effect Identification from Multiple Incomplete Data Sources: A General Search-based Approach

Abstract: Causal effect identification considers whether an interventional probability distribution can be uniquely determined without parametric assumptions from measured source distributions and structural knowledge on the generating system. While complete graphical criteria and procedures exist for many identification problems, there are still challenging but important extensions that have not been considered in the literature. To tackle these new settings, we present a search algorithm directly over the rules of do-calculus. Due to generality of do-calculus, the search is capable of taking more advanced data-generating mechanisms into account along with an arbitrary type of both observational and experimental source distributions. The search is enhanced via a heuristic and search space reduction techniques. The approach, called do-search, is provably sound, and it is complete with respect to identifiability problems that have been shown to be completely characterized by do-calculus. When extended with additional rules, the search is capable of handling missing data problems as well. With the versatile search, we are able to approach new problems such as combined transportability and selection bias, or multiple sources of selection bias. We perform a systematic analysis of bivariate missing data problems and study causal inference under case-control design. We also present the R package dosearch that provides an interface for a C++ implementation of the search.
Comments: This is the version published in the Journal of Statistical Software
Subjects: Machine Learning (stat.ML); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Journal reference: Journal of Statistical Software, 99(5):1-40, 2021
DOI: 10.18637/jss.v099.i05
Cite as: arXiv:1902.01073 [stat.ML]
  (or arXiv:1902.01073v5 [stat.ML] for this version)

Submission history

From: Santtu Tikka [view email]
[v1] Mon, 4 Feb 2019 08:12:04 GMT (234kb,D)
[v2] Thu, 28 Feb 2019 08:19:27 GMT (235kb,D)
[v3] Thu, 5 Dec 2019 13:36:36 GMT (267kb,D)
[v4] Mon, 26 Oct 2020 10:00:07 GMT (268kb,D)
[v5] Fri, 27 Aug 2021 09:41:11 GMT (579kb,D)

Link back to: arXiv, form interface, contact.