We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ML

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Statistics > Machine Learning

Title: Accurate Shapley Values for explaining tree-based models

Abstract: Although Shapley Values (SV) are widely used in explainable AI, they can be poorly understood and estimated, implying that their analysis may lead to spurious inferences and explanations. As a starting point, we remind an invariance principle for SV and derive the correct approach for computing the SV of categorical variables that are particularly sensitive to the encoding used. In the case of tree-based models, we introduce two estimators of Shapley Values that exploit the tree structure efficiently and are more accurate than state-of-the-art methods. Simulations and comparisons are performed with state-of-the-art algorithms and show the practical gain of our approach. Finally, we discuss the ability of SV to provide reliable local explanations. We also provide a Python package that computes our estimators at this https URL
Comments: Accepted at the 25th International Conference on Artificial Intelligence and Statistics (AISTATS), 2022. V2: The section on Active Shapley Values has been removed in this updated version
Subjects: Machine Learning (stat.ML); Machine Learning (cs.LG)
Journal reference: AISTATS 2022
Cite as: arXiv:2106.03820 [stat.ML]
  (or arXiv:2106.03820v2 [stat.ML] for this version)

Submission history

From: Salim I. Amoukou [view email]
[v1] Mon, 7 Jun 2021 17:35:54 GMT (1034kb,D)
[v2] Thu, 24 Mar 2022 16:37:58 GMT (2143kb,D)

Link back to: arXiv, form interface, contact.