16 papers · ranked by Valyu relevance
Gabriel Laberge, Yann Pequignot
Shapley values are ubiquitous in interpretable Machine Learning due to their strong theoretical background and efficient implementation in the SHAP library. Computing these values previously induced an exponential cost with respect to the number of input features of an opaque model. Now, with efficient implementations…
Rory Mitchell, Eibe Frank, Geoffrey Holmes
SHAP (SHapley Additive exPlanation) values (Lundberg and Lee, 2017) provide a game theoretic interpretation of the predictions of machine learning models based on Shapley values (Shapley, 1953). While exact calculation of SHAP values is computationally intractable in general, a recursive polynomialtime algorithm called…
Jilei Yang
SHAP (SHapley Additive exPlanation) values are one of the leading tools for interpreting machine learning models, with strong theoretical guarantees (consistency, local accuracy) and a wide availability of implementations and use cases. Even though computing SHAP values takes exponential time in general, TreeSHAP takes…
Akshat Dubey, Aleksandar Anžel, Georges Hattab
The field of health informatics has been profoundly influenced by the development of random forest models, which have led to significant advances in the interpretability of feature interactions. These models are characterized by their robustness to overfitting and parallelization, making them particularly useful in…
Di Fan, Ayan Biswas, James Ahrens
Wildfires present intricate challenges for prediction, necessitating the use of sophisticated machine learning techniques for effective modeling[1]. In our research, we conducted a thorough assessment of various machine learning algorithms for both classification and regression tasks relevant to predicting wildfires.…
Peng Yu, Chao Xu, Albert Bifet, Jesse Read
Decision trees are well-known due to their ease of interpretability. To improve accuracy, we need to grow deep trees or ensembles of trees. These are hard to interpret, offsetting their original benefits. Shapley values have recently become a popular way to explain the predictions of tree-based machine learning models.…
Giulia Di Teodoro, Marta Monaci, Laura Palagi
The interpretability of models has become a crucial issue in Machine Learning because of algorithmic decisions' growing impact on real-world applications. Tree ensemble methods, such as Random Forests or XgBoost, are powerful learning tools for classification tasks. However, while combining multiple trees may provide…
T. Campbell, Heinrich Röder, Robert W. Georgantas, Joanna Roder
- 1. Thomas W. Campbell, Biodesix. Contributions: conceptualization, formal analysis, methodology, software, writing – original draft, writing – review and editing. Corresponding Author. Address: 2970 Wilderness Place, Suite 100, Boulder, CO 80301 - 2. Heinrich Roder, Biodesix. Contributions: conceptualization, writing…
M. Mayer
An important technique to explore a black-box machine learning (ML) model is called SHAP (SHapley Additive exPlanation). SHAP values decompose predictions into contributions of the features in a fair way. We will show that for a boosted trees model with some or all features being additively modeled, the SHAP dependence…
Abhineet Agarwal, Yan Shuo Tan, Omer Ronen, Chandan Singh + 1 more
Tree-based models such as decision trees and random forests (RF) are a cornerstone of modern machine-learning practice. To mitigate overfitting, trees are typically regularized by a variety of techniques that modify their structure (e.g. pruning). We introduce Hierarchical Shrinkage (HS), a post-hoc algorithm that does…
Houtao Deng
Tree ensembles such as random forests and boosted trees are accurate but difficult to understand, debug and deploy. In this work, we provide the inTrees (interpretable trees) framework that extracts, measures, prunes and selects rules from a tree ensemble, and calculates frequent variable interactions. An rule-based…
Bénard, Clément
Tree ensembles have demonstrated state-of-the-art predictive performance across a wide range of problems involving tabular data. Nevertheless, the black-box nature of tree ensembles is a strong limitation, especially for applications with critical decisions at stake. The Hoeffding or ANOVA functional decomposition is a…
Nathan Wycoff
Regression trees have emerged as a preeminent tool for solving real-world regression problems due to their ability to deal with nonlinearities, interaction effects and sharp discontinuities. In this article, we rather study regression trees applied to well-behaved, differentiable functions, and determine the…
Lorenzo Bonasera, Emilio Carrizosa
Integer Programming Authors: ['Lorenzo Bonasera' 'Emilio Carrizosa'] Tree ensemble methods represent a popular machine learning model, known for their effectiveness in supervised classification and regression tasks. Their performance derives from aggregating predictions of multiple decision trees, which are renowned…
Tim Räz
The interpretability of ML models is important, but it is not clear what it amounts to. So far, most philosophers have discussed the lack of interpretability of black-box models such as neural networks, and methods such as explainable AI that aim to make these models more transparent. The goal of this paper is to…
Sondag, Max, Meinecke, Christofer + 6 more
Fig. 1: An overview of the visual analytics system. A) shows global information such as feature and class distribution, model accuracy, and predictions. Furthermore, it allows control of the number of clusters that are displayed and shows a 2D projection of the clustered trees. B) The Feature Plot (top) shows how often…