Fast Estimation of Partial Dependence Functions using Trees

Jinyang Liu*, Tessa Steensgaard, Marvin Wright, Niklas Pfister, Munir Hiabu

*Corresponding author af dette arbejde

Publikation: Bidrag til bog/antologi/rapportKonferencebidrag i proceedingsForskningpeer review

6 Downloads (Pure)

Abstract

Many existing interpretation methods are based on Partial Dependence (PD) functions that, for a pre-trained machine learning model, capture how a subset of the features affects the predictions by averaging over the remaining features. Notable methods include Shapley additive explanations (SHAP) which computes feature contributions based on a game theoretical interpretation and PD plots (i.e., 1-dim PD functions) that capture average marginal main effects. Recent work has connected these approaches using a functional decomposition and argues that SHAP values can be misleading since they merge main and interaction effects into a single local effect. However, a major advantage of SHAP compared to other PD-based interpretations has been the availability of fast estimation techniques, such as TreeSHAP. In this paper, we propose a new tree-based estimator, FastPD*, which efficiently estimates arbitrary PD functions. We show that FastPD consistently estimates the desired population quantity – in contrast to path-dependent TreeSHAP which is inconsistent when features are correlated. For moderately deep trees, FastPD improves the complexity of existing methods from quadratic to linear in the number of observations. By estimating PD functions for arbitrary feature subsets, FastPD can be used to extract PD-based interpretations such as SHAP, PD plots and higher-order interaction effects.

OriginalsprogEngelsk
TitelProceedings of the 42nd International Conference on Machine Learning
Antal sider39
ForlagPMLR
Publikationsdato2025
Sider39496-39534
StatusUdgivet - 2025
Begivenhed42nd International Conference on Machine Learning, ICML 2025 - Vancouver, Canada
Varighed: 13 jul. 202519 jul. 2025

Konference

Konference42nd International Conference on Machine Learning, ICML 2025
Land/OmrådeCanada
ByVancouver
Periode13/07/202519/07/2025
NavnProceedings of Machine Learning Research
Vol/bind267
ISSN2640-3498

Bibliografisk note

Publisher Copyright:
© 2025, by the authors.

Citationsformater