Back

Understanding binary classifier model structure based on Shapley feature interaction patterns

Zhao, B.

2021-03-30 systems biology
10.1101/2021.03.29.437591 bioRxiv
Show abstract

There is increasing emphasis on the interpretability of machine learning models, including in understanding biological systems. The well-known Shapley value framework based on game theory works in principle with any models to attribute feature importance. While feature interactions are critical to understand and can be interpreted within this framework, much attention is paid in practice on global feature importance and general trends of interactions. The inter-relationships between underlying model structure and Shapley value and its decomposition is less clear. Here we use binary classifiers to systematically examine how logical and additive interactions affect marginal contributions. These decomposed main and interaction effects are reflected in resulting Shapley dependence plots. The directionality of inequalities or logical/additive operators influence independently the main and marginal/interaction effects. Lastly, we show that these principles are applicable for models with noise.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.