Identifying intervention strategies from machine learning models with COALA: a counterfactual optimization framework
Han, B.; Duan, Q.; Hu, T.
Show abstract
MotivationMachine learning models in biomedicine have become increasingly complex, often functioning as black boxes. However, understanding contributors to disease and making actionable health interventions requires interpretable models. Common explainable AI methods like SHAP focus on feature importance but fall short in explaining why features contribute in certain patterns or what interventions to take. Counterfactual explanations address this by proposing "what if" scenarios but current tools focus on individual predictions and fail to generalize complex trends. ResultsWe propose the framework Counterfactual Optimization for Actionable interpretabiLity in AI (COALA). COALA interprets models by identifying optimal counterfactuals across user-defined mutable feature subsets and constraining remaining features to reveal how constraint features determine what interventions are optimal. By analyzing counterfactual profiles of features rather than individual features, COALA reveals holistic patterns. Using synthetic and real datasets, COALA reveals simple and complex model trends and provides more intuitive, multi-feature interventions than SHAP. Availability and ImplementationCode for COALA implementation, synthetic data, models trained on synthetic data, and code to replicate results and figures are available at https://github.com/brt-solo/COALA.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Enabling interpretable machine learning for biological data with reliability scores 95%
- Using random forests to uncover the predictive power of distance-varying cell interactions in tumor microenvironments 94%
- Artificial neural networks for model identification and parameter estimation in computational cognitive models 93%
Similar papers in this journal
- Phenotype Driven Data Augmentation Methods for Transcriptomic Data 94%
- Prediction of the infecting organism in peritoneal dialysis patients with acute peritonitis using interpretable Tsetlin Machines 93%
- Beyond synthetic lethality in large-scale metabolic and regulatory network models via genetic minimal intervention sets 93%
Similar papers in this journal
- HARVESTMAN: A framework for hierarchical featurelearning and selection from whole genome sequencingdata 95%
- COMIC: Explainable Drug Repurposing via Contrastive Masking for Interpretable Connections 94%
- DAGBagM: Learning directed acyclic graphs of mixed variables with an application to identify prognostic protein biomarkers in ovarian cancer 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.