NEXIM: A Nash Equilibrium-Based Framework for Stable Explainable AI in Medical Applications
Upadhyaya, D. P.; Sahoo, S. S.; Prantzalos, K.; Golnari, P.
Show abstract
Reliable explanations are important for trustworthy medical applications of artificial intelligence (AI), but attribution-based explanations can vary across model randomization and small analytic changes. We present NEXIM (Nash Equilibrium-based Explainability and Interpretability Model), implemented here as an accuracy-constrained, equilibrium-inspired model-selection framework that jointly evaluates held-out prediction error, explanation stability, and cross-model connectivity. The implementation evaluated ten GradientBoostingRegressor models per prediction horizon, differing only by random seed (0-9), using a fixed 75/25 patient split. Kernel SHAP attribution vectors were compared using Spearman rank correlation, and graph connectivity summarized whether each model belonged to a dense explanation-similarity region. Candidate models within 0.02 Montreal Cognitive Assessment points of the best root mean squared error (RMSE) were ranked using a multiplicative Explanation Equilibrium Score. In longitudinal Parkinson's Progression Markers Initiative data, NEXIM selected the RMSE-optimal model at the one- and three-year horizons. At the two-year horizon, it selected Model 4 rather than the RMSE-only Model 8, increasing scaled stability from 0.8757 to 0.8847 and normalized graph connectivity from 0.889 to 1.000 while increasing RMSE by only 0.0014. The two models retained the same top-20 feature set but differed modestly in feature order, illustrating that NEXIM primarily acted as a reproducibility screen rather than identifying clinically contradictory explanations. Stability and consensus are treated as reproducibility criteria, not evidence of causal faithfulness, clinical usefulness, or improved patient outcomes. NEXIM may therefore serve as a governance checkpoint for model refresh and documentation, but external validation, stronger model-family baselines, and prospective clinical evaluation remain necessary.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Interpretable deep learning approach for extracting cognitive features from hand-drawn images of intersecting pentagons in older adults 93%
- Continuous-Time and Dynamic Suicide Attempt Risk Prediction with Neural Ordinary Differential Equations 93%
- Conformal prediction enables disease course prediction and allows individualized diagnostic uncertainty in multiple sclerosis 92%
Similar papers in this journal
- Assessing the transportability of clinical prediction models for cognitive impairment using causal models 94%
- Comparing randomized trial designs to estimate treatment effect in rare diseases with longitudinal models: a simulation study showcased by Autosomal Recessive Cerebellar Ataxias using the SARA score 91%
- Completeness of reporting of clinical prediction models developed using supervised machine learning: A systematic review 90%
Similar papers in this journal
- Evaluation of Domain Generalization and Adaptation on Improving Model Robustness to Temporal Dataset Shift in Clinical Medicine 93%
- Risk factors for severe COVID-19 differ by age: a retrospective study of hospitalized adults 92%
- Machine learning approach to dynamic risk modeling of mortality in COVID-19: a UK Biobank study 92%
Similar papers in this journal
- Enhancing Early Detection of Cognitive Decline in the Elderly through Ensemble of NLP Techniques: A Comparative Study Utilizing Large Language Models in Clinical Notes 93%
- Predicting the functional effects of voltage-gated potassium channel missense variants with multi-task learning 90%
- Consistent Performance of GPT-4o in Rare Disease Diagnosis Across Nine Languages and 4967 Cases 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.