Discernibility in explanations: an approach to designing more acceptable and meaningful machine learning models for medicine
Wang, H.; Aligon, J.; May, J.; Doumard, E.; Labroche, N.; Delpierre, C.; Soule-Dupuy, C.; Casteilla, L.; Planat, V.; Monsarrat, P.
Show abstract
BackgroundAlthough the benefits of machine learning (ML) are undeniable in health-care, explainability plays a vital role in improving transparency and understanding the most decisive and persuasive variables for prediction. The challenge is to identify explanations that make sense to the biomedical expert. This work proposes discernibility as a new approach to faithfully reflect human cognition, with the users perception of a relationship between explanations and data for a given variable. MethodsA total of 50 participants (19 biomedical and 31 data scientists) evaluated their perception of the discernibility of explanations from both synthetic and human-based dataset (National Health and Nutrition Examination Survey). The inter-rater reliability was tested through the intraclass correlation coefficient (ICC). 13 statistical coefficients were considered to be able to capture for a given variable the relationship between its values and its explanations. A Passing-Bablok regression was performed for each user to highlight the consistency between user rating and each coefficient. FindingsThe low inter-rater reliability of discernibility (ICC{inverted exclamation}0.5) with no difference between areas of expertise or level of education underlines the need for an objective metric of discernibility. Among all evaluated metrics, dcor metric was found to be the most suitable to capture the intra-individual reliability of discernibility perceived by users (median slope closer to 1 and a narrower confidence interval width for the Passing-Bablok regression with the lowest differential bias between the most and least discernible values). Interpretationdcor was shown to be a reliable metric for assessing the discernibility of explanations, effectively capturing the clarity of the relationship between the data and their explanations, and providing clues to the underlying pathophysiological mechanisms that are not immediately apparent when examining individual predictors. Discernibility can also serve as an evaluation metric for model quality, used to prevent overfitting or aid in feature selection, providing medical practitioners with more accurate and persuasive results.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of Myocardial Infarction (MI) Probability from Imbalanced Medical Survey Data: An Artificial Neural Network (ANN) with Explainable AI (XAI) Insights 96%
- A machine-learning Approach for Stress Detection Using Wearable Sensors in Free-living Environments 95%
- Unsupervised Discovery of Risk Profiles on Negative and Positive COVID-19 Hospitalized Patients 95%
Similar papers in this journal
Similar papers in this journal
- Empirical methods for the validation of Time-To-Event mathematical models taking into account uncertainty and variability: Application to EGFR+ Lung Adenocarcinoma. 93%
- Leveraging Permutation Testing to Assess Confidence in Positive-Unlabeled Learning Applied to High-Dimensional Biological Datasets 92%
- Fast and robust imputation for miRNA expression data using constrained least squares 92%
Similar papers in this journal
- Benchmarking feature selection and feature extraction methods to improve the performances of machine-learning algorithms for patient classification using metabolomics biomedical data. 95%
- iMDA-BN: Identification of miRNA-Disease Associations based on the Biological Network and Graph Embedding Algorithm 93%
- Topological embedding and directional feature importance in ensemble classifiers for multi-class classification 92%
Similar papers in this journal
- Prediction of Sepsis Mortality in ICU Patients Using Machine Learning Methods 96%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 95%
- Combining symbolic regression with the Cox proportional hazards model improves prediction of heart failure deaths 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.