A Systematic Fairness Evaluation of Racial Bias in Alzheimers Disease Diagnosis Using Machine Learning Models
Baddam, N. G.; Pijani, B. A.; Bozdag, S.
Show abstract
INTRODUCTIONAlzheimers disease (AD) is a major global health concern, expected to affect 12.7 million Americans by 2050. Machine learning (ML) algorithms have been developed for AD diagnosis and progression prediction, but the lack of racial diversity in clinical datasets raises concerns about their generalizability across demographic groups, particularly underrepresented populations. Studies show ML algorithms can inherit biases from data, leading to biased AD predictions. METHODSThis study investigates the fairness of ML models in AD diagnosis. We hypothesize that models trained on a single racial group perform well within that group but poorly in others. We employ feature selection and model training techniques to improve fairness. RESULTSOur findings support our hypothesis that ML models trained on one group underperform on others. We also demonstrated that applying fairness techniques to ML models reduces their bias. DISCUSSIONThis study highlights the need for racial diversity in datasets and fair models for AD prediction.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Random forest model for feature-based Alzheimer's disease conversion prediction from early mild cognitive impairment subjects 94%
- Compressive Big Data Analytics: An Ensemble Meta-Algorithm for High-dimensional Multisource Datasets 94%
- Robust Disease Prognosis via Diagnostic Knowledge Preservation: A Sequential Learning Approach 93%
Similar papers in this journal
- Enhancing Fairness in Disease Prediction by Optimizing Multiple Domain Adversarial Networks 95%
- Uncovering the effects of model initialization on deep model generalization: A study with adult and pediatric chest X-ray images 94%
- Evaluating and mitigating unfairness in multimodal remote mental health assessments 94%
Similar papers in this journal
- Selecting the most important self-assessed features for predicting conversion to Mild Cognitive Impairment with Random Forest and Permutation-based methods 94%
- Mitigating Machine Learning Bias Between High Income and Low-Middle Income Countries for Enhanced Model Fairness and Generalizability 94%
- Machine learning for classifying chronic kidney disease and predicting creatinine levels using at-home measurements 94%
Similar papers in this journal
- Characterizing subgroup performance of probabilistic phenotype algorithms within older adults: A case study for dementia, mild cognitive impairment, and Alzheimer’s and Parkinson’s diseases 97%
- Modeling physician variability to prioritize relevant medical record information 94%
- A Study of Calibration as a Measurement of Trustworthiness of Large Language Models in Biomedical Research 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.