Back

A Systematic Fairness Evaluation of Racial Bias in Alzheimers Disease Diagnosis Using Machine Learning Models

Baddam, N. G.; Pijani, B. A.; Bozdag, S.

2025-10-02 bioinformatics
10.1101/2025.09.30.678854 bioRxiv
Show abstract

INTRODUCTIONAlzheimers disease (AD) is a major global health concern, expected to affect 12.7 million Americans by 2050. Machine learning (ML) algorithms have been developed for AD diagnosis and progression prediction, but the lack of racial diversity in clinical datasets raises concerns about their generalizability across demographic groups, particularly underrepresented populations. Studies show ML algorithms can inherit biases from data, leading to biased AD predictions. METHODSThis study investigates the fairness of ML models in AD diagnosis. We hypothesize that models trained on a single racial group perform well within that group but poorly in others. We employ feature selection and model training techniques to improve fairness. RESULTSOur findings support our hypothesis that ML models trained on one group underperform on others. We also demonstrated that applying fairness techniques to ML models reduces their bias. DISCUSSIONThis study highlights the need for racial diversity in datasets and fair models for AD prediction.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.