Assessment of machine learning algorithms to predict medical specialty choice
Alvarez, D. V.; Abbiati, M.; Bornet, A.; Savoldelli, G.; Bajwa, N.; Teodoro, D.
Show abstract
Equitable distribution of physicians across specialties is a significant public health challenge. While previous studies primarily relied on classic statistics models to estimate factors affecting medical students career choices, this study explores the use of machine learning techniques to predict decisions early in their studies. We evaluated various supervised models, including support vector machines, artificial neural networks, extreme gradient boosting (XGBoost), and CatBoost using data from 399 medical students from medical faculties in Switzerland and France. Ensemble methods outperformed simpler models, with CatBoost achieving a macro AUROC of 76%. Post-hoc interpretability methods revealed key factors influencing predictions, such as motivation to become a surgeon and psychological traits like extraversion. These findings show that machine learning could be used for predicting medical career paths and inform better workforce planning.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 93%
- An AI-based approach to predict delivery outcome based on measurable factors of pregnant mothers 93%
- Cardiology Knowledge Assessment of Retrieval-Augmented Open versus Proprietary Large Language Models 91%
Similar papers in this journal
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 96%
- Prediction of Sepsis Mortality in ICU Patients Using Machine Learning Methods 95%
- Combining symbolic regression with the Cox proportional hazards model improves prediction of heart failure deaths 94%
Similar papers in this journal
- A Machine Learning-Based Prediction of Hospital Mortality in Mechanically Ventilated ICU Patients 94%
- Factors and mediators impacting the number of undergraduate research mentees at a research-intensive Hispanic-serving institution 94%
- ChatGPT-Enhanced ROC Analysis (CERA): A Shiny Web Tool for Finding Optimal Cutoff in Biomarker Analysis 93%
Similar papers in this journal
- Evaluation of the performance of GPT-3.5 and GPT-4 on the Medical Final Examination 94%
- A multipurpose machine learning approach to predict COVID-19 negative prognosis in Sao Paulo, Brazil 93%
- Machine learning for classifying chronic kidney disease and predicting creatinine levels using at-home measurements 93%
Similar papers in this journal
- Predicting mortality in SARS-COV-2 (COVID-19) positive patients in the inpatient setting using a Novel Deep Neural Network 93%
- Performance of Advanced Large Language Models (GPT-4o, GPT-4, Gemini 1.5 Pro, Claude 3 Opus) on Japanese Medical Licensing Examination: A Comparative Study 92%
- Image and structured data analysis for prognostication of health outcomes in patients presenting to the Emergency Department during the COVID-19 pandemic 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.