Prediction of Hospital Outpatient Attendance in UK Hospitals: A Retrospective Study Applying Machine Learning to Routinely Collected Data for Patients of All Ages
Holdship, J.; Dhanoa, H.; Hopper, A.; Steves, C. J.; Butler, M.; Wolfe, I.; Tucker, K.; Cooper, C.; Yates, J.
Show abstract
ObjectivesPatient non-attendance at outpatient appointments is a major concern for healthcare providers. Non-attendances increase waiting lists, reduce access to care and may be detrimental not for the patient who did not attend. We aim to produce a model which can accurately predict which appointments will be attended. SettingA teaching hospital in London, UK combining secondary and tertiary care. ParticipantsA set of 9.6 million outpatient appointments between April 2015 and September 2019 including all ages and specialities. Primary and secondary outcome measuresArea under the receiver operating characteristic curve (AU-ROC) for prediction of outpatient appointment non-attendances. ResultsThe model uses 27 predictors to achieve an AUROC score of 0.768 (95% CI: 0.767-0.769) and accuracy of 89.2% (95% CI: 89.16%-89.24%) on test data. We find that the waiting period between booking and the appointment, the patients past attendance behaviour, and the levels of deprivation in their local area are important factors in predicting future attendance. ConclusionOur model successfully predicts patient attendance at outpatient appointments. Its performance on both patients who did not appear in the training data and appointments from a different time period which covers the Covid-19 pandemic indicate it generalized well across both face to face and virtual appointments and could be used to target resources and intervention towards those patients who are likely to miss an appointment. Moreover, it highlights the impact of deprivation on patient access to healthcare Strengths and Limitation of this StudyO_LIWe make use of a large dataset which enables us to use complex machine learning algorithms. C_LIO_LIWe validate the model on two large, distinct datasets giving high confidence in our model performance. C_LIO_LIAn unknown amount of patient data is missing due to a nearby hospital which shares patients with the study setting. C_LI
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 96%
- A Multi-Granular Stacked Regression for Forecasting Long-Term Demand in Emergency Departments 95%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 94%
Similar papers in this journal
- From theoretical models to practical deployment: A perspective and case study of opportunities and challenges in AI-driven healthcare research for low-income settings 95%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Modular Clinical Decision Support Networks (MoDN)—Updatable, Interpretable, and Portable Predictions for Evolving Clinical Environments 93%
Similar papers in this journal
- Mitigating Machine Learning Bias Between High Income and Low-Middle Income Countries for Enhanced Model Fairness and Generalizability 96%
- Emergency department admissions during COVID-19: explainable machine learning to characterise data drift and detect emergent health risks 96%
- Using explainable machine learning to identify patients at risk of reattendance at discharge from emergency departments 95%
Similar papers in this journal
- A Deep Learning Method to Detect Opioid Prescription and Opioid Use Disorder from Electronic Health Records 94%
- Assessing the effects of data drift on the performance of machine learning models used in clinical sepsis prediction 94%
- Synthetic Data Generation in Healthcare: A Scoping Review of reviews on domains, motivations, and future applications 93%
Similar papers in this journal
- Modeling physician variability to prioritize relevant medical record information 95%
- Characterizing subgroup performance of probabilistic phenotype algorithms within older adults: A case study for dementia, mild cognitive impairment, and Alzheimer’s and Parkinson’s diseases 94%
- Trajectories: a framework for detecting temporal clinical event sequences from health data standardized to the OMOP Common Data Model 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.