Towards An Ai-Driven Registry For Postoperative Complica-Tions: A Proof-Of-Concept Study Evaluating The Opportunities And Challenges Of Ai-Models
Dencker, E. E.; Millarch, A. S.; Bonde, A.; Troelsen, A.; Jensen, J. W.; Sillesen, M.
Show abstract
BackgroundContinuous quality improvement is essential in surgery, with clinical registries and quality improvement programs (QIPs) playing a key role. Postoperative complications (PCs) require substantial resources to manage, yet traditional QIPs are expensive and often lays a significant labor burden on clinicians in data collection. Artificial intelligence (AI), particularly natural language processing (NLP), offers a potential solution by automating and streamlining these processes, but models can be optimized for optimal sensitivity or positive predictive value. This study aimed to develop a mock-up automated registry for PCs using NLP algorithms and evaluate the effects of optimization strategies for surgical quality control. We hypothesized using NLP to obtain longitudinal overviews of key quality metrics is feasible, but that optimization strategies impacted on the observed rate of PCs and thus how quality management and surveillance would be affected in a real-world setting. MethodsWe analyzed 100,505 surgical cases from 12 Danish hospitals between 2016 and 2022. Previously validated NLP models were applied to detect seven types of PCs, using two different threshold settings: a set of thresholds optimized for positive predictive value (PPV or Precision), referred to as F-score of 0.5, and a set of thresholds optimized for sensitivity, referred to as F-score of 2. Trends in PC rates over time were assessed, and hospital-level variations were examined using logistic regression models adjusted for age, sex, and comorbidity. ResultsThe NLP models detected 8,512 or 15,892 PCs, depending on threshold selection, corresponding to total PC rates of 9.14% and 17.1%, respectively. Most PCs showed stable or increasing trends over time, regardless of threshold setting. Hospital-level analyses similarly revealed stable or rising PC rates in most institutions. Regression analyses demonstrated that threshold selection significantly influenced findings, impacting hospital comparisons. ConclusionThis study demonstrates that NLP can be used for automated PC detection in surgical quality monitoring. However, threshold selection and additional performance metrics, such as precision-recall curves (PPV-Sensitivity curves), must be carefully considered to ensure reliable and meaningful results beyond traditional Receiver Operator Area Under the Curve (ROC AUC) evaluation.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Unplanned Hospital Visits after Ambulatory Surgical Care 94%
- Regional performance variation in external validation of four prediction models for severity of COVID-19 at hospital admission: An observational multi-centre cohort study 94%
- Postoperative mortality analysis of national Japanese Diagnosis Procedure Combination database with a focus on regional comparisons and changes over time 94%
Similar papers in this journal
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 96%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 95%
- Development of a data-driven COVID-19 prognostication tool to inform triage and step-down care for hospitalised patients in Hong Kong: A population based cohort study 93%
Similar papers in this journal
- Development and validation of automated computer aided-risk score for predicting the risk of in-hospital mortality using first electronically recorded blood test results and vital signs for COVID-19 hospital admissions: a retrospective development and validation study 95%
- Surgery & COVID-19: A rapid scoping review of the impact of COVID-19 on surgical services during public health emergencies 94%
- Use of the first National Early Warning Score recorded within 24 hours of admission to estimate the risk of in-hospital mortality in unplanned COVID-19 patients: a retrospective cohort study 94%
Similar papers in this journal
- Impact of the Federated Data Platform's digital surgery scheduling system on elective theatre utilisation at an NHS Trust: an interrupted time series analysis 93%
- Development of a customised data management system for a COVID-19-adapted colorectal cancer pathway 93%
- The performance of national COVID-19 ‘Symptom Checkers’: A comparative case simulation study 93%
Similar papers in this journal
- The impact of atypical intrahospital transfers on patient outcomes: a mixed methods study 94%
- AKI Risk Score (AKI-RiSc): Developing an Interpretable Clinical Score for Early Identification of Acute Kidney Injury for Patients Presenting to the Emergency Department 93%
- An Observational Study of COVID-19 from A Large Healthcare System in Northern New Jersey: Diagnosis, Clinical Characteristics, and Outcomes 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.