Plugging the holes in the Swiss Cheese by learning from those who fall through the cracks of our medical system's multiple lines of defense in Hong Kong: Data-based analytic protocol for identifying the tertiary prevention needs of a population with machine learning pipeline
Leung, E.; Lee, A.; Guan, J.; Ching, C. C.; Lam, O.; Wong, B.; Law, C. B.; Tsang, H. W.
Show abstract
IntroductionThe objective of tertiary prevention is to reduce re-hospitalization, as re-hospitalization puts patients at unnecessary risk, delays those in the population requiring timely care, and incurs financial burdens on healthcare systems. Nevertheless, it is challenging to stratify the needs of tertiary prevention in a population. Hence, we advance an analytic protocol to identify from an inpatient population the clinical and service-utilization profiles of those re-hospitalized within 28 days. Methods and analysisThe protocol is based on implementing unsupervised and supervised machine learning (ML) in tandem with an inpatient populations electronic health records. The unsupervised ML will cluster the population into segments of maximized within-segment similarity and between-segment dissimilarity, across the dimensions of clinical diagnoses, acuity, complexity, chronicity, and multimorbidity. Within each clinically similar segment identified, a 28-day re-hospitalization outcome-supervised decision tree will classify the segment into a series of binarily-split subgroups with profile and service utilization-related features. The order of selected features reflects relative importance to the outcome. Two subgroups originated from a selected feature are statistically different in re-hospitalization outcomes. So, the subgroups lacking selected services while realizing the highest re-hospitalization rates will potentially benefit most from tertiary prevention of the selected services. Thus, they are fit for further assessment and corresponding interventions. Ethics and disseminationThe Survey and Behaviour Research Ethics Committee of the Chinese University of Hong Kong, Hospital Authority Data Collaboration Lab, and a local ethics committee have approved this protocol and its ongoing and forthcoming validations in territory-wide HK through centralized data access and in local clinical management systems, respectively. In addition to disseminating through publications, presentations, and other communications, the protocol is also implementable in different systems as part of the decision support mechanism to inform the venue-based sampling of patients with a high risk of re-hospitalization. Article SummaryO_ST_ABSStrengths and Limitations of this studyC_ST_ABSStrength #1This will be a data-based machine-learning analytic methodology to align population-based cohort study with the conceptual framework of the Swiss Cheese Model for patient safety. The Swiss Cheese Model is a conceptual framework that elevates us from the paradigm of linear causality to conceptualizing safety events as having multiple lines of defense simultaneously broken through. However, the predominant linear model research is incompatible with identifying a medical systems multiple lines of defense and their potential breakpoints, due to much less prioritizing impacts and mapping interactions on outcomes. Strength #2This protocol no long assumes independence among clinical profiles and service utilization-related factors. As a result, the absence and ensembles of post-acute services of the entire population can be handled appropriately. Strength #3Subgroups of statistically significant differences in re-hospitalization outcomes will be identified and articulable as a portfolio of clinical profiles and selected services. LimitationsHeterogeneity may still exist after the clustering and classification based on clinical homogeneity. Like all studies based on clinical databases, the heterogeneity can be attributable to exogenous factors absent in clinical information systems, e.g., psychosocial factors, and the narrowly defined tertiary prevention needs of the supervisory outcome.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Demonstrating the Consequences of Learning Missingness Patterns in Early Warning Systems for Preventative Health Care: A Novel Simulation and Solution 94%
- Signal from the Noise: A Mixed Methods Process Mining Approach to Evaluate Care Pathways. 94%
- Mining for Equitable Health: Assessing the Impact of Missing Data in Electronic Health Records 93%
Similar papers in this journal
- Implicit bias in Critical Care Data: Factors affecting sampling frequencies and missingness patterns of clinical and biological variables in ICU Patients 95%
- Development of a data-driven COVID-19 prognostication tool to inform triage and step-down care for hospitalised patients in Hong Kong: A population based cohort study 94%
- A Multi-Granular Stacked Regression for Forecasting Long-Term Demand in Emergency Departments 94%
Similar papers in this journal
- Development and validation of a machine learning model for predicting illness trajectory and hospital resource utilization of COVID-19 hospitalized patients - a nationwide study 94%
- Measure what matters: counts of hospitalized patients are a better metric for health system capacity planning for a reopening 94%
- Use of unstructured text in prognostic clinical prediction models: a systematic review 93%
Similar papers in this journal
- Synthetic Data Generation in Healthcare: A Scoping Review of reviews on domains, motivations, and future applications 93%
- Development and Evaluation of MADDIE: Method to Acquire Delivery Date Information from Electronic Health Records 93%
- Completion of electronic nursing documentation of inpatient admission assessment: insights from Australian metropolitan hospitals 92%
Similar papers in this journal
- Nowcasting and Forecasting the Spread of COVID-19 and Healthcare Demand In Turkey, A Modelling Study 94%
- A Data-Driven Framework for Identifying Intensive Care Unit Admissions Colonized with Multidrug-Resistant Organisms 93%
- Predicting hospital demand during the COVID-19 outbreak in Bogota, Colombia 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.