Standardized Multicenter Critical Care Database Integrating Minute-Level Vital Signs, Laboratory Tests, Interventions, and Outcomes: Profile of the OneICU Database
Kinoshita, T.; Umemura, Y.; Watanabe, M.; Uchida, K.; Nishimoto, Y.; Katayama, S.; Inoue, Y.; Kurosawa, H.; Nakamori, Y.
Show abstract
IntroductionStandrdized intensive care unit (ICU) databases with high frequency data collection from multicenter electronic medical records remain scarce. We aimed to describe the profile of newly developed OneICU database, which includes minute-level recordings of vital signs, laboratory values, interventions, and diagnosis codes, and to evaluate the importance of vital sign measurement frequency and laboratory data completeness for developing machine learning-based clinical decision support. MethodsThis retrospective, multicenter observational study collected critically ill patient data from 12 tertiary care hospitals from 2013 to 2025. Patient demographics, measurement frequency of vital signs and laboratory tests, as well as the availability of the Sequential Organ Failure Assessment (SOFA) score components were compared across three large ICU databases: OneICU, Medical Information Mart for Intensive Care (MIMIC)-IV and eICU. We then evaluated the prediction accuracy of five machine learning models to forecast hypotensive events 60-120 minutes in advance, defined as a median invasive mean arterial pressure (MAP) < 65 mmHg or vasopressor initiation, using different MAP sampling frequencies: OneICU 1 minute, OneICU 5 minute, OneICU hourly, MIMIC IV hourly, and eICU 5 minute. ResultsOneICU currently includes 152,269 ICU stays from 127,757 unique patients. Compared with MIMIC-IV and eICU, OneICU captured more frequent vital signs (minute-level vs. hourly in MIMIC-IV and every five minutes in eICU) and provided broader availability of SOFA score components. In particular, the respiratory component was available for 73.6 % of stays in OneICU versus 37.8 % in MIMIC IV and 30.9 % in eICU, and the liver component for 93.3 % versus 45.2 % and 41.3 %, respectively. The test-set area under the receiver operating characteristic curve was highest for the OneICU 1-minute model (0.942), followed by OneICU 5-minute model (0.939), OneICU hourly model (0.901), eICU 5-minute model (0.899), and MIMIC IV hourly model (0.799). ConclusionsA high resolution, multicenter ICU database integrating minute level vital sign recordings with comprehensive SOFA score coverage is feasible and was associated with superior hypotension prediction performance. OneICU enables detailed analyses of ICU trajectories and addresses the current scarcity of large scale ICU data from Asian populations.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 95%
- Development of a Risk Prediction Model for Sepsis-Related Delirium Based on Multiple Machine Learning Approaches and an Online Calculator 95%
- SOFA score performs worse than age for predicting mortality in patients with COVID-19 93%
Similar papers in this journal
- Identification of physiological adverse events using continuous vital signs monitoring during paediatric critical care transport: a novel data-driven approach 94%
- Generalizability Challenges of Mortality Risk Prediction Models: A Retrospective Analysis on a Multi-center Database 93%
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 93%
Similar papers in this journal
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 96%
- ARDSFlag: An NLP/Machine Learning Algorithm to Visualize and Detect High-Probability ARDS Admissions Independent of Provider Recognition and Billing Codes 95%
- Sepsis prediction via the clinical data integration system in the ICU 95%
Similar papers in this journal
- Development and Prospective Implementation of a Large Language Model based System for Early Sepsis Prediction 95%
- A comprehensive ML-based Respiratory Monitoring System for Physiological Monitoring & Resource Planning in the ICU 95%
- CT-based Rapid Triage of COVID-19 Patients: Risk Prediction and Progression Estimation of ICU Admission, Mechanical Ventilation, and Death of Hospitalized Patients 94%
Similar papers in this journal
- Predicting bloodstream infection outcome using machine learning 96%
- Imputation of PaO2 from SpO2 values from the MIMIC-III Critical Care Database Using Machine-Learning Based Algorithms 95%
- Evaluation of Domain Generalization and Adaptation on Improving Model Robustness to Temporal Dataset Shift in Clinical Medicine 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.