A Deep Learning Approach for Culture-Free Bacterial Meningitis Diagnosis and ICU Outcome Prediction
Chen, R.; Cai, Y.; Zhang, S.; Huo, Z.; Song, M.; Li, W.; Yang, D.; Zhang, X.
Show abstract
BackgroundCerebrospinal fluid (CSF) culture is the diagnostic gold standard for neuroinfectious diseases such as bacterial meningitis, but its sensitivity is limited and results are often delayed. Natural language processing (NLP) offers a powerful approach to extract meaningful clinical signals from unstructured data such as chief complaints and ICD notes. This study applies machine learning, including BioBERT-enhanced NLP models (not traditional TF-IDF approaches), to support early diagnosis and outcome prediction in ICU patients. MethodsTraining and validation datasets were derived from MIMIC-IV (internal) and MIMIC-III/eICU (external) databases. Fully connected neural network (FCNN) and other machine learning models were trained to predict CSF culture results using structured lab features. Labels were refined using clinical criteria to reduce false negatives. For ICU survival prediction, three multimodal deep learning architectures (mCNN, mFCNN, and mLSTM) were developed using two ICU survival cohorts with different inclusion criteria. The Strict ICU Survival Cohort included CSF culture results as an input feature, while the Lenient ICU Survival Cohort excluded this requirement, allowing for a broader patient population. In both cohorts, models integrated structured variables with unstructured text encoded by BioBERT, a deep contextual language model, rather than simpler methods like TF-IDF, effectively capturing clinical meaning from free-text ICD entries and chief complaints. ResultsFor CSF culture prediction (training n = 9261), the FCNN model achieved the highest performance (AUROC = 0.853) in independent validation. For ICU survival prediction in the Strict ICU Survival Cohort (training n = 5,795), the mCNN model achieved an AUROC of 0.889 in external validation. In the expanded Lenient ICU Survival Cohort (n = 58,615), the same model achieved an AUROC of 0.974 and an AUPRC of 0.868 during external validation. During model training and development, the predictive performance declined when text features were excluded (AUROC from 0983 to 0.946) or when ICD entries were converted from free-text (BioBERT-encoded) to coded format (AUROC to 0.947). ConclusionsMultimodal machine learning models, enhanced by advanced NLP through BioBERT embeddings of clinical free text, effectively predicted CSF culture results and ICU survival outcomes.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development and Prospective Implementation of a Large Language Model based System for Early Sepsis Prediction 95%
- CT-based Rapid Triage of COVID-19 Patients: Risk Prediction and Progression Estimation of ICU Admission, Mechanical Ventilation, and Death of Hospitalized Patients 94%
- A comprehensive ML-based Respiratory Monitoring System for Physiological Monitoring & Resource Planning in the ICU 93%
Similar papers in this journal
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 94%
- Modular Clinical Decision Support Networks (MoDN)—Updatable, Interpretable, and Portable Predictions for Evolving Clinical Environments 92%
- Community-acquired pneumonia identification from electronic health records in the absence of a gold standard: a Bayesian latent class analysis 92%
Similar papers in this journal
- AI-MET: A Deep Learning-based Clinical Decision Support System for Distinguishing Multisystem Inflammatory Syndrome in Children from Endemic Typhus 95%
- Machine Learning Interpretability Methods to Characterize the Importance of Hematologic Biomarkers in Prognosticating Patients with Suspected Infection 94%
- Whole slide image representation in bone marrow cytology 93%
Similar papers in this journal
- A comparison of machine learning models versus clinical evaluation for mortality prediction in patients with sepsis 95%
- Development of a Risk Prediction Model for Sepsis-Related Delirium Based on Multiple Machine Learning Approaches and an Online Calculator 94%
- Machine learning in predicting respiratory failure in patients with COVID-19 pneumonia - challenges, strengths, and opportunities in a global health emergency 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.