Integrating a host transcriptomic biomarker with a large language model for diagnosis of lower respiratory tract infection
Phan, H. V.; Spottiswoode, N.; Lydon, E. C.; Chu, V. T.; Cuesta, A.; Kazberouk, A. D.; Richmond, N. L.; Calfee, C. S.; Langelier, C.
Show abstract
BACKGROUNDLower respiratory tract infections (LRTIs) are a leading cause of mortality worldwide and can be difficult to diagnose in critically ill patients, as non-infectious causes of respiratory failure can present with similar clinical features. METHODSWe developed a LRTI diagnostic method combining the pulmonary transcriptomic biomarker FABP4 with electronic medical record (EMR) text assessment using the large language model Generative Pre-trained Transformer 4 (GPT-4). We evaluated this approach in a prospective cohort of critically ill adults with acute respiratory failure from whom tracheal aspirate FABP4 expression was measured by RNA sequencing. Patients with LRTI or non-infectious conditions were identified using retrospective, multi-physician clinical adjudication. We then confirmed our findings by applying this method to an independent validation cohort of 115 adults with acute respiratory failure. RESULTSIn the derivation cohort, a combined classifier incorporating FABP4 expression and GPT-4- assisted EMR analysis achieved an AUC of 0.93 ({+/-}0.08) and an accuracy of 84%, outperforming FABP4 expression alone (AUC 0.84 {+/-} 0.11) and GPT-4-based analysis alone (AUC 0.83 {+/-} 0.07). By comparison, the primary medical teams admission diagnosis had an accuracy of 72%. In the validation cohort, the combined classifier yielded an AUC of 0.98 ({+/-}0.04) and an accuracy of 96%. CONCLUSIONSIntegrating a host transcriptional biomarker with EMR text analysis using a large language model may offer a promising new approach to improving the diagnosis of LRTIs in critically ill adults. DescriptionWe present the novel use of a host transcriptional biomarker combined with artificial intelligence analysis of electronic medical record data to diagnose lower respiratory tract infections in a derivation cohort of critically ill adults, then the validation of this approach in a second, fully independent, cohort. This approach demonstrated high diagnostic accuracy compared to a gold standard of post-hoc multi-physician adjudication.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Transformer-based deep learning model for the diagnosis of suspected lung cancer in primary care based on electronic health record data 92%
- Characterizing Long COVID: Deep Phenotype of a Complex Condition 90%
- The COVID-19 Pandemic as an Opportunity for Unravelling the Causative Association between Respiratory Viruses and Pneumococcus-Associated Disease in Young Children: A Prospective Study 90%
Similar papers in this journal
- Biomarkers to distinguish bacterial from viral pediatric clinical pneumonia in a malaria endemic setting 94%
- SARS-CoV-2 RNAaemia predicts clinical deterioration and extrapulmonary complications from COVID-19 93%
- Multimorbidity Profiles and Severe In-Hospital Outcomes in Adults with Respiratory Syncytial Virus 93%
Similar papers in this journal
- Pulmonary function and survival one year after dupilumab treatment of acute moderate to severe COVID-19: A follow up study from a Phase IIa trial 92%
- A Predictive Model to Identify Complicated Clostridiodes difficile Infection 91%
- Discriminatory ability of gas chromatography-ion mobility spectrometry to identify patients hospitalised with COVID-19 and predict prognosis 91%
Similar papers in this journal
- Evaluation of Domain Generalization and Adaptation on Improving Model Robustness to Temporal Dataset Shift in Clinical Medicine 93%
- Predicting bloodstream infection outcome using machine learning 93%
- Imputation of PaO2 from SpO2 values from the MIMIC-III Critical Care Database Using Machine-Learning Based Algorithms 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.