Back

Host gene-expression signatures accurately distinguish bacterial, viral, and inflammatory diseases in febrile children across multiple cohorts

Viz-Lasheras, S.; Dacosta, A.; Rivero-Calle, I.; Martinon-Torres, F.; EUCLIDS, GENDRES, PERFORM, and DIAMONDS consortia, ; Gomez-Carballa, A.; Salas, A.

2026-08-18 genomics
10.64898/2026.08.11.744042 bioRxiv
Show abstract

Accurate discrimination between viral, bacterial, and inflammatory diseases in febrile children remains a major clinical challenge that contributes to diagnostic uncertainty, inappropriate antimicrobial use, and suboptimal clinical management. Host blood transcriptomics offer a promising strategy to improve diagnostic precision. The present study represents the largest integrative multi-cohort pediatric study of transcriptomic biomarker discovery, validation, and confirmation reported to date, integrating harmonized public transcriptomic datasets with an independent confirmation cohort comprising well-phenotyped patients to identify parsimonious host-response signatures for differentiating viral, bacterial, and inflammatory diseases. Transcriptomic signatures were derived from an integrated retrospective microarray multi-cohort (n=1,683), independently validated in a retrospective RNA-seq cohort (n=767), and confirmed by digital PCR in an independent cohort (n=29), demonstrating reproducibility across patient populations, transcriptomic technologies, and analytical platforms. The analysis identified binary signatures and a unified multiclass classifier that consistently achieved high diagnostic accuracy across all three study phases and outperformed more than 30 published host transcriptomic signatures. Decision curve analysis showed substantially greater clinical net benefit than C-reactive protein across clinically relevant decision thresholds. These findings provide a strong foundation for clinically deployable molecular diagnostics to improve patient triage, antimicrobial stewardship, and precision medicine in childhood infections.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.