Back

Development and Validation of a MALDI-TOF-Based Model to Predict Extended-Spectrum Beta-Lactamase and/or Carbapenemase-Producing in Klebsiella pneumoniae Clinical Isolates

Guerrero-Lopez, A.; Candela, A.; Sevilla-Salcedo, C.; Hernandez-Garcia, M.; Martinez-Olmos, P.; Canton, R.; Munoz, P.; del Campo, R.; Gomez-Verdejo, V.; Rodriguez-Sanchez, B.

2021-10-05 microbiology
10.1101/2021.10.04.463058 bioRxiv
Show abstract

Matrix-Assisted Laser Desorption Ionization Time-Of-Flight (MALDI-TOF) Mass Spectrometry (MS) is a reference method for microbial identification and it can be used to predict Antibiotic Resistance (AR) when combined with artificial intelligence methods. However, current solutions need time-costly preprocessing steps, are difficult to reproduce due to hyperparameter tuning, are hardly interpretable, and do not pay attention to epidemiological differences inherent to data coming from different centres, which can be critical. We propose using a multi-view heterogeneous Bayesian model (KSSHIBA) for the prediction of AR using MALDI-TOF MS data together with their epidemiological differences. KSSHIBA is the first model that removes the ad-hoc preprocessing steps that work with raw MALDI-TOF data. In addition, due to its Bayesian probabilistic nature, it does not require hyperparameter tuning, provides interpretable results, and allows exploiting local epidemiological differences between data sources. To test the proposal, we used data from 402 Klebsiella pneumoniae isolates coming from two different domains and 20 different hospitals located in Spain and Portugal. KSSHIBA outperforms current state-of-the-art approaches in antibiotic susceptibility prediction, obtaining a 0.78 AUC score in Wild Type classification and a 0.90 AUC score in Extended-Spectrum Beta-Lactamases (ESBL)+Carbapenemases (CP)-producers. The proposal consistently removes the need for ad-hoc preprocessing by working with raw MALDI-TOF data, which, in turn, reduces the time needed to obtain the results of the resistance mechanism in microbiological laboratories. The proposed model implementation as well as both data domains are publicly available.

Matching journals

The top 10 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.