Genus-level transfer learning of Matrix-assisted laser desorption/ionization time-of-flight mass spectrometry data predicts antibiotic resistance with greater accuracy
Chen, M.-I.; Chen, Y.-H.; Hsu, Y.-L.; Shih, Y.-T.
Show abstract
Bacterial resistance, driven by excessive antibiotic use, has rendered many traditional antibiotics ineffective. Despite the advantages of applying matrix-assisted laser desorption/ionization time-of-flight mass spectrometry (MALDI-TOF MS) to predict bacterial antimicrobial resistance, limited databases and a lack of high-quality data hinder this effort. This study aimed to address this lack of data by integrating MALDI-TOF MS technology with transfer learning approaches, utilizing genus-level data, to predict species-specific bacterial antimicrobial resistance. The data were retrieved from the DRIAMS dataset. A multilayer perceptron deep neural network was applied for pretraining and fine-tuning, integrating genus-level data to enhance the accuracy of the species-specific predictions. Notably, fine-tuned models enhanced antibiotic resistance prediction. In Staphylococcus aureus, the area under the receiver operating characteristic curve (AUROC) improved from 0.95 to 0.96 and the area under the precision-recall curve (AUPRC) improved from 0.85 to 0.88 for oxacillin; for fusidic acid, the AUROC improved from 0.80 to 0.81 and the AUPRC from 0.32 to 0.36; and for ciprofloxacin, the AUROC improved from 0.80 to 0.82 while the AUPRC remained 0.57. In Klebsiella pneumoniae, the AUROC held steady at 0.75, while the AUPRC declined from 0.45 to 0.41 for ciprofloxacin, indicating potential limitations for certain pathogen-antibiotic combinations. Despite such variability, the overall improvements across multiple strains highlight the potential of transfer learning in antimicrobial resistance prediction and underscore the need to expand genus-level databases to improve species-level diagnostic accuracy, thereby enhancing clinical diagnostics.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 93%
- Uncovering the effects of model initialization on deep model generalization: A study with adult and pediatric chest X-ray images 91%
- A flexible framework for minimal biomarker signature discovery from clinical omics studies without library size normalisation 91%
Similar papers in this journal
- Using Deep Learning for Gene Detection and Classification in Raw Nanopore Signals 93%
- Low-Cost In-House Re-formulated Brain Heart Infusion Medium for Effective Planktonic Growth and Early Detection of Bloodstream Bacterial Pathogens 92%
- Pixel-based machine learning and image reconstitution for dot-ELISA pathogen serodiagnosis 92%
Similar papers in this journal
- MACI: A machine learning-based approach to identify drug classes of antibiotic resistance genes from metagenomic data 96%
- ToxinPred 3.0: An improved method for predicting the toxicity of peptides 94%
- Machine Learning Interpretability Methods to Characterize the Importance of Hematologic Biomarkers in Prognosticating Patients with Suspected Infection 93%
Similar papers in this journal
- Comparative Performance Of Two Automated Machine Learning Platforms For COVID-19 Detection By Maldi-Tof-Ms 95%
- Interpretable machine learning with tree-based Shapley additive explanations: application to metabolomics datasets for binary classification 94%
- Improving prediction of drug-target interactions based on fusing multiple features with data balancing and feature selection techniques 94%
Similar papers in this journal
- Designing of thermostable proteins with a desired melting temperature 94%
- DeepInsight-3D for precision oncology: an improved anti-cancer drug response prediction from high-dimensional multi-omics data with convolutional neural networks 93%
- ProtAlign-ARG: Antibiotic Resistance Gene Characterization Integrating Protein Language Models and Alignment-Based Scoring 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.