Back

Resistance Gene Association and Inference Network (ReGAIN): A Bioinformatics Pipeline for Assessing Probabilistic Co-Occurrence Between Resistance Genes in Bacterial Pathogens

Bring Horvath, E. R.; Stein, M. G.; Mulvey, M. A.; Hernandez, E. J.; Winter, J. M.

2024-03-01 bioinformatics
10.1101/2024.02.26.582197 bioRxiv
Show abstract

The rampant rise of multidrug resistant (MDR) bacterial pathogens poses a severe health threat, necessitating innovative tools to unravel the complex genetic underpinnings of antimicrobial resistance. Despite significant strides in developing genomic tools for detecting resistance genes, a gap remains in analyzing organism-specific patterns of resistance gene co-occurrence. Addressing this deficiency, we developed the Resistance Gene Association and Inference Network (ReGAIN), a novel web-based and command line genomic platform that uses Bayesian network structure learning to identify and map resistance gene networks in bacterial pathogens. ReGAIN not only detects resistance genes using well- established methods, but also elucidates their complex interplay, critical for understanding MDR phenotypes. Focusing on ESKAPE pathogens, ReGAIN yielded a queryable database for investigating resistance gene co-occurrence, enriching resistome analyses, and providing new insights into the dynamics of antimicrobial resistance. Furthermore, the versatility of ReGAIN extends beyond antibiotic resistance genes to include assessment of co-occurrence patterns among heavy metal resistance and virulence determinants, providing a comprehensive overview of key gene relationships impacting both disease progression and treatment outcomes. Graphical Abstract O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=190 SRC="FIGDIR/small/582197v1_ufig1.gif" ALT="Figure 1"> View larger version (53K): org.highwire.dtl.DTLVardef@158a667org.highwire.dtl.DTLVardef@114c965org.highwire.dtl.DTLVardef@1b24504org.highwire.dtl.DTLVardef@d112af_HPS_FORMAT_FIGEXP M_FIG C_FIG

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.