Identification of polyketide biosynthetic gene clusters that harbor self-resistance target genes
Vandova, G. A.; Nivina, A.; Davis, R. W.; Khosla, C.; Fischer, C. R.; Hillenmeyer, M. E.
Show abstract
BackgroundPolyketide secondary metabolites have been a rich source of antibiotic discovery for decades. Thousands of novel polyketide synthase (PKS) gene clusters have been identified in recent years with advances in DNA sequencing. However, experimental characterization of novel and useful PKS activities remains complicated. As a result, computational tools to analyze sequence data are essential to identify and prioritize potentially novel PKS activities. Here we exploit the concept of genetically-encoded self-resistance to identify and rank biosynthetic gene clusters for their potential to encode novel antibiotics. ResultsTo identify PKS genes that are likely to produce an antibacterial compound, we developed an automated method to identify and catalog clusters that harbor potential self-resistance genes. We manually curated a list of known self-resistance genes and searched all NCBI genome databases for homologs of these self-resistance genes in biosynthetic gene clusters. The algorithm takes into account (1) the distance of the potential self-resistance gene to a core enzyme in the biosynthetic gene cluster; (2) the presence of a duplicated housekeeping copy of the self-resistance gene; (3) the presence of close homologs of the biosynthetic gene cluster in diverse species also harboring the putative self-resistance gene; (4) evidence for coevolution of the self-resistance gene and core biosynthetic gene; and (5) self-resistance gene ubiquity. We generated a catalog of 190 unique PKS clusters whose products likely target known enzymes of antibacterial importance. We also present an expanded set of putative self-resistance genes that may be useful in identifying small molecules active against novel microbial targets. ConclusionsWe developed a bioinformatic approach to identify and rank biosynthetic gene clusters that likely harbor self-resistance genes and may produce compounds with antibacterial properties. We compiled a list of putative self-resistance genes for novel antibacterial targets, and of orphan PKS clusters harboring these targets. These catalogues are a resource for discovery of novel antibiotics.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Biosynthetic gene cluster profiling predicts the positive association between antagonism and phylogeny in Bacillus 94%
- Convergence of resistance and evolutionary responses in Escherichia coli and Salmonella enterica co-inhabiting chicken farms in China 94%
- Construction of a complete set of Neisseria meningitidis defined mutants - the NeMeSys 2.0 collection - and its use for the phenotypic profiling of the genome of an important human pathogen 94%
Similar papers in this journal
- From sequence to molecules: Feature sequence-based genome mining uncovers the hidden diversity of bacterial siderophore pathways 94%
- Antimicrobial activity of iron-depriving pyoverdines against human opportunistic pathogens 94%
- Treatment history shapes the evolution of complex carbapenem-resistant phenotypes in Klebsiella spp. 93%
Similar papers in this journal
- RRE-Finder: A Genome-Mining Tool for Class-Independent RiPP Discovery 95%
- Elucidation of independently modulated genes in Streptococcus pyogenes reveals carbon sources that control its expression of hemolytic toxins 93%
- Pangenome analytics reveal two-component systems as conserved targets in ESKAPEE pathogens 93%
Similar papers in this journal
- Leveraging a synthetic biology approach to enhance BCG-mediated expansion of Vγ9Vδ2 T cells 91%
- Promiscuous structural cross-compatibilities between major shell components of Klebsiella pneumoniae bacterial microcompartments 91%
- A resource for improved predictions of Trypanosoma and Leishmania protein three-dimensional structure 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.