Back

Genome mining unveils a class of ribosomal peptides with two amino termini

Ren, H.; Dommaraju, S. R.; Huang, C.; Cui, H.; Pan, Y.; Nesic, M.; Zhu, L.; Sarlah, D.; Mitchell, D. A.; Zhao, H.

2023-03-08 biochemistry
10.1101/2023.03.08.531785 bioRxiv
Show abstract

The era of inexpensive genome sequencing and improved bioinformatics tools has reenergized the study of natural products, including the ribosomally synthesized and post-translationally modified peptides (RiPPs). In recent years, RiPP discovery has challenged preconceptions about the scope of post-translational modification chemistry, but genome mining of new RiPP classes remains an unsolved challenge. Here, we report a RiPP class defined by an unusual (S)-N2,N2-dimethyl-1,2-propanediamine (Dmp)-modified C-terminus, which we term the daptides. Nearly 500 daptide biosynthetic gene clusters (BGCs) were identified by analyzing the RiPP Recognition Element (RRE), a common substrate-binding domain found in half of prokaryotic RiPP classes. A representative daptide BGC from Microbacterium paraoxydans DSM 15019 was selected for experimental characterization. Derived from a C-terminal threonine residue, the class-defining Dmp is installed over three steps by an oxidative decarboxylase, aminotransferase, and methyltransferase. Daptides uniquely harbor two positively charged termini, and thus we suspect this modification could aid in membrane targeting, as corroborated by hemolysis assays. Our studies further show that the oxidative decarboxylation step requires a functionally unannotated accessory protein. Fused to the C-terminus of the accessory protein is an RRE domain, which delivers the unmodified substrate peptide to the oxidative decarboxylase. This discovery of a class-defining post-translational modification in RiPPs may serve as a prototype for unveiling additional RiPP classes through genome mining.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.