Back

Machine Learning of Pseudomonas aeruginosa transcriptomes identifies independently modulated sets of genes associated with known transcriptional regulators

Rajput, A.; Tsunemoto, H.; Sastry, A. V.; Szubin, R.; Rychel, K.; Sugie, J.; Pogliano, J.; Palsson, B.

2021-07-28 bioinformatics
10.1101/2021.07.28.454220 bioRxiv
Show abstract

The transcriptional regulatory network (TRN) of Pseudomonas aeruginosa plays a critical role in coordinating numerous cellular processes. We extracted and quality controlled all publicly available RNA-sequencing datasets for P. aeruginosa to find 281 high-quality transcriptomes. We produced 83 new RNAseq data sets under critical conditions to generate a comprehensive compendium of 364 transcriptomes. We used this compendium to reconstruct the TRN of P. aeruginosa using independent component analysis (ICA). We identified 104 independently modulated sets of genes (called iModulons), among which 81 (78%) reflect the effects of known transcriptional regulators. We show that iModulons: 1) play an important role in defining the genomic boundaries of biosynthetic gene clusters (BGCs); 2) show increased expression of the BGCs and associated secretion systems in conditions that emulate cystic fibrosis (CF); 3) show the presence of a novel BGC named RiPP (bacteriocin producer) which might have a role in worsening CF outcomes; 4) exhibit the interplay of amino acid metabolism regulation and central metabolism across carbon sources, and 5) clustered according to their activity changes to define iron and sulfur stimulons. Finally, we compare the iModulons of P. aeruginosa with those of E. coli to observe conserved regulons across two gram negative species. This comprehensive TRN framework covers almost every aspect of the transcriptional regulatory machinery in P. aeruginosa, and thus could prove foundational for future research of its physiological functions.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.