Back

Gene exchange networks define species-like units in marine prokaryotes

Stepanauskas, R.; Brown, J. M.; Mai, U.; Bezuidt, O.; Pachiadaki, M.; Brown, J.; Biller, S.; Berube, P. M.; Record, N. R.; Mirarab, S.

2020-09-10 microbiology
10.1101/2020.09.10.291518 bioRxiv
Show abstract

Post-submission note. Since the original submission of this manuscript to bioRxiv, we discovered that some of our results may be impacted by the limitations of some of the comparative genomics tools used in this study. We are working on a revised version of the manuscript. Although horizontal gene transfer is recognized as a major evolutionary process in Bacteria and Archaea, its general patterns remain elusive, due to difficulties tracking genes at relevant resolution and scale within complex microbiomes. To circumvent these challenges, we analyzed a randomized sample of >12,000 genomes of individual cells of Bacteria and Archaea in the tropical and subtropical ocean - a well-mixed, global environment. We found that marine microorganisms form gene exchange networks (GENs) within which transfers of both flexible and core genes are frequent, including the rRNA operon that is commonly used as a conservative taxonomic marker. The data revealed efficient gene exchange among genomes with <28% nucleotide difference, indicating that GENs are much broader lineages than the nominal microbial species, which are currently delineated at 4-6% nucleotide difference. The 42 largest GENs accounted for 90% of cells in the tropical ocean microbiome. Frequent gene exchange within GENs helps explain how marine microorganisms maintain millions of rare genes and adapt to a dynamic environment despite extreme genome streamlining of their individual cells. Our study suggests that sharing of pangenomes through horizontal gene transfer is a defining feature of fundamental evolutionary units in marine planktonic microorganisms and, potentially, other microbiomes.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.