Identifying Core Operons in Metagenomic Data
Hu, X.; Friedberg, I.
Show abstract
An operon is a functional unit of DNA whose genes are co-transcribed on polycistronic mRNA, in a co-regulated fashion. Operons are a powerful mechanism of introducing functional complexity in bacteria, and are therefore of interest in microbial genetics, physiology, biochemistry, and evolution. Here we present a Pipeline for Operon Exploration in Metagenomes or POEM. At the heart of POEM lies the concept of a core operon, a functional unit enabled by a predicted operon in a metagenome. Using a series of benchmarks, we show the high accuracy of POEM, and demonstrate its use on a human gut metagenome sample. We conclude that POEM is a useful tool for analyzing metagenomes beyond the genomic level, and for identifying multi-gene functionalities and possible neofunctionalization in metagenomes. Availability: https://github.com/Rinoahu/POEM_py3k
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Pre- and post-sequencing recommendations for functional annotation of human fecal metagenomes 96%
- GenAPI: a tool for gene absence-presence identification in fragmented bacterial genome sequences 96%
- Read-SpaM: assembly-free and alignment-free comparison of bacterial genomes with low sequencing coverage 95%
Similar papers in this journal
- HiCBin: Binning metagenomic contigs and recovering metagenome-assembled genomes using Hi-C contact maps 97%
- Efficient inference of large pangenomes with PanTA 97%
- MetaBinner: a high-performance and stand-alone ensemble binning method to recover individual genomes from complex microbial communities 96%
Similar papers in this journal
- AMRomics: a scalable workflow to analyze large microbial genome collection 96%
- Inferring directional relationships in microbial communities using signed Bayesian networks 95%
- MeShClust v3.0: High-quality clustering of DNA sequences using the mean shift algorithm and alignment-free identity scores 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.