Vast human gut virus diversity uncovered by combined short- and long-read sequencing
Chen, W.; Chen, J.; Sun, C.; Dong, Y.; Jin, M.; Lai, S.; Jia, L.; Zhao, X.; Gao, N. L.; Liu, Z.; Bork, P.; Zhao, X.-M.
Show abstract
Current metagenome-assembled human phage catalogs contained mostly fragmented genomes. Here, we developed a vigorous phage detection method involving phage enrichment and long-read sequencing and applied to 135 fecal samples. With ~10 times more efficient in obtaining complete genomes (~34%) than the Gut Virome Database, we identified the first megabasephage (~1.03Mb), and revealed the hidden diversity of the gut phageome including dozens of phages more prevalent than the crAssphages and Gubaphages.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- The genetic diversity and populational specificity of the human gut virome at single nucleotide resolution 98%
- Viromes vs. mixed community metagenomes: choice of method dictates interpretation of viral community ecology 94%
- Strain-resolved de-novo metagenomic assembly of viral genomes and microbial 16S rRNAs 93%
Similar papers in this journal
- Long-read sequencing reveals extensive DNA methylations in human gut phagenome contributed by prevalently phage-encoded methyltransferases 97%
- Microbial general model: Leveraging large language model for contextualized microbiome analysis 92%
- Spatiotemporal transcriptomic profiling reveals the dynamic immunological landscape of alveolar echinococcosis 90%
Similar papers in this journal
- Genome-wide association study between SARS-CoV-2 single nucleotide polymorphisms and virus copies during infections 95%
- Virus-Host Interactions Predictor (VHIP): machine learning approach to resolve microbial virus-host interaction networks 94%
- De novo virus inference and host prediction from metagenome using CRISPR spacers 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.