Pannagram: unbiased pangenome alignment and the Mobilome calling
Igolkina, A. A.; Bezlepsky, A. D.; Nordborg, M.
Show abstract
Pannagram is a toolkit for unbiased pangenome alignment and Mobilome detection based on full-genome assemblies. It consists of two components: a command-line interface for performing the multiple genome alignment in different modes, extracting features such as structural variants (SVs), and searching sequences between datasets; and an R library that provides tools for analyzing alignments and SVs, as well as visualizing every step of the Mobilome analysis pipeline. As a proof of concept, we applied Pannagram to 12 Cucumis sativus genomes, identified SVs, and constructed the graph of nestedness for these variants. Within this graph, we identified families of actually mobile elements belonging to known transposable element families (LINEs, LTRs, and TIRs) and uncovered candidates for a novel type of mobile elements. Our results highlight the power of Pannagram in Mobilome discovery while ensuring an unbiased approach to pangenome analysis.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- HiC-TE: a computational pipeline for Hi-C data analysis shows a possible role of repeat family interactions in the genome 3D organization 95%
- HiTea: a computational pipeline to identify non-reference transposable element insertions in Hi-C data 94%
- The String Decomposition Problem and its Applications to Centromere Assembly 93%
Similar papers in this journal
Similar papers in this journal
- High-fidelity (repeat) consensus sequences from short reads using combined read clustering and assembly 94%
- LtrDetector: A modern tool-suite for detecting long terminal repeat retrotransposons de-novo on the genomic scale 94%
- Towards a better understanding of the low recall of insertion variants with short-read based variant callers 94%
Similar papers in this journal
- Benchmarking Transposable Element Annotation Methods for Creation of a Streamlined, Comprehensive Pipeline 96%
- SyRI: finding genomic rearrangements and local sequence differences from whole-genome assemblies 94%
- DIVE: a reference-free statistical approach to diversity-generating and mobile genetic element discovery 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.