HAMRLNC: A Comprehensive Pipeline for High-throughput Analysis of Modified Ribonucleotides and Long Non-Coding Ribonucleic Acids
Obih, C. E.; Li, J.; melandri, G.; Pauli, D.; Lyons, E.; Nelson, A. D. L.; Gregory, B. D.
Show abstract
As sequencing technologies advance and costs decline, there has been a surge in the application of RNA sequencing (RNA-seq) in understanding biological processes. In addition to the typical uses of RNA-seq for transcriptomics, gene annotation, novel gene discovery, and network analysis, these data can enable a deeper understanding of cellular processes through the identification of RNA modifications (epitranscriptome) and long non-coding RNAs (lncRNAs). To expedite discovery, we developed a portable, centralized computational pipeline for High-throughput Annotation of Modified Ribonucleotides and Long Non-Coding ribonucleic acids (HAMRLNC). HAMRLNC differs from existing methods by incorporating three workflows for quantifying transcript abundance, inferring RNA modifications, and lncRNA annotation using the same RNA- seq pre-processing and mapping steps. This facilitates reproducibility across multiple analyses and allows researchers to perform post-hoc analyses of archived sequencing data. In addition, we include novel analysis features to enable downstream visualization of annotated modified RNAs. HAMRLNC generates over a dozen well-defined and labeled figures as output, including gene ontology heatmaps, modification enrichment landscape, and modification clustering statistics. Availability and ImplementationHAMRLNC is an open-source software, and the source code is available at https://github.com/bdgregory/HAMRLNC. The pipeline can be installed and used through a docker container (https://hub.docker.com/r/chosenobih/HAMRLNC/tags). HAMRLNC is also available as an app in the CyVerse Discovery Environment https://de.cyverse.org/. Supplementary informationSupplementary data are available at BioRxiv online.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- racoon_clip - a complete pipeline for single-nucleotide analyses of iCLIP and eCLIP data 95%
- tinyRNA: precision analysis of small RNA-seq data with user-defined hierarchical selection rules 95%
- AnnSQL: A Python SQL-based package for fast large-scale single-cell genomics analysis using minimal computational resources 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.