PLncFire: A Scalable Pipeline for Transcriptome-wide Discovery of Plant lncRNAs
Mistry, S. D.; Saxena, S.; Patel, D.; Rizvi, A. Z.
Show abstract
Long non-coding RNAs (lncRNAs) are key regulators of plant biology, yet their discovery is hindered by low sequence conservation and a lack of comprehensive annotations. To overcome these challenges, we developed PLncFire, a modular computational pipeline that automates the genome-wide identification and annotation of lncRNAs from standard RNA-seq data. PLncFire integrates quality control, transcript assembly, and a robust consensus coding-potential assessment using CPC2, PlantLncPipe, and FEELnc to generate high-confidence predictions. It classifies lncRNAs as known or novel, facilitates their prioritisation through differential expression analysis, and is designed for scalability and reproducibility across diverse plant species. PLncFire provides a standardised framework to empower large-scale lncRNA discovery and advance comparative functional genomics. The source code is available at https://github.com/ahsan-rizvi/PLncFire.git.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Organ-level Gene Regulatory Network models enable the identification of central transcription factors in Solanum lycopersicum 94%
- HEMU: an integrated Andropogoneae comparative genomics database and analysis platform 93%
- Transcriptome brings variations of gene expression, alternative splicing, and structural variations into gene-scale trait dissection in soybean 92%
Similar papers in this journal
- The Soybean Expression Atlas v2: a comprehensive database of over 5000 RNA-seq samples 95%
- MINI-AC: Inference of plant gene regulatory networks using bulk or single-cell accessible chromatin profiles 95%
- Systematic analysis of 1,298 RNA-Seq samples and construction of a comprehensive soybean (Glycine max) expression atlas 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.