SALRR: Scalable Analysis of Long-Read RNA-Seq Enables Comprehensive Transcriptome Profiling in Human Brain
Kouam, C.; Mingle, J.; Alvarez Jerez, P.; Evans, A.; Moller, A.; Baker, B.; Weller, C.; Paquette, K.; Brooks, J.; Grant, S. M.; Ayuketah, A.; Meredith, M.; Palade, J.; Malik, L.; Hise, K.; Raphael Gibbs, J.; Anderson, J.; Ding, J.; Harbert, R.; Fu, Y.; Zheng, X.; Garcia-Ruiz, S.; Gustavsson, E. K.; Blauwendraat, C.; Ryten, M.; Sedlazeck, F.; Ferrucci, L.; Reed, X.; Nalls, M. A.; Cookson, M. R.; Van Keuren-Jensen, K.; Hutchins, E.; Jain, M.; Billingsley, K. J.
Show abstract
Isoform-resolved transcriptomics is fundamental to decoding the molecular complexity of the human brain, yet population-scale long-read RNA sequencing has remained inaccessible due to labor-intensive library preparation, sensitivity to RNA degradation in postmortem tissue, and the absence of integrated, reproducible analysis pipelines. Here we present SALRR (Scalable Analysis of Long-Read RNA-seq), an integrated wet-lab and computational platform designed to overcome these barriers. Automated ONT long-read cDNA library preparation on the Hamilton Microlab NGS STAR platform reduces hands-on time by 67% and enables 24 libraries per operator per day while maintaining performance across RNA integrity values. A modular, Snakemake-based pipeline performs end-to-end processing from ONT signal data to isoform-level quantification, incorporating SIRV spike-in calibration, multi-stage quality control, and stringent isoform validation. Applied to 10 postmortem frontal cortex samples from the North American Brain Expression Consortium, SALRR identified 31,607 high-confidence isoforms from 10,075 genes, including 8,532 novel splice variants absent from GENCODE v49, and complex splicing events systematically missed by short-read sequencing at neurodegeneration-relevant loci, including GBA1, CCNF, CHCHD10, and TREM2. All protocols and code are openly available, providing a scalable, community-ready framework for isoform-resolved transcriptomics in neurodegeneration, aging, and complex brain disease.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Characterising tandem repeat complexities across long-read sequencing platforms with TREAT and otter 94%
- Allele-specific splicing modulates protein isoforms and Alzheimer's risk 93%
- Quantitative mitochondrial DNA copy number determination using droplet digital PCR with single cell resolution: a focus on aging and cancer 92%
Similar papers in this journal
- Human-lineage-specific genomic elements: relevance to neurodegenerative disease and APOE transcript usage 93%
- An integrative systems-biology approach defines mechanisms of Alzheimer's disease neurodegeneration 93%
- Molecular Signatures of Resilience to Alzheimer's Disease in Neocortical Layer 4 Neurons 93%
Similar papers in this journal
- MitoDelta: identifying mitochondrial DNA deletions at cell-type resolution from single-cell RNA sequencing data 92%
- lute: estimating the cell composition of heterogeneous tissue with varying cell sizes using gene expression 90%
- scRAPID-web: a web server for predicting protein-RNA interactions from single-cell transcriptomics 90%
Similar papers in this journal
- Accurate characterization of expanded tandem repeat length and sequence through whole genome long-read sequencing on PromethION. 94%
- Cell-type specific inference from bulk RNA-sequencing data by integrating single cell reference profiles via EPIC-unmix 92%
- Predicting Disease-Specific Histone Modifications and Functional Effects of Non-coding Variants by Leveraging DNA Language Models 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.