Mitochondrial mRNA is a stable and convenient reference for the normalization of RNA-seq data
Liu, Y.; Zhang, Y.; Lu, F.; Wang, J.
Show abstract
The normalization of high-throughput RNA sequencing (RNA-seq) data is needed to accurately analyze gene expression levels. Traditional normalization methods can either correct the differences in sequencing depth, or correct both the sequencing depth and other unwanted variations introduced during sequencing library preparation through exogenous spike-ins1-4. However, the exogenous spike-ins are prone to variation5,6. Therefore, a better normalization approach with a more appropriate reference is an ongoing demand. In this study, we demonstrated that mitochondrial mRNA (mRNA encoded by mitochondria genome) can serve as a steady endogenous reference for RNA-seq data analysis, and performs better than exogenous spike-ins. We also found that using mitochondrial mRNA as a reference can reduce batch effects for RNA-seq data. These results provide a simple and practical normalization strategy for RNA-seq data, which will serve as a valuable tool widely applicable to transcriptomic studies.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Single-molecule long-read sequencing reveals a conserved selection mechanism determining intact long RNA and miRNA profiles in sperm 95%
- Maternal mRNA deadenylation is defective in in vitro matured mouse and human oocytes 95%
- Reference-free assembly of long-read transcriptome sequencing data with RNA-Bloom2 94%
Similar papers in this journal
Similar papers in this journal
- Multi-sample Full-length Transcriptome Analysis of 22 Breast Cancer Clinical Specimens with Long-Read Sequencing 94%
- Targeted Transcriptome Analysis using Synthetic Long Read Sequencing Uncovers Isoform Reprograming in the Progression of Colon Cancer 93%
- ATAC-seq with unique molecular identifiers improves quantification and footprinting 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.