Back

Utilising Nanopore direct RNA sequencing of blood from patients with sepsis for discovery of co- and post-transcriptional disease biomarkers

He, J.; Ganesamoorthy, D.; Chang, J. J.-Y.; Zhang, J.; Trevor, S. L.; Gibbons, K. S.; McPherson, S. J.; Kling, J. C.; Schlapbach, L. J.; Blumenthal, A.; The RAPIDS Study Group, ; Coin, L. J. M.

2024-12-14 genetic and genomic medicine
10.1101/2024.12.13.24318230 medRxiv
Show abstract

BackgroundRNA sequencing of whole blood has been increasingly employed to find transcriptomic signatures of disease states. These studies traditionally utilize short-read sequencing of cDNA, missing important aspects of RNA expression such as differential isoform abundance and poly(A) tail length variation. MethodsWe used Oxford Nanopore Technologies long-read sequencing to sequence native mRNA extracted from whole blood from 12 patients with suspected bacterial and viral sepsis, and compared with results from matching Illumina short-read cDNA sequencing data. Additionally, we explored poly(A) tail length variation, novel transcript identification and differential transcript usage. ResultsThe correlation of gene count data between Illumina cDNA and Nanopore RNA-sequencing strongly depended on the choice of analysis pipeline; NanoCount for Nanopore and Kallisto for Illumina data yielded the highest mean Pearsons correlation of 0.93 at gene level and 0.74 at transcript isoform level. We identified 18 genes significantly differentially polyadenylated and 4 genes with significant differential transcript usage between bacterial and viral infection. Gene ontology gene set enrichment analysis of poly(A) tail length revealed enrichment of long tails in signal transduction and short tails in oxidoreductase molecular functions. Additionally, we detected 594 non-artifactual novel transcript isoforms, including 9 novel isoforms for Immunoglobulin lambda like polypeptide 5 (IGLL5). ConclusionsNanopore RNA- and Illumina cDNA-gene counts are strongly correlated, indicating that both platforms are suitable for discovery and validation of gene count biomarkers. Nanopore direct RNA-seq provides additional advantages by uncovering additional post- and co-transcriptional biomarkers, such as poly(A) tail length variation and transcript isoform usage.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

1
NAR Genomics and Bioinformatics
242 papers in training set
Top 0.1%
34.6%
2
Genome Biology
637 papers in training set
Top 2%
6.8%
3
BMC Genomics
406 papers in training set
Top 1.0%
5.6%
4
Scientific Reports
3612 papers in training set
Top 18%
5.2%
50% of probability mass above
5
iScience
1154 papers in training set
Top 5%
3.5%
6
Nucleic Acids Research
1281 papers in training set
Top 6%
3.3%
7
Frontiers in Immunology
638 papers in training set
Top 4%
3.3%
8
RNA Biology
78 papers in training set
Top 0.4%
3.2%
9
Bioinformatics
1204 papers in training set
Top 6%
2.8%
10
NAR Molecular Medicine
22 papers in training set
Top 0.1%
2.5%
11
Frontiers in Genetics
230 papers in training set
Top 3%
1.7%
12
Genome Medicine
183 papers in training set
Top 3%
1.7%
13
RNA
189 papers in training set
Top 0.9%
1.5%
14
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 32%
1.3%
15
International Journal of Molecular Sciences
494 papers in training set
Top 11%
1.1%
16
Nature Communications
5641 papers in training set
Top 50%
1.1%
17
Briefings in Bioinformatics
354 papers in training set
Top 6%
1.1%
18
Genome Research
468 papers in training set
Top 5%
1.1%
19
PLOS ONE
5266 papers in training set
Top 57%
1.0%
20
Cell Genomics
172 papers in training set
Top 4%
0.9%
21
Life Science Alliance
285 papers in training set
Top 7%
0.9%
22
Computational and Structural Biotechnology Journal
242 papers in training set
Top 7%
0.9%
23
Genes
144 papers in training set
Top 5%
0.6%