Targeted long-read RNA sequencing reveals the complexity of CLN3 transcription and the consequences of the most common 1-kb deletion in patients with juvenile CLN3 disease
Minnis, C. J.; Zhang, H.-Y.; Gustavsson, E.; Lee, J.; Anderson, C.; Gardner, E.; Schulz, A.; Nickel, M.; Arno, G.; Jurkute, N.; Webster, A. R.; Gammaldi, N.; Santorelli, F. M.; Gissen, P.; Mills, P.; Ryten, M.; Mole, S. E.
Show abstract
Most genes are not yet fully annotated, and the extent of their transcript diversity and the roles and significance of specific isoforms is not understood. This information is therefore lacking for disease genes. The CLN3 gene underlies classic juvenile CLN3 disease, also known as juvenile neuronal ceroid lipofuscinosis, a rare paediatric neurodegenerative disorder. The most common cause of this biallelic disorder is a 1-kb intragenic deletion that removes two internal coding exons (exons 7 and 8). Here, we report findings from the first long-read RNA sequencing targeting CLN3 in blood samples derived from control individuals and from patients clinically and genetically diagnosed with juvenile CLN3 disease. We find that CLN3 transcription is complex, with >80 different transcripts encoding >35 different open reading frames (ORF) of different lengths, and no dominantly expressed transcript. The 1-kb deletion has direct consequences on this. This is consistent across patients, with total loss of some transcripts including those encoding the canonical 438 amino acid protein and other significant smaller isoforms. The highest expressed disease transcripts include those lacking exons 7 and 8 and encoding a 181 amino acid protein isoform, and other novel isoforms that lack additional exons and encode longer ORFs. The different effects on transcription of other CLN3 disease-causing variants are revealed in single patients. Together, these findings confirm the complexity of transcription at the CLN3 locus, reveal the impact of the 1-kb deletion and other variants on isoform abundance, and highlight the importance of understanding the contribution of these isoforms to CLN3 function in health and disease. Moreover, they impact the future design and development of personalised therapeutics and the design and generation of disease models. Finally, they underline the importance of full annotation for disease genes.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A KLHL40 3’ UTR splice-altering variant causes milder NEM8, an under-appreciated disease mechanism 94%
- Familial ALS/FTD-associated RNA-Binding deficient TDP-43 mutants cause neuronal and synaptic transcript dysregulation in vitro 92%
- Genomic dissection of 43 serum urate-associated loci provides multiple insights into molecular mechanisms of urate control. 92%
Similar papers in this journal
- Using single molecule Molecular Inversion Probes as a cost-effective, high-throughput sequencing approach to target all genes and loci associated with macular diseases 94%
- Type II Alexander disease caused by splicing errors and aberrant overexpression of an uncharacterized GFAP isoform. 93%
- Whole genome sequencing of ‘mutation-negative’ individuals with Cornelia de Lange Syndrome 93%
Similar papers in this journal
- Biallelic pathogenic variants in TRMT1 disrupt tRNA modification and induce a syndromic neurodevelopmental disorder 93%
- MRSD: a novel quantitative approach for assessing suitability of RNA-seq in the clinical investigation of mis-splicing in Mendelian disease 93%
- Clinical validation of RNA sequencing for Mendelian disorder diagnostics 93%
Similar papers in this journal
- Genotype–phenotype correlations and novel molecular insights into the DHX30 -associated neurodevelopmental disorders 94%
- Clustering of predicted loss-of-function variants in genes linked with monogenic disease can explain incomplete penetrance 92%
- Mendelian gene identification through mouse embryo viability screening 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.