Standardized transcriptome analysis improves rare disease diagnosis in the pan-European Solve-RD consortium
Yepez, V. A.; Luknarova, R.; Beijer, D.; Estevez-Arias, B.; Mei, D.; Morsy, H.; Mueller, J. S.; Polavarapu, K.; Demidov, G.; Doornbos, C.; Ellwanger, K.; Krass, L.; Laurie, S.; Matalonga, L.; Abdelrazek, I. M.; Astuti, G.; Bisulli, F.; Brechtmann, F.; Dabad, M.; Denomme Pichon, A. S.; Drakos, M.; Eddafir, Z.; Garrabou, G.; Guerrini, R.; Johari, M.; Kegele, J.; Kilicarslan, O. A.; Koelbel, H.; Kolen, I. H. M.; Licchetta, L.; Lochmueller, H.; Maassen, K.; Macken, W.; Mertes, C.; Milisenda, J. C.; Minardi, R.; Mostacci, B.; Neveling, K.; Oud, M. M.; Park, J.; Pujol, A.; Roos, A.; Sagath, L.; van
Show abstract
RNA sequencing (RNA-seq) provides a powerful complement to DNA sequencing for uncovering pathogenic defects affecting gene expression and splicing in individuals with genetically undiagnosed rare disorders. However, as large rare disease consortia adopt RNA-seq, challenges arise due to cohort heterogeneity, variability in tissues and sample sizes, and differences in interpretation practices. Here, we present a harmonized analytical and interpretation framework developed by the pan-European Solve-RD consortium to address these challenges. We analyzed 521 RNA-seq samples from whole blood, fibroblasts, muscle and peripheral blood mononuclear cells collected across more than 30 clinics and five European Reference Networks. Aberrant expression and splicing events were identified using OUTRIDER and FRASER 2.0 and analysed through a standardized four-level scoring framework that encompassed RNA-seq outlier reliability, phenotype relevance, variant mechanism, and segregation evidence, captured in structured reports for interpretation. Regular meetings, and collaborative "Solvathon" workshops were used to evaluate variant pathogenicity. This effort resulted in molecular diagnoses for 19 families out of 248 (7.7%) for whom DNA analyses had been inconclusive. Furthermore, three cases diagnosed using DNA analyses were confirmed, and 49 candidate events and five novel candidate disease genes were identified in the remaining families. Our results demonstrate the feasibility and impact of large-scale, standardized RNA-seq analysis in a transnational research setting. This framework provides a model for other international initiatives such as the Undiagnosed Diseases Network and ERDERA, paving the way for broader clinical implementation of transcriptome-based rare disease diagnostics.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Advancing long-read nanopore genome assembly and accurate variant calling for rare disease detection 97%
- HiFi long-read genomes for difficult-to-detect clinically relevant variants 96%
- MRSD: a novel quantitative approach for assessing suitability of RNA-seq in the clinical investigation of mis-splicing in Mendelian disease 96%
Similar papers in this journal
- A systematic analysis of splicing variants identifies new diagnoses in the 100,000 Genomes Project. 95%
- A polyclonal allelic expression assay for detecting regulatory effects of transcript variants 95%
- Systematic analysis of genetic and phenotypic characteristics reveals antisense oligonucleotide therapy potential for one-third of neurodevelopmental disorders 94%
Similar papers in this journal
- Comprehensive reanalysis for CNVs in ES data from unsolved rare disease cases results in new diagnoses 93%
- WEGS: a cost-effective sequencing method for genetic studies combining high-depth whole exome and low-depth whole genome 93%
- Discordance between a deep learning model and clinical-grade variant pathogenicity classification in a rare disease cohort 93%
Similar papers in this journal
- Dominant variants in major spliceosome U4 and U5 small nuclear RNA genes cause neurodevelopmental disorders through splicing disruption 95%
- Biallelic variants in RNU2-2 cause the most prevalent known recessive neurodevelopmental disorder 93%
- Inferring compound heterozygosity from large-scale exome sequencing data 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.