A Systematic Re-Analysis Of Copy Number Losses Of Uncertain Clinical Significance
Burghel, G.; Ellingford, J. M.; Wright, R.; Bradford, L.; Miller, J.; Watt, C.; Edgerley, J.; Naeem, F.; Banka, S.
Show abstract
BackgroundRe-analysis of whole exome/genome data improves diagnostic yield. However, the value of re-analysis of clinical array comparative genomic hybridisation (aCGH) data has never been investigated. Case-by-case re-analysis is impractical in busy diagnostic laboratories. Methods and ResultsWe harmonised historical post-natal clinical aCGH results from [~]16,000 patients tested via our diagnostic laboratory over [~]7 years with current clinical guidance. This led to identification of 33,857 benign, 2,173 class 3, and 979 pathogenic copy number losses (CNLs). We found benign CNLs to be significantly less likely to encompass haploinsufficient genes compared to the pathogenic or class 3 CNLs in our database. Using this observation, we developed a re-analysis pipeline (using up-to-date disease association data and haploinsufficiency scores) and shortlisted 207 class 3 CNLs encompassing at least one autosomal dominant disease-gene associated with haploinsufficiency or loss-of-function mechanism. Clinical scientist review led to reclassification of 7.2% shortlisted class 3 CNLs as pathogenic or likely pathogenic. This included first cases of CNV-mediated disease for some genes where all previously described cases involved only point variants. Interestingly, some CNLs could not be re-classified because the phenotypes of patients with CNLs seemed distinct from the known clinical features resulting from point variants, thus raising questions about accepted underlying disease mechanisms. Several potential novel disease-genes were identified that would need further validation. ConclusionsRe-analysis of clinical aCGH data increases diagnostic yield and demonstrates their research value. In future, the aCGH reanalysis program should be expanded to include other copy number variant types.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Phasing of de novo mutations using a scaled-up multiple amplicon long-read sequencing approach 94%
- Matching whole genomes to rare genetic disorders: Identification of potential causative variants using phenotype-weighted knowledge in the CAGI SickKids5 clinical genomes challenge 94%
- Whole genome sequencing of ‘mutation-negative’ individuals with Cornelia de Lange Syndrome 93%
Similar papers in this journal
- Exome copy number variant detection, analysis and classification in a large cohort of families with undiagnosed rare genetic disease 95%
- HiFi long-read genomes for difficult-to-detect clinically relevant variants 94%
- The impact of 22q11.2 copy number variants on human traits in the general population 94%
Similar papers in this journal
- Next-generation phenotyping in Nigerian children with Cornelia de Lange Syndrome 93%
- Rare variants found in multiplex families with orofacial clefts: Does expanding the phenotype make a difference? 93%
- Resolving the diagnostic odyssey in inherited retinal dystrophies through long-read genome sequencing 93%
Similar papers in this journal
- Exome sequencing as a first-tier test for copy number variant detection : retrospective evaluation and prospective screening in 2418 cases 94%
- A comparative medical genomics approach may facilitate the interpretation of rare missense variation 92%
- Assessing performance of pathogenicity predictors using clinically-relevant variant datasets 91%
Similar papers in this journal
- Assessment of the variant prioritisation strategy for genomic newborn screening in the Generation Study 95%
- The Importance of Automation in Genetic Diagnosis: Lessons from Analyzing an Inherited Retinal Degeneration Cohort with the Mendelian Analysis Toolkit (MATK) 94%
- A gene pathogenicity tool 'GenePy' identifies missed biallelic diagnoses in the 100,000 Genomes Project 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.