A rare splice-site variant in cardiac troponin-T (TNNT2): The need for ancestral diversity in genomic reference datasets
Butters, A.; Thomson, K.; Harrington, F.; Henden, N.; McGuire, K.; Byrne, A.; Bryen, S.; Leask, M.; Ackerman, M. J.; Atherton, J. J.; Bos, J. M.; Caleshu, C.; Parikh, V. N.; Day, S.; Tardiff, J. C.; Dunn, K.; Hayes, I.; Juang, J.; McGaughran, J.; Nowak, N.; Ronan, A.; Semsarian, C.; Tiemensma, M.; Merriman, T.; Skinner, J. R.; MacArthur, D. G.; Siggs, O. M.; Bagnall, R. D.; Ingles, J.
Show abstract
The underrepresentation of different ancestry groups in large genomic datasets creates difficulties in interpreting the pathogenicity of monogenic variants. Genetic testing for individuals with non-European ancestry results in higher rates of uncertain variants and a greater risk of misclassification. We report a rare variant in the cardiac troponin T gene, TNNT2; NM_001001430.3: c.571-1G>A (rs483352835) identified via research-based whole exome sequencing in two unrelated probands of Oceanian ancestry with cardiac phenotypes. The variant disrupts the canonical splice acceptor site, activating a cryptic acceptor and resulting in an in-frame deletion (p.Gln191del). The variant is rare in gnomAD v4.0.0 (13/780,762; 0.002%), with the highest frequency in South Asians (5/74,486; 0.007%) and has 16 ClinVar assertions (13 diagnostic clinical laboratories classify as variant of uncertain significance). There are at least 28 reported cases, many with Oceanian ancestry and diverse cardiac phenotypes. Indeed, among Oceanian-ancestry-matched datasets, the allele frequency ranges from 2.9-8.8% and is present in 2/4 (50%) Indigenous Australian alleles in Genome Asia 100K, with one participant being homozygous. With Oceanians deriving greater than 3% of their DNA from archaic genomes, we found c.571-1G>A in Vindija and Altai Neanderthal, but not the Altai Denisovan, suggesting an origin post Neanderthal divergence from modern humans 130-145 thousand years ago. Based on these data, we classify this variant as benign, and conclude it is not a monogenic cause of disease. Even with ongoing efforts to increase representation in genomics, we highlight the need for caution in assuming rarity of genetic variants in largely European datasets. Efforts to enhance diversity in genomic databases remain crucial.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Rare Genetic Variants Associated with Sudden Cardiac Arrest in the Young: A Prospective, Population-Based Study 95%
- Prevalence and disease expression of pathogenic and likely pathogenic variants associated with inherited cardiomyopathies in the general population 95%
- Clinical genetic risk variants inform a functional protein interaction network for tetralogy of Fallot 95%
Similar papers in this journal
- The Clinical Genome Resource (ClinGen) Familial Hypercholesterolemia Variant Curation Expert Panel consensus guidelines for LDLR variant classification 94%
- Disease-specific variant pathogenicity prediction significantly improves variant interpretation in inherited cardiac conditions 93%
- Variant Curation Expert Panel Recommendations for RYR1 Pathogenicity Assertions in Malignant Hyperthermia Susceptibility 93%
Similar papers in this journal
- Genomic rare variant mechanisms for congenital cardiac laterality defect: A digenic model approach 94%
- Identification of actionable genetic variants in 4,198 Scottish volunteers from the Viking Genes research cohort and implementation of return of results 94%
- The penetrance of rare variants in cardiomyopathy-associated genes: a cross-sectional approach to estimate penetrance for secondary findings 93%
Similar papers in this journal
- Performance of a Protein Language Model for Variant Annotation in Cardiac Disease 95%
- Pathogenic and uncertain genetic variants have clinical cardiac correlates in diverse biobank participants 95%
- The effect of sex and underlying disease on the genetic association of QT interval and sudden cardiac death 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.