Back

Evaluation of ACMG Rules for In silico Evidence Strength Using An Independent Computational Tool Absent of Circularities on ATM and CHEK2 Breast Cancer Cases and Controls

Young, C. C.; Feng, B.-J.; Mackenzie, C. B.; Girard, E.; Hu, D.; Iwasaki, Y.; Momozawa, Y.; Lesueur, F.; Ziv, E. C.; Neuhausen, S. L.; Tavtigian, S. V.

2019-11-08 genetics
10.1101/835264 bioRxiv
Show abstract

The American College of Medical Genetics and Genomics (ACMG) guidelines for sequence variant classification include two criteria, PP3 and BP4, for combining computational data with other evidence types contributing to sequence variant classification. PP3 and BP4 assert that computational modeling can provide \"Supporting\" evidence for or against pathogenicity within the ACMG framework. Here, leveraging a meta-analysis of ATM and CHEK2 breast cancer case-control mutation screening data, we evaluate the strength of evidence determined from the relatively simple computational tool Align-GVGD. Importantly, application of Align-GVGD to these ATM and CHEK2 data is free of logical circularities, hidden multiple testing, and use of other ACMG evidence types. For both genes, rare missense substitutions that are assigned the most severe Align-GVGD grade exceed a \"Moderate pathogenic\" evidence threshold when analyzed in a Bayesian framework; accordingly, we argue that the ACMG classification rules be updated for well-calibrated computational tools. Additionally, congruent with previous analyses of ATM and CHEK2 case-control mutation screening data, we find that both genes have a considerable burden of pathogenic missense substitutions, and that severe ATM rare missense have increased odds ratios compared to truncating and splice junction variants, indicative of a potential dominant-negative effect for those missense substitutions.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.