Why variant effect predictors and multiplexed assays agree and disagree
Livesey, B. J.; Marsh, J. A.
Show abstract
Computational variant effect predictors (VEPs) and multiplexed assays of variant effect (MAVEs) are two key tools for assessing the functional consequences of genetic variants. While their outputs are often concordant, there are also many differences. Here, we analyse missense MAVE data from 37 different human proteins, comparing them to five state-of-the-art VEPs in order to quantify and explain their points of agreement and disagreement. We find that discordance is not random but reflects fundamental differences in how each method infers functional impact. VEPs, which rely heavily on sequence conservation and basic structural features, tend to overcall pathogenicity at buried and hydrophobic residues, while underestimating impact in disordered regions and on charged surface residues. MAVEs, by contrast, capture context-specific mechanisms more accurately, but can miss pathogenic variants when the assay fails to reflect disease biology, or be subject to high levels of experimental noise. By comparing both global patterns and specific clinically relevant variants, we show how protein features, assay design, and variant type shape predictive discordance. Our findings provide a framework for interpreting when and why VEPs and MAVEs diverge and point toward strategies for improving variant interpretation through integrated, mechanism-aware approaches.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- GeneBreaker: Variant simulation to improve the diagnosis of Mendelian rare genetic diseases 94%
- Matching whole genomes to rare genetic disorders: Identification of potential causative variants using phenotype-weighted knowledge in the CAGI SickKids5 clinical genomes challenge 93%
- Spectrum of pathogenic variants and multiple founder effects in amelogenesis imperfecta associated with MMP20 93%
Similar papers in this journal
Similar papers in this journal
- StabilitySort: assessment of protein stability changes on a genome-wide scale to prioritise potentially pathogenic genetic variation 96%
- Rhapsody: Pathogenicity prediction of human missense variants based on protein sequence, structure and dynamics 95%
- DrivR-Base: A Feature Extraction Toolkit For Variant Effect Prediction Model Construction 94%
Similar papers in this journal
- Updated benchmarking of variant effect predictors using deep mutational scanning 98%
- Using deep mutational scanning to benchmark variant effect predictors and identify disease mutations 95%
- VEFill: a model for accurate and generalizable deep mutational scanning score imputation across protein domains 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.