Back

YDS-Ternoplex: Surpassing AlphaFold 3-Type Models for Molecular Glue-Mediated Ternary Complex Prediction

Che, X.; The Statistical Mechanics Inference Team,

2025-02-03 biophysics
10.1101/2024.12.23.630090 bioRxiv
Show abstract

Molecular glues represent an innovative class of drugs that enable previously impossible protein-protein interactions, but their rational design remains challenging, a problem that accurate ternary complex modeling can significantly address. Here we present YDS-GlueFold, a novel computational approach that enhances AlphaFold 3-type models by incorporating guided diffusion during inference to accurately predict molecular glue-induced ternary complex structures. We demonstrate YDS-GlueFolds capabilities across eight diverse test cases, including both E3 ligase-based systems (VHL, CRBN complexes with mTOR-FRB, NEK7, and VAV1-SH3c, and KBTBD4 complexes with HDAC1) and non-E3 ligase complexes (FKBP12 complexes with mTOR-FRB, BRD9 and QDPR). The model achieves remarkable accuracy with RMSD values consistently below 2.5 [A] compared to experimental structures. Importantly, 7 out of 8 test cases involve protein-protein pairs that were not present in the AlphaFold 3 training set, providing a rigorous test of the models ability to generalize beyond its training data. Notably, in the FKBP12 case, YDS-GlueFold correctly predicts a novel interface configuration instead of defaulting to known interactions present in training data, further demonstrating true generalization rather than mere memorization. Our results suggest that strategic enhancement of the inference process through guided diffusion can significantly improve ternary complex prediction accuracy, potentially accelerating the development of molecular glue therapeutics for previously undruggable targets.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.