Back

Single-sequence deep learning delivers crystal-quality models of covalent K-Ras G12 hotspot complexes

Jung, S.; Zheng, Q.; Shokat, K.

2025-10-13 bioinformatics
10.1101/2025.09.16.676163 bioRxiv
Show abstract

Structure-based design of covalent drugs has achieved tremendous success by understanding and leveraging the three-dimensional interactions between small-molecule drug candidates and their protein targets. However, this approach traditionally relies on high-resolution co-complex structures obtained by X-ray crystallography, NMR, or cryo-EM, methods that are time-consuming and resource-intensive. Here we show that Chai-1, a publicly available structure prediction tool that accepts user-defined ligands, is able to accurately predict covalent K-Ras(G12C) complexes without using a multiple sequence alignment (MSA). Chai-1 yields pocket-aligned RMSDs <2 [A] for chemically diverse K-Ras(G12C) inhibitors, ranging from ARS-853 to BBO-8520. In addition to the conventional acrylamide-based covalent K-Ras(G12C) inhibitors, Chai-1 with a covalent-bond restraint successfully reproduced the binding poses of covalent K-Ras(G12D) and K-Ras(G12S) inhibitors, while showing limitations in capturing chemical details such as accounting for leaving-groups, bond properties, and stereochemistry. Chai-1 also provides [~]40-fold higher throughput than state-of-the-art AlphaFold3 while maintaining comparable pose accuracy. Together, these findings establish Chai-1 as an accessible and computationally efficient tool for covalent protein-ligand co-complex structure prediction, with its covalent-restraint mode offering a unique solution to accelerate covalent drug discovery, especially for challenging targets beyond cysteine. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=79 SRC="FIGDIR/small/676163v2_ufig1.gif" ALT="Figure 1"> View larger version (37K): org.highwire.dtl.DTLVardef@bcbad5org.highwire.dtl.DTLVardef@8e2baborg.highwire.dtl.DTLVardef@1d507a2org.highwire.dtl.DTLVardef@e851ef_HPS_FORMAT_FIGEXP M_FIG C_FIG

Published in IUBMB Life · not in our set (fewer than 10 published preprints to learn from) · training set

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.