CovSite: A High-Throughput Blind Covalent Screening Framework for Reactive Site Detection
Hu, A.; Bailey, J. S.; Spina, S. C.; Rajagopal, G.; Phan, N.; Kimmel, B. R.
Show abstract
Targeted covalent inhibitors are a powerful, yet underexplored, class of therapeutics, and current computational covalent screeners are constrained in early drug discovery due to the need for prior knowledge of the target site and limited throughput. We present CovSite, a blind covalent screening tool that identifies candidate reactive residues across the entire protein surface, utilizing only the protein structure and electrophile SMILES. CovSite applies a pipeline of four orthogonal physicochemical filters (nucleophile identification, solvent accessibility, environment-dependent deprotonation prediction, and semi-quantum-mechanical reactivity ranking) to identify potential small-molecule candidate inhibitors. Validated against 2,062 diverse covalent protein-ligand complexes spanning six nucleophilic residue types, CovSite achieves a 98.5% blind target site hit on a held-out benchmark set of 207 cysteine-targeted complexes while reducing the search space by 97.8%. The target-site hit detection exceeds the 53-62% accuracy of popular covalent screening tools operating under non-blind conditions on the same benchmark set. By extending nucleophilic coverage beyond cysteine to include serine, threonine, lysine, histidine, and tyrosine, and completing a screening of a 200-residue protein in two to three minutes on standard hardware, CovSite serves as a platform technology with the potential to address critical gaps in throughput, generalizability, and accuracy in this field of covalent screening. We demonstrate this capability by using CovSite as a blind, ligand-specific approach that enables iterative, machine-learning-driven covalent inhibitor generation that is impractical with existing tools, establishing a foundation for computationally guided covalent drug discovery for novel and understudied targets. TOC Figure O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=100 SRC="FIGDIR/small/739288v1_ufig1.gif" ALT="Figure 1"> View larger version (27K): org.highwire.dtl.DTLVardef@17abeeforg.highwire.dtl.DTLVardef@18d6299org.highwire.dtl.DTLVardef@14437eaorg.highwire.dtl.DTLVardef@1b2ec81_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Deep Docking - a Deep Learning Approach for Virtual Screening of Big Chemical Datasets 94%
- The Interplay of Electrostatics and Chemical Positioning in the Evolution of Antibiotic Resistance in TEM β-Lactamases 93%
- A High-Throughput Screen Reveals the Structure-Activity Relationship of the Antimicrobial Lasso Peptide Ubonodin 93%
Similar papers in this journal
- Integrative x-ray structure and molecular modeling for the rationalization of procaspase-8 inhibitor potency and selectivity 94%
- Rationalizing diverse binding mechanisms to the same protein fold: in-sights for ligand recognition and biosensor design 94%
- Analogs of the Dopamine Metabolite 5,6-Dihydroxyindole Bind Directly to and Activate the Nuclear Receptor Nurr1 (NR4A2) 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.