Back

SpecLig: Energy-Guided Hierarchical Model for Target-Specific 3D Ligand Design

Zhang, P.; Han, R.; Kong, X.; Chen, T.; Ma, J.

2026-02-19 bioinformatics
10.1101/2025.11.06.687093 bioRxiv
Show abstract

Structure-based generative models often optimize single-target affinity with ignorance of specificity, resulting in the generation of high-affinity candidates that exhibit promiscuous binding across unrelated targets. This decoupling of affinity and specificity not only compromises therapeutic efficacy but also elevates off-target risks that constrain translational potential. Therefore, we introduce SpecLig, a unified structure-based framework that jointly generates small molecules and peptides with improved target affinity and specificity. SpecLig represents a complex as a block-based graph, combining a hierarchical SE(3)-equivariant variational autoencoder with an energy-guided geometric latent-diffusion model. Chemical priors derived from block-block contact statistics are explicitly incorporated, biasing generation towards pocket-complementary fragment combinations. We benchmark SpecLig on peptide and small-molecule tasks using standard public datasets and propose precision/breadth testing paradigms to quantify specificity. Across multiple evaluations, ligand candidates generated by SpecLig usually bind to the target pocket with high specificity and affinity while maintaining competitive advantages in other attributes. Ablations indicate that both hierarchical representation and energy guidance contribute to success. Finally, we present multiple real applications that demonstrate how SpecLig improves ligands in natural complexes to mitigate potential off-target risks. SpecLig, therefore, provides a practical route to prioritize higher-specificity designs for downstream experimental validation. The codes are available at: https://github.com/CQ-zhang-2016/SpecLig.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.