PXDesign: Fast, Modular, and Accurate De Novo Design of Protein Binders
Ren, M.; Sun, J.; Guan, J.; Liu, C.; Gong, C.; Wang, Y.; Wang, L.; Cai, Q.; Chen, X.; Xiao, W.
Show abstract
PXDesign achieves nanomolar binder hit rates of 17-82% across six of seven diverse protein targets, surpassing prior methods such as AlphaProteo. This experimental success rate is enabled by advances in both binder generation and filtering. We develop both a diffusion-based generative model (PXDesign-d) and a hallucination-based approach (PXDesign-h), each showing strong in silico performance that outperforms existing models. Beyond generation, we systematically analyze confidence-based filtering and ranking strategies from multiple structure predictors, comparing their accuracy, efficiency, and complementarity on datasets spanning de novo binders and mutagenesis. Finally, we validate the full design process experimentally, achieving high hit rates and multiple nanomolar binders. To support future research and broaden community adoption, we release the full PXDesign pipeline (https://github.com/bytedance/PXDesign), provide public access to PXDesign through a dedicated web server (https://protenix-server.com), and make all designed binder sequences available at the project page (https://protenix.github.io/pxdesign). O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=75 SRC="FIGDIR/small/670450v3_ufig1.gif" ALT="Figure 1"> View larger version (25K): org.highwire.dtl.DTLVardef@2a414aorg.highwire.dtl.DTLVardef@248f85org.highwire.dtl.DTLVardef@4a6afdorg.highwire.dtl.DTLVardef@1b622dd_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Neural Network-Derived Potts Models for Structure-Based Protein Design using Backbone Atomic Coordinates and Tertiary Motifs 96%
- COLLAPSE: A representation learning framework for identification and characterization of protein structural sites 96%
- AlphaFold Model Quality Self-Assessment Improvement Via Deep Graph Learning 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.