Back

ComplexDesign: sequence-hallucination design of protein binders bridging multiple proteins

Xu, J.; Ren, M.; Qi, N.; Zhang, X.; He, Z.; Yu, C.; Bu, D.

2026-06-24 bioinformatics
10.64898/2026.06.21.733655 bioRxiv
Show abstract

MotivationDesigning multichain protein complexes requires coordinating the folding of component proteins with the formation of their interfaces. The existing methods, however, remain limited in their ability to satisfy these requirements simultaneously, especially for trimeric and tetrameric complexes. As an important practical scenario, designing a binder that bridges two target proteins into a ternary complex requires flexibility in the relative arrangement of the two targets, adding an additional challenge to existing design methods. ResultsWe present ComplexDesign, a hallucination-based approach for multichain protein design. ComplexDesign performs structure-prediction-guided sequence optimization to simultaneously fold each protein chain and form inter-chain interactions that bind them together. To provide the flexibility required to appropriately arrange these target proteins, ComplexDesign introduces a specialized masking mechanism that enables exploration of possible relative arrangements rather than being limited to the predefined ones. Across a comprehensive set of benchmarks with various chain lengths, ComplexDesign outperformed existing methods in the unconditional design of dimers, trimers, and tetramers, achieving a high design success rate exceeding 50%, supporting its capability for multichain complex design. Furthermore, in the case of multi-target binder design, ComplexDesign produced high-confidence, self-consistent ternary complexes for 8 out of 10 target pairs. These results establish ComplexDesign as an effective tool for multichain protein design, with particular utility for designing binders that bridge two target proteins. Availability and implementationThe source code of ComplexDesign will be made publicly available upon publication.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
Bioinformatics
1204 papers in training set
Top 0.6%
30.3%
2
Journal of Chemical Information and Modeling
238 papers in training set
Top 0.5%
11.6%
3
Protein Science
246 papers in training set
Top 0.3%
10.8%
50% of probability mass above
4
Journal of Molecular Biology
232 papers in training set
Top 0.5%
4.7%
5
Briefings in Bioinformatics
354 papers in training set
Top 2%
4.2%
6
Bioinformatics Advances
203 papers in training set
Top 1%
4.2%
7
PLOS Computational Biology
1863 papers in training set
Top 9%
3.4%
8
Computational and Structural Biotechnology Journal
242 papers in training set
Top 2%
3.2%
9
BMC Bioinformatics
457 papers in training set
Top 3%
3.1%
10
Journal of Chemical Theory and Computation
140 papers in training set
Top 0.7%
2.1%
11
Journal of Cheminformatics
29 papers in training set
Top 0.3%
2.1%
12
Nature Communications
5641 papers in training set
Top 44%
1.9%
13
Proteins: Structure, Function, and Bioinformatics
88 papers in training set
Top 0.8%
1.7%
14
Structure
193 papers in training set
Top 2%
1.4%
15
Protein Engineering, Design and Selection
15 papers in training set
Top 0.2%
1.0%
16
ACS Omega
105 papers in training set
Top 3%
1.0%
17
Scientific Reports
3612 papers in training set
Top 71%
1.0%
18
Cell Systems
201 papers in training set
Top 5%
0.8%
19
PLOS ONE
5266 papers in training set
Top 66%
0.6%
20
Nucleic Acids Research
1281 papers in training set
Top 15%
0.6%
21
Communications Chemistry
48 papers in training set
Top 2%
0.6%