Back

AGZArank: Investigating epitope-conditioned antibody binder ranking with structure-derived synthetic supervision

Sadykov, Z.; Khamidullina, A.; Sultankulov, B.; Seitkali, D.

2026-06-11 bioinformatics
10.64898/2026.06.08.730711 bioRxiv
Show abstract

AO_SCPLOWBSTRACTC_SCPLOWComputational antibody design methods can generate large libraries of candidate binders for a target epitope, but prioritizing which candidates to test experimentally remains a major bottleneck. Existing scoring approaches, including physics-based affinity estimators, structure-prediction-derived confidence measures, and inverse-folding likelihood models, provide useful proxy signals but are not explicitly optimized for early enrichment of binders among many structurally similar candidates. Here we investigate epitope-conditioned antibody binder ranking as a dedicated learning problem and introduce AGZArank, a geometric deep learning framework trained with structure-derived synthetic supervision based on normalized pseudo-energy targets. On a benchmark of 45 experimentally validated antibody-antigen interfaces, AGZArank recovered the true binder within the top ten candidates in 44.4% of cases and showed stronger generalization on post-2021 structures than ProteinMPNN, ESM-IF, and PRODIGY. Ablation experiments indicate that ranking performance depends primarily on training scale and alignment between the optimization objective and retrieval-based evaluation, rather than architectural complexity alone. These results support candidate prioritization as a distinct and tractable problem in computational antibody design.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Nature Communications
5641 papers in training set
Top 16%
11.6%
2
mAbs
32 papers in training set
Top 0.1%
11.6%
3
Cell Systems
201 papers in training set
Top 0.4%
7.7%
4
Bioinformatics
1204 papers in training set
Top 3%
7.7%
5
PLOS Computational Biology
1863 papers in training set
Top 6%
6.6%
6
Nature Machine Intelligence
70 papers in training set
Top 0.5%
5.3%
50% of probability mass above
7
Briefings in Bioinformatics
354 papers in training set
Top 2%
5.3%
8
Journal of Chemical Information and Modeling
238 papers in training set
Top 1%
3.9%
9
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 14%
3.9%
10
Protein Science
246 papers in training set
Top 2%
2.3%
11
PRX Life
42 papers in training set
Top 0.4%
2.1%
12
Bioinformatics Advances
203 papers in training set
Top 3%
1.9%
13
Computational and Structural Biotechnology Journal
242 papers in training set
Top 4%
1.7%
14
Communications Biology
993 papers in training set
Top 15%
1.7%
15
Patterns
78 papers in training set
Top 1%
1.7%
16
Nature Methods
385 papers in training set
Top 4%
1.7%
17
eLife
5828 papers in training set
Top 50%
1.7%
18
Journal of Chemical Theory and Computation
140 papers in training set
Top 0.8%
1.6%
19
Frontiers in Immunology
638 papers in training set
Top 6%
1.6%
20
Communications Chemistry
48 papers in training set
Top 0.9%
1.3%
21
Cell Reports Methods
165 papers in training set
Top 3%
1.1%
22
Scientific Reports
3612 papers in training set
Top 70%
1.0%
23
Nature Computational Science
55 papers in training set
Top 1%
1.0%
24
Structure
193 papers in training set
Top 2%
0.8%
25
Advanced Science
286 papers in training set
Top 10%
0.8%
26
Nature Structural & Molecular Biology
18 papers in training set
Top 0.5%
0.8%
27
Nucleic Acids Research
1281 papers in training set
Top 14%
0.8%
28
ImmunoInformatics
12 papers in training set
Top 0.2%
0.8%
29
Cell Reports
1498 papers in training set
Top 30%
0.6%