Benchmarking Generative Models for Antibody Design
Ucar, T.; Malherbe, C.; Gonzalez Hernandez, F.
Show abstract
Generative models trained on antibody sequences and structures have shown great potential in advancing machine learning-assisted antibody engineering and drug discovery. Current state-of-the-art models are primarily evaluated using two categories of in silico metrics: sequence-based metrics, such as amino acid recovery (AAR), and structure-based metrics, including root-mean-square deviation (RMSD), predicted alignment error (pAE), and interface predicted template modeling (ipTM). While metrics such as pAE and ipTM have been shown to be useful filters for experimental success, there is no evidence that they are suitable for ranking, particularly for antibody sequence designs. Furthermore, no reliable sequence-based metric for ranking has been established. In this work, using real-world experimental data from fourteen diverse datasets, we extensively benchmark a range of generative models, including LLM-style, diffusion-based, and graph-based models. We show that log-likelihood scores from these generative models have promising correlation with experimentally measured binding affinities, suggesting that log-likelihood can potentially serve as a reliable metric for ranking antibody sequence designs. Additionally, we scale up one of the diffusion-based models by training it on a large and diverse synthetic dataset, significantly enhancing its ability to rank antibodies based on their binding affinities. We also evaluate non-log-likelihood-based metrics on ten datasets and find that, while they are less consistent for ranking, they provide complementary information. Structure-, energy-, and sequence-based scores appear to be orthogonal and may be used together to increase the likelihood of experimental success. Our implementation is available at: https://github.com/AstraZeneca/DiffAbXL
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Learning Context-aware Structural Representations to Predict Antigen and Antibody Binding Interfaces 97%
- ParaSurf: A Surface-Based Deep Learning Approach for Paratope-Antigen Interaction Prediction 97%
- Paragraph - Antibody paratope prediction using Graph Neural Networks with minimal feature vectors 96%
Similar papers in this journal
- Prediction of Antibody Non-Specificity using Protein Language Models and Biophysical Parameters 94%
- Ab-Ligity: Identifying sequence-dissimilar antibodies that bind to the same epitope 94%
- Towards generalizable prediction of antibody thermostability using machine learning on sequence and structure features 94%
Similar papers in this journal
- Guiding a language-model based protein design method towards MHC Class-I immune-visibility profiles for vaccines and therapeutics 97%
- DoggifAI: a transformer based approach for antibodycaninisation 96%
- Do Domain-Specific Protein Language Models Outperform General Models on Immunology-Related Tasks? 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.