Accelerating Antibody Development: Sequence and Structure-Based Models for Predicting Developability Properties through Size Exclusion Chromatography
Abeer, A. N. M. N.; Boroumand, M.; Sermadiras, I.; Caldwell, J. G.; Stanev, V.; Mody, N.; Kaplan, G.; Savery, J.; Croasdale-Wood, R.; Pouryahya, M.
Show abstract
Experimental screening for biopharmaceutical developability properties typically relies on resource-intensive, and time-consuming assays such as size exclusion chromatography (SEC). This study highlights the potential of in silico models to accelerate the screening process by exploring sequence and structure-based machine learning techniques. Specifically, we compared surrogate models based on pre-computed features extracted from sequence and predicted structure with sequence-based approaches using protein language models (PLMs) like ESM-2. In addition to different end-to-end fine-tuning strategies for PLM, we have also investigated the integration of the structural information of the antibodies into the prediction pipeline through graph neural networks (GNN). We applied these different methods for predicting protein aggregation propensity using a dataset of approximately 1200 Immunoglobulin G (IgG1) molecules. Through this empirical evaluation, our study identifies the most effective in silico approach for predicting developability properties for SEC assays, thereby adding insights to existing screening efforts for accelerating the antibody development process.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Gated Graph Transformer for Protein ComplexStructure Quality Assessment and its Performancein CASP15 96%
- Identifying B-cell epitopes using AlphaFold2 predicted structures and pretrained language model 96%
- Improving sequence-based modeling of protein families using secondary structure quality assessment 96%
Similar papers in this journal
- SPDesign: protein sequence designer based on structural sequence profile using ultrafast shape recognition 96%
- EGRET: Edge Aggregated Graph Attention Networks and Transfer Learning Improve Protein-Protein Interaction Site Prediction 96%
- Kinase-Inhibitor Binding Affinity Prediction with Pretrained Graph Encoder and Language Model 95%
Similar papers in this journal
- DISTEMA: distance map-based estimation of single protein model accuracy with attentive 2D convolutional neural network 95%
- Struct2Graph: A graph attention network for structure based predictions of protein-protein interactions 95%
- Multi-Head Attention-based U-Nets for Predicting Protein Domain Boundaries Using 1D Sequence Features and 2D Distance Maps 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.