Ab-VS: Evaluating Large Language Models for Virtual Antibody Screening via Antibody-Antigen Interaction Prediction
Zhang, Y.; Tsuda, K.
Show abstract
We present Ab-VS-Bench, a new benchmark for evaluating large language models (LLMs) on antibody virtual screening (VS) tasks through natural language instructions. Unlike prior efforts that focus on structural annotation, binding prediction, or developability using frozen antibody models, Ab-VS-Bench targets the end-to-end VS workflow and leverages LLMs instruction-following capabilities. The benchmark comprises three core tasks inspired by small-molecule VS: (1) Scoring--predicting antibody-antigen binding affinity, (2) Ranking--ordering antibodies by affinity or thermostability, and (3) Screening--identifying high-affinity binders from large antibody libraries. We convert experimental datasets into instruction-tuning format and evaluate multiple model variants, including zero-shot LLMs, instruction-finetuned LLMs, and multimodal models enhanced with antibody-specific embeddings. Ab-VS-Bench provides a unified framework to benchmark LLMs for antibody discovery, aiming to catalyze progress in language-guided biotherapeutic design.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Humatch - fast, gene-specific joint humanisation of antibody heavy and light chains 94%
- AlphaBind, a Domain-Specific Model to Predict and Optimize Antibody-Antigen Binding Affinity 94%
- BioPhi: A platform for antibody design, humanization and humanness evaluation based on natural antibody repertoires and deep learning 94%
Similar papers in this journal
- Tokenized and Continuous Embedding Compressions of Protein Sequence and Structure 94%
- Generating hard-to-obtain information from easy-to-obtain information: applications in drug discovery and clinical inference 93%
- Focused learning by antibody language models using preferential masking of non-templated regions 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.