LightMHC: A Light Model for pMHC Structure Prediction with Graph Neural Networks
Delaunay, A. P.; Fu, Y.; Gorbushin, N.; McHardy, R.; Djermani, B. A.; Copoiu, L.; Rooney, M.; Lang, M.; Tovchigrechko, A.; Sahin, U.; Beguir, K.; Lopez Carranza, N.
Show abstract
The peptide-major histocompatibility complex (pMHC) is a crucial protein in cell-mediated immune recognition and response. Accurate structure prediction is potentially beneficial for protein interaction prediction and therefore helps immunotherapy design. However, predicting these structures is challenging due to the sequential and structural variability. In addition, existing pre-trained models such as AlphaFold 2 require expensive computation thus inhibiting high throughput in silico peptide screening. In this study, we propose LightMHC: a lightweight model (2.2M parameters) equipped with attention mechanisms, graph neural networks, and convolutional neural networks. LightMHC predicts full-atom pMHC structures from amino-acid sequences alone, without template structures. The model achieved comparable or superior performance to AlphaFold 2 and ESMFold (93M and 15B parameters respectively), with five-fold acceleration (6.65 seconds/sample for LightMHC versus 36.82 seconds/sample for AlphaFold 2), potentially offering a valuable tool for immune protein structure prediction and immunotherapy design.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Guiding a language-model based protein design method towards MHC Class-I immune-visibility profiles for vaccines and therapeutics 95%
- Structural pre-training improves physical accuracy of antibody structure prediction using deep learning. 94%
- Epitopedia: identifying molecular mimicry between pathogens and known immune epitopes 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.