Lightweight Fine-tuning a Pretrained Protein Language Model for Protein Secondary Structure Prediction
Yang, W.; Liu, C.; Li, Z.
Show abstract
Pretrained large-scale protein language models, such as ESM-1b and ProtTrans, are becoming the fundamental infrastructure for various protein-related biological modeling tasks. Existing works use mainly pretrained protein language models in feature extraction. However, the knowledge contained in the embedding features directly extracted from a pretrained model is task-agnostic. To obtain task-specific feature representations, a reasonable approach is to fine-tune a pretrained model based on labeled datasets from downstream tasks. To this end, we investigate the fine-tuning of a given pretrained protein language model for protein secondary structure prediction tasks. Specifically, we propose a novel end-to-end protein secondary structure prediction framework involving the lightweight fine-tuning of a pretrained model. The framework first introduces a few new parameters for each transformer block in the pretrained model, then updates only the newly introduced parameters, and then keeps the original pretrained parameters fixed during training. Extensive experiments on seven test sets, namely, CASP12, CASP13, CASP14, CB433, CB634, TEST2016, and TEST2018, show that the proposed framework outperforms existing predictors and achieves new state-of-the-art prediction performance. Furthermore, we also experimentally demonstrate that lightweight fine-tuning significantly outperforms full model fine-tuning and feature extraction in enabling models to predict secondary structures. Further analysis indicates that only a few top transformer blocks need to introduce new parameters, while skipping many lower transformer blocks has little impact on the prediction accuracy of secondary structures.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Improved model quality assessment using sequence and structural information by enhanced deep neural networks 97%
- SPDesign: protein sequence designer based on structural sequence profile using ultrafast shape recognition 97%
- GraphGPSM: a global scoring model for protein structure using graph neural networks 97%
Similar papers in this journal
- DeepUMQA: Ultrafast Shape Recognition-based Protein Model Quality Assessment using Deep Learning 97%
- Combining protein sequences and structures with transformers and equivariant graph neural networks to predict protein function 97%
- Identifying B-cell epitopes using AlphaFold2 predicted structures and pretrained language model 97%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.