RNASTOP: A Deep Learning Framework for mRNA Chemical Stability Prediction and Optimization
Lin, S.; Chen, J.; Sun, H.; Zhang, Y.; Yang, W.; tan, h.; Wei, D.-Q.; Jiang, Q.; Xiong, Y.
Show abstract
Messenger RNA (mRNA) vaccines offer promising therapeutics for combating various diseases, yet their inherent chemical instability hampers their long-term efficacy. Although several methods have been developed to predict mRNA degradation, they exhibit limited accuracy and lack the capability for rational sequence optimization. Here, we propose RNASTOP, a novel framework integrating deep learning with heuristic search to simultaneously predict and optimize mRNA chemical stability. RNASTOP achieves a 13% accuracy improvement over the top-performing model on the Stanford OpenVaccine competition dataset and demonstrates robust generalization in predicting full-length mRNA degradation. Applied to mRNA codon optimization, RNASTOP reduces the minimum free energy of the Varicella-Zoster Virus vaccine sequence by 75.73% while maintaining high translation efficiency. Overall, RNASTOP serves as a powerful tool for predicting and optimizing mRNA chemical stability, poised to expedite the development of mRNA therapeutics. The source code of RNASTOP can be accessed at https://github.com/xlab-BioAI/RNASTOP.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- EternaBrain: Automated RNA design through move sets from an Internet-scale RNA videogame 94%
- THLANet: A Deep Learning Framework for Predicting TCR-pHLA Binding in Immunotherapy Applications 93%
- Improving deep models of protein-coding potential with a Fourier-transform architecture and machine translation task 93%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.