Mechanistic machine learning enables interpretable and generalizable prediction of prime editing outcomes
Hsu, A.; Chen, P. J.; Li, A. H.; Hemez, C. F.; Gao, X. D.; Terrey, M.; Nelson, C.; Selvam, V.; Cristian, A.; McElroy, A. N.; Steinbeck, B. J.; Mahadeshwar, G. K.; Pandey, S.; Barsdale, Z.; Chen, P. Z.; Sousa, A. A.; Sakai, H. A.; Silverstein, R. A.; Morad, I.; Krueger, R. K.; Shen, M. W.; Kleinstiver, B. P.; Lutz, C. M.; Tolar, J.; Blazar, B. R.; Osborne, M.; Liu, D.
Show abstract
Although prime editing (PE) can effect virtually any specified local change to genomic DNA in living systems, its efficient application currently requires extensive optimization of prime editing guide RNA (pegRNA) sequences. We present OptiPrime, a machine learning model of PE efficiency based on our current understanding of the mechanism of prime editing. OptiPrime achieves state-of-the-art accuracy on PE efficiency prediction and also enables prediction of nicking guide RNA (PE3) and dual pegRNA (twinPE) outcomes. We validated that OptiPrime has learned the determinants of mammalian mismatch repair (MMR), and is therefore well suited for nominating MMR-evasive silent edits that improve PE efficiency. We demonstrate the utility of OptiPrime in a variety of prospective therapeutic contexts, including in primary human and mouse cells. Finally, we show how OptiPrime can be used to achieve highly streamlined and efficient in vivo correction of a pathogenic mutation in the brain of a mouse model of KIF1A-associated neurological disorder.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- UPF1 mutants with intact ATPase but deficient helicase activities promote efficient nonsense-mediated mRNA decay 96%
- Cooperativity between Cas9 and hyperactive AID establishes broad and diversifying mutational footprints in base editors 96%
- Precise and efficient C-to-U RNA Base Editing with SNAP-CDAR-S 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.