Back

Comprehensively Testing the Function of Missense Variation in the STK11 Tumour Suppressor

Zimmerman, D.; Cote, A.; van Loggerenberg, W.; Gebbia, M.; Kishore, N.; Weile, J.; Li, R.; Reno, C.; Marsh, A.; Hernandez, F.; Shahagadkar, P.; Grove, L.; Meier, S.; Wu, H.-J.; Fengolia, S.; Ahronian, L.; Teng, T.; Waters, A. J.; Seward, D.; Taipale, M.; Aronson, M.; Richardson, M. E.; Adams, D.; Roth, F.

2025-07-18 cancer biology
10.1101/2025.07.14.664734 bioRxiv
Show abstract

The tumor suppressor gene STK11 encoding Serine/Threonine Kinase 11 (STK11) is associated with Peutz-Jeghers Syndrome (PJS), a heritable gastrointestinal disease that increases lifetime cancer risk, and with somatic variation that contributes to [~]30% of lung and 20% of cervical cancers. Although identifying pathogenic variants is clinically actionable, over 94% of STK11 missense variants that have been observed clinically lack a definitive classification. We therefore measured the impact of STK11 variants at scale in a mammalian cell-based assay, scoring 6,026 (73% of all possible) amino acid substitutions across the full-length gene. Functional scores--which were consistent with biochemical properties, smaller-scale assays, and pathogenicity annotations--identified a subset of PJS patients with germline STK11 variants diagnosed later in life, as well as somatic STK11 variants found in cancer patients that had comparable overall survival estimates to wild-type STK11. Our scores provided new evidence for 350 annotated VUS STK11 missense variants and [~]80% of missense variants that have not yet been reported clinically, but we might expect to observe in the future. Thus, our effect map provides a proactive resource for gaining sequence-structure-function insights and evidence for actionable interpretation of clinical missense variants.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.