Back

Genetic drivers and clinical consequences of mosaic chromosomal alterations in 1 million individuals

Zhao, K.; Pershad, Y.; Poisner, H. M.; Ma, X.; Quade, K.; Vlasschaert, C.; Mack, T.; Khankari, N. K.; von Beck, K.; Brogan, J.; Kishtagari, A.; Corty, R.; Li, Y.; Xu, Y.; Reiner, A. P.; Scheet, P.; Auer, P.; Bick, A. G.

2025-03-06 genetic and genomic medicine
10.1101/2025.03.05.25323443 medRxiv
Show abstract

Mosaic chromosomal alterations of the autosomes (aut-mCAs) are large structural somatic mutations which cause clonal hematopoiesis and increase cancer risk. Here, we detected aut-mCAs in 1,011,269 participants across four biobanks. Through integrative analysis of the minimum critical region and inherited genetic variation, we found that proto-oncogenes exclusively drive chromosomal gains, tumor suppressors drive losses, and copy-neutral events can be driven by either. We identified three novel inherited risk loci in CHI3L2, HLA class II, and TERT that modulate aut-mCA risk and ten novel aut-mCA-specific loci. We found specific aut-mCAs are associated with cardiovascular, cerebrovascular, or kidney disease incidence. High-risk aut-mCAs were associated with elevated plasma protein levels of therapeutically actionable targets: NPM1, PARP1, and TACI. Participants with multiple high-risk features such as high clonal fraction, more than one aut-mCA, and abnormal red cell morphology had a 50% cumulative incidence of blood count abnormalities over 2 years. Leveraging inherited variation, we causally established aut-mCAs as premalignant lesions for chronic lymphocytic leukemia. Together, our findings provide a framework integrating somatic mosaicism, germline genetics, and clinical phenotypes to identify individuals who could benefit from preventative interventions.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.