Back

Pharming: Joint Clonal Tree Reconstruction of SNV and CNAEvolution from Single-cell DNA Sequencing of Tumors

Weber, L. L.; Hart, A.; Ochoa, I.; El-Kebir, M.

2024-11-18 bioinformatics
10.1101/2024.11.17.623950 bioRxiv
Show abstract

Cancer arises through an evolutionary process in which somatic mutations, including single nucleotide variants (SNVs) and copy number aberrations (CNAs), drive the development of a malignant, heterogeneous tumor. Reconstructing this evolutionary history from sequencing data is critical for understanding the order in which mutations are acquired and the dynamic interplay between different types of alterations. Advances in modern whole genome single-cell sequencing now enable the accurate inference of copy number profiles in individual cells. However, the low sequencing coverage of these low pass sequencing technologies poses a challenge for reliably inferring the presence or absence of SNVs within tumor cells, limiting the ability to simultaneously study the evolutionary relationships between SNVs and CNAs. In this work, we introduce a novel tumor phylogeny inference method, PO_SCPLOWHARMINGC_SCPLOW, that jointly infers the evolutionary histories of SNVs and CNAs. Our key insight is to leverage the high accuracy of copy number inference methods and the fact that SNVs co-occur in regions with CNAs in order to enable more precise tumor phylogeny reconstruction for both alteration types. We demonstrate via simulations that PO_SCPLOWHARMINGC_SCPLOW outperforms state-of-the-art single-modality tumor phylogeny inference methods. Additionally, we apply PO_SCPLOWHARMINGC_SCPLOW to a triple-negative breast cancer case, achieving high-resolution, joint reconstruction of CNA and SNV evolution, including the de novo detection of a clonal whole-genome duplication event. Thus, PO_SCPLOWHARMINGC_SCPLOW offers the potential for more comprehensive and detailed tumor phylogeny inference for high-throughput, low-coverage single-cell DNA sequencing technologies compared to existing approaches. Availabilityhttps://github.com/elkebir-group/Pharming

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.