Back

Modeling mitochondrial inheritance enables high-precision single-cell lineage tracing in humans

Gao, T.; Weng, C.; Johnson, I.; Poeschla, M.; Gudera, J.; King, E.; Rouya, C.; Donovan, A.; Bourke, L.; Shao, Y.; Marquez, E.; Tyag, R.; Zon, L. I.; Weissman, J. S.; Sankaran, V. G.

2026-02-13 genomics
10.64898/2026.02.12.705660 bioRxiv
Show abstract

Somatic mutations in mitochondrial DNA (mtDNA) provide natural barcodes that enable engineering-free lineage tracing in human tissues, but the complex dynamics of mtDNA inheritance across cell divisions and incomplete sampling of mtDNA introduce uncertainty in reconstructed lineages. Here, we present MitoDrift, a probabilistic framework that integrates Wright-Fisher drift dynamics with sparse single-cell measurements to produce confidence-refined lineage trees enriched for accurate clonal relationships. Validation with gold-standard lentiviral barcoding and whole-genome sequencing demonstrates that MitoDrift outperforms existing tree reconstruction methods in precision while maintaining high clonal recovery, enabling robust analyses linking lineage to cell state. Applying MitoDrift to human hematopoiesis reveals an age-associated decline in clonal diversity with differential impact across cell types and identifies heritable regulatory programs in hematopoietic stem cells in vivo, linking AP-1/stress-associated programs to clonal expansions. In multiple myeloma, MitoDrift captures therapy-associated clonal remodeling undetectable by copy number analysis, revealing phenotypic transitions and linking gene regulatory programs to differential drug sensitivity. Collectively, MitoDrift enables high-precision lineage tracing at scale and establishes quantitative lineage-state analysis in primary human tissues, linking clonal history to transcriptional and epigenetic programs in tissue homeostasis, aging, and disease.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.