Back

fastdemux: Robust SNP-based demultiplexing of single-cell population genomics data

Ranjbaran, A.; Luca, F.; Pique-Regi, R.

2026-02-11 bioinformatics
10.64898/2026.02.10.705082 bioRxiv
Show abstract

Sample multiplexing reduces cost and batch effects in population based large-scale single-cell genomics studies but requires accurate and scalable computational demultiplexing. Existing genotype-based methods, such as demuxlet, provide high accuracy but can be computationally slow and memory intensive as the number of cells, donors, and informative variants increases. Here, we introduce fastdemux, a scalable genotype-based demultiplexing framework based on a diagonal linear discriminant analysis (DLDA) model that substantially improves computational efficiency while maintaining accurate donor assignment. Using a pooled single-cell RNA-seq dataset from unrelated donors, we benchmarked fastdemux against demuxlet, vireo, and demuxalot.fastdemux achieved comparable or improved demultiplexing accuracy while reducing runtime and peak memory usage by orders of magnitude relative to alternative methods. Performance remained robust across varying sequencing depths and genotype SNP filtering thresholds. In addition, the DLDA framework naturally extends to doublet and higher-order multiplet detection. We also show thatfastdemux works well with scATAC-seq data where genetic variants are more sparsely covered. Together, these results establish fastdemux as an efficient and scalable solution for genetic demultiplexing of pooled single-cell datasets.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.