ActSort: An active-learning accelerated cell sorting algorithm for large-scale calcium imaging datasets
Jiang, Y.; Akengin, H. O.; Zhou, J.; Aslihak, M. A.; Li, Y.; Hernandez, O.; Ebrahimi, S.; Zhang, Y.; Inan, H.; Jaidar, O.; Miranda, C.; Dinc, F.; Blanco-Pozo, M.; Schnitzer, M. J.
Show abstract
Recent advances in calcium imaging enable simultaneous recordings of up to a million neurons in behaving animals, producing datasets of unprecedented scales. Although individual neurons and their activity traces can be extracted from these videos with automated algorithms, the results often require human curation to remove false positives, a laborious process called cell sorting. To address this challenge, we introduce ActSort, an active-learning algorithm for sorting large-scale datasets that integrates features engineered by domain experts together with data formats with minimal memory requirements. By strategically bringing outlier cell candidates near the decision boundary up for annotation, ActSort reduces human labor to about 1-3% of cell candidates and improves curation accuracy by mitigating annotator bias. To facilitate the algorithms widespread adoption among experimental neuroscientists, we created a user-friendly software and conducted a first-of-its-kind benchmarking study involving about 160,000 annotations. Our tests validated ActSorts performance across different experimental conditions and datasets from multiple animals. Overall, ActSort addresses a crucial bottleneck in processing large-scale calcium videos of neural activity and thereby facilitates systems neuroscience experiments at previously inaccessible scales.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- uniPort: a unified computational framework for single-cell data integration with optimal transport 96%
- scMODAL: A general deep learning framework for comprehensive single-cell multi-omics data alignment with feature links 96%
- Learning interpretable cellular and gene signature embeddings from single-cell transcriptomic data 96%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.