Back

scAdapt: Virtual adversarial domain adaptation network for single cell RNA-seq data classification across platforms and species

Zhou, X.; Chai, H.; Zeng, Y.; Zhao, H.; Luo, C.-H.; Yang, Y.

2021-01-19 bioinformatics
10.1101/2021.01.18.427083 bioRxiv
Show abstract

MotivationIn single cell analyses, cell types are conventionally identified based on known marker gene expressions. Such approaches are time-consuming and irreproducible. Therefore, many new supervised methods have been developed to identify cell types for target datasets using the rapid accumulation of public datasets. However, these approaches are sensitive to batch effects or biological variations since the data distributions are different in cross-platforms or species predictions. ResultsWe developed scAdapt, a virtual adversarial domain adaptation network to transfer cell labels between datasets with batch effects. scAdapt used both the labeled source and unlabeled target data to train an enhanced classifier, and aligned the labeled source centroid and pseudo-labeled target centroid to generate a joint embedding. We demonstrate that scAdapt outperforms existing methods for classification in simulated, cross-platforms, cross-species, and spatial transcriptomic datasets. Further quantitative evaluations and visualizations for the aligned embeddings confirm the superiority in cell mixing and preserving discriminative cluster structure present in the original datasets. Availabilityhttps://github.com/zhoux85/scAdapt. Contactangyd25@mail.sysu.edu.cn or luojinx5@mail.sysu.edu.cn

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.