Inferring physical cell-cell communication networks from scRNAseq data using univariate linear models.
Hameed, S. A.; Iglesias-Martinez, L. F.; kolch, W.; Zhernovkov, V.
Show abstract
Cells in tissues interact by direct physical contact or over short and long distances via secreted mediators. Cell-cell communication inference has now become routine in downstream scRNAseq analysis but this mostly fails to capture physical cell-cell interactions due to tissue dissociation. Multiplets (mostly doublets) in scRNAseq may represent undissociated physically attached cells that become sequenced together. Hence, identifying multiplets may serve as a good starting point to harness scRNAseq data for physical cell-cell interaction inference. In this study, we develop a computational method which utilizes univariate linear models (ULM) to identify multiplets in scRNAseq datasets, predict their cellular compositions, and infer physical cell-cell interaction networks. Indeed, our method showed good sensitivity ([~]56%) with excellent precision ([~]99%) in predicting FACS sorted doublet cell pairs with known constituents (ground truth), recording comparable or superior performance over two other existing methods. Also, applying it to scRNAseq data of partially dissociated tissues containing real multiplets unraveled physical networks which recapitulated the microanatomical structures of the tested tissues. This further underscores the accuracy in our predictions to capture biologically meaningful interactions. Finally, we tested our method on classical scRNAseq datasets and obtained biologically reasonable results. For example, when tested on classical cancer scRNAseq datasets, we recovered important interactions which followed biologically plausible cell interactions, validated by cell-cell colocalization in matched spatial transcriptomics datasets. This reassured the accuracy of our method in depicting physical interactions only between cells that were truly in close proximity in tissues.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Rescue Blood Contamination in scRNA-Seq Data by Originator, a Computational Deciphering Tool using Genetic and Contextual Information 97%
- Identifying tumor cells at the single cell level 96%
- scCDC: a computational method for gene-specific contamination detection and correction in single-cell and single-nucleus RNA-seq data 96%
Similar papers in this journal
- Inferring cell diversity in single cell data using consortium-scale epigenetic data as a biological anchor for cell identity 96%
- CelLink: integrating single-cell multi-omics data with weak feature linkage and imbalanced cell populations 96%
- STAN, a computational framework for inferring spatially informed transcription factor activity across cellular contexts 96%
Similar papers in this journal
- SSMD: A semi-supervised approach for a robust cell type identification and deconvolution of mouse transcriptomics data 95%
- Scalable batch-correction method for integrating large-scale single-cell transcriptomes 95%
- StereoMM: A Graph Fusion Model for Integrating Spatial Transcriptomic Data and Pathological Images 95%
Similar papers in this journal
- STANCE: a unified statistical model to detect cell-type-specific spatially variable genes in spatial transcriptomics 96%
- FastCCC: A permutation-free framework for scalable, robust, and reference-based cell-cell communication analysis in single cell transcriptomics studies 96%
- CellScope: High-Performance Cell Atlas Workflow with Tree-Structured Representation 96%
Similar papers in this journal
- Integrative Analysis of Spatial Transcriptome with Single-cell Transcriptome and Single-cell Epigenome in Mouse Lungs after Immunization 94%
- IReNA: integrated regulatory network analysis of single-cell transcriptomes 94%
- Implicating Gene and Cell Networks Responsible for Differential COVID-19 Host Responses via an Interactive Single Cell Web Portal 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.