DBSOMA: A Machine Learning Method that Identifies Chemical Modulators of Transcriptional States Uncovers Effectors of Beta-Cell Maturation
Kunz, T. R.; Rivera-Feliciano, J.
Show abstract
The effects of perturbation on a biological system can be readily measured in terms of transcriptional changes. However, despite a wealth of transcriptional perturbation response data, there are currently few methods to draw equivalence between the many biological systems used to generate that data and a specific system of interest. Here we use density analysis of transcriptional correlations to computationally predict whether a given perturbation readout is relevant to Stem Cell derived islet (SC-Islet) maturation. The approach, Density Based Self-Organizing Map Analysis (DBSOMA), first learns patterns of gene expression represented in scRNA-seq sets by clustering genes with the Self-Organizing-Map (SOM) algorithm. Perturbation expression profiles and other gene lists are then projected onto the SOM grid, where the degree of clustering is determined by the Density-Based Spatial Clustering of Applications with Noise (DBSCAN) algorithm. We applied DBSOMA to SC-Islet maturation and identified known and novel regulators of {beta}-cell maturation. This workflow can be applied broadly to biological systems where single-cell RNA-sequencing data is available, and a desired outcome can be represented in transcriptional changes.
Matching journals
The top 12 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Inferring cell diversity in single cell data using consortium-scale epigenetic data as a biological anchor for cell identity 93%
- Flexible comparison of batch correction methods for single-cell RNA-seq using BatchBench 92%
- CelLink: integrating single-cell multi-omics data with weak feature linkage and imbalanced cell populations 92%
Similar papers in this journal
- Integrating temporal single-cell gene expression modalities for trajectory inference and disease prediction 93%
- scCDC: a computational method for gene-specific contamination detection and correction in single-cell and single-nucleus RNA-seq data 93%
- Biology-inspired data-driven quality control for scientific discovery in single-cell transcriptomics 93%
Similar papers in this journal
- Projecting genetic associations through gene expression patterns highlights disease etiology and drug mechanisms 94%
- eQTL mapping in fetal-like pancreatic progenitor cells reveals early developmental insights into diabetes risk 93%
- On the discovery of population-specific state transitions from multi-sample multi-condition single-cell RNA sequencing data 93%
Similar papers in this journal
- Building, Benchmarking, and Exploring Perturbative Maps of Transcriptional and Morphological Data 93%
- GAN-Enhanced Machine Learning and Metabolic Modeling Identify Reprogramming in Pancreatic Cancer 92%
- Network models of protein phosphorylation, acetylation, and ubiquitination connect metabolic and cell signaling pathways in lung cancer 92%
Similar papers in this journal
- Chromatin accessibility differences between alpha, beta, and delta cells identifies common and cell type-specific enhancers 93%
- Probability of stealth multiplets in sample-multiplexing for droplet-based single-cell analysis 92%
- Predicting substrates for orphan Solute Carrier Proteins using multi-omics datasets 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.