Ingres: from single-cell RNA-seq data to single-cell probabilistic Boolean networks
Victori, P.; Buffa, F. M.
Show abstract
MotivationThe current explosion of omics data has provided scientists with an unique opportunity to elucidate the inner workings of biological processes that remained opaque. For this, computational models are essential. Gene regulatory networks (GRN) have long been used as a way to integrate heterogeneous data into a discrete model, and are very useful to generate actionable hypotheses on the mechanisms governing these biological processes. Boolean networks are particularly popular for this kind of discrete models. When working with single-cell RNA-seq datasets a main focus of analysis is the differential expression between subpopulations of cells. Boolean networks are limited in this task, since they cannot easily represent different levels of expression, confined as they are to binary states. We set out to develop an algorithm that can fit Boolean networks with this kind of data, maintaining both the heterogeneity of the data and the simplicity and computational efficiency of these type of networks. ResultsHere we present Ingres (Inferring Probabilistic Boolean Networks of Gene Regulation Using Protein Activity Enrichment Scores) an open-source tool that uses single-cell sequencing data and prior knowledge GRNs to produce a probabilistic Boolean network (PBN) per each cell and/or cluster of cells in the dataset. Ingres allows to better capture the differences between cell phenotypes, using a continuous measure of protein activity while still confined to the simplicity of a GRN. We believe Ingres will be useful to better understand the heterogeneous makeup of cell populations, to gain insight into the specific circuits that drive certain phenotypes, and to use expression and other omics to infer computational cellular models in bulk or single-cell data. Availability and implementationIngres has been implemented as an R package, and it is publicly available at https://github.com/CBigOxf/ingres. It is currently being submitted to the public repository CRAN too. works seamlessly with existing software for single-cell RNA-seq analysis, and for network analysis, modelling and visualization. Contactpedro.victori@oncology.ox.ac.uk; francesca.buffa@unibocconi.it; francesca.buffa@imm.ox.ac.uk Supplementary informationSupplementary data are available online. Software documentation and an explanatory vignette are available on GitHub and as part of the R package.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Beyond synthetic lethality in large-scale metabolic and regulatory network models via genetic minimal intervention sets 95%
- scAnnotate: an automated cell type annotation tool for single-cell RNA-sequencing data 94%
- GeneSNAKE: a Python package for benchmarking and simulation of gene regulatory networks and perturbation-induced expression data 94%
Similar papers in this journal
- Benchmarking imputation methods for network inference using a novel method of synthetic scRNA-seq data generation 94%
- HARVESTMAN: A framework for hierarchical featurelearning and selection from whole genome sequencingdata 94%
- CoGAPS 3: Bayesian non-negative matrix factorization for single-cell analysis with asynchronous updates and sparse data structures 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.