BULLKpy: An AnnData-Inspired Unified Framework for Comprehensive Bulk OMICs Analysis
Malumbres, M.
Show abstract
Bulk OMICs data, such as RNA-seq and proteomics, remain foundational in biomedical and cancer research. While single-cell transcriptomics has revolutionized our understanding of cellular heterogeneity, bulk OMICs continue to provide a solid ground for large-scale studies, integrative analyses, and robust biomedical correlations. Yet, analytical workflows for bulk OMICs are frequently fragmented across disparate tools, programming languages, and data formats, in contrast to the single-cell field, which benefits from comprehensive and standardized frameworks like Scanpy and the broader scverse ecosystem. Here, we introduce BULLKpy, a Python-based, scverse-inspired framework designed to deliver similar integration, flexibility, and scalability to bulk RNA-seq analysis, with a particular emphasis on cancer research. Beyond standard preprocessing and visualization utilities, BULLKpy offers systematic evaluation of genemetadata associations, categorical enrichment analyses, clustering stability metrics, and advanced visualization strategies that facilitate intuitive exploration of tumor heterogeneity in expression profiles and metaprograms. Cancer-specific visualizations, such as oncoprints, are natively supported and seamlessly integrated with transcriptomic and clinical features, allowing users to relate somatic alterations to expression-derived phenotypes. By standardizing workflows within the Python ecosystem and aligning with AnnData objects and the scverse, BULLKpy aims to democratize advanced transcriptomic and proteomic analyses, facilitate integrative cancer research, and support future developments at the intersection of computational biology, multi-omics integration, and artificial intelligence.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Xenomake: a pipeline for processing and sorting xenograft reads from spatial transcriptomic experiments 95%
- SpatialRNA: a Python package for easy application of Graph Neural Network models on single-molecule spatial transcriptomicsdataset 95%
- Scalable Integration of Multiomic Single Cell Data Using Generative Adversarial Networks 95%
Similar papers in this journal
- Neighborhood nonnegative matrix factorization identifies patterns and spatially-variable genes in large-scale spatial transcriptomics data 95%
- Sfaira accelerates data and model reuse in single cell genomics 95%
- Giotto, a toolbox for integrative analysis and visualization of spatial expression data 95%
Similar papers in this journal
- DeepSpaceDB: a spatial transcriptomics atlas for interactive in-depth analysis of tissues and tissue microenvironments 95%
- Interpretable deep learning for chromatin-informed inference of transcriptional programs driven by somatic alterations across cancers 95%
- Multi-resolution characterization of molecular taxonomies in bulk and single-cell transcriptomics data 94%
Similar papers in this journal
- stLearn: integrating spatial location, tissue morphology and gene expression to find cell types, cell-cell interactions and spatial trajectories within undissociated tissues 95%
- Scarf: A toolkit for memory efficient analysis of large-scale single-cell genomics data 94%
- Identification of Relevant Genetic Alterations in Cancer using Topological Data Analysis 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.