Dynamically Assembling Biological Intelligence to Predict Novel Cellular Phenotypes
Pu, H.; Long, Y.
Show abstract
In this work, we introduce Bio-AMLM (Biological Adaptive Modular Learning Model), a new framework designed to address out-of-distribution (OOD) generalization challenges in predicting cellular responses. Unlike monolithic deep learning models or simple data retrieval methods, which struggle to predict the effects of novel genetic or chemical perturbations, Bio-AMLM dynamically constructs a bespoke analytical pipeline for each biological query. It leverages a library of pre-trained, functionally specialized biological modules (e.g., for genomic, proteomic, and metabolic analysis). Guided by a biological context encoder, an adaptive inference planner selects, configures, and links these modules to form an optimal analysis chain. In experiments on several challenging bio-simulation benchmarks, including Gene-Edit-Bench, Drug-Response-Bench, and Toxicity-Bench, Bio-AMLM consistently outperformed state-of-the-art approaches, producing more reliable, robust, and interpretable predictions of cellular behavior in complex OOD scenarios.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- PandoGen: Generating complete instances of future SARS-CoV-2 sequences using Deep Learning 96%
- Learning Genetic Perturbation Effects with Variational Causal Inference 95%
- Highly Accurate Cancer Phenotype Prediction with AKLIMATE, a Stacked Kernel Learner Integrating Multimodal Genomic Data and Pathway Knowledge 95%
Similar papers in this journal
- seqgra: Principled Selection of Neural Network Architectures for Genomics Prediction Tasks 96%
- Graph Convolutional Networks for Epigenetic State Prediction Using Both Sequence and 3D Genome Data 95%
- Consensus Label Propagation with Graph Convolutional Networks for Single-Cell RNA Sequencing Cell Type Annotation 95%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.