Enformation Theory: A Framework for Evaluating Genomic AI
Robson, E. S.; Ioannidis, N. M.
Show abstract
The nascent field of genomic AI is rapidly expanding with new models, benchmarks, and findings. As the field diversifies, there is an increased need for a common set of measurement tools and perspectives to standardize model evaluation. Here, we present a statistically grounded framework for performance evaluation, visualization, and interpretation using the prominent sequence-based deep learning models Enformer and Borzoi as case studies. The Enformer model has been used for applications ranging from understanding regulatory mechanisms to variant effect prediction, but what makes it better or worse than precedent models? Does its follow-up, Borzoi, offer improved performance and more informative embeddings as well as finer resolution? Our goal is to propose a general blueprint for answering such questions and evaluating new models. We start by contrasting the few-shot performance of Enformer and Borzoi to precedent models on the GUANinE benchmark, which emphasizes complex genome interpretation tasks. We then examine Enformer and Borzoi intermediate embeddings in model-subjective principal component space, where we identify limiting aspects that affect model generalization. Finally, we present an interpretable decomposition of Enformer and Borzoi, which allows for global model interpretability and partial backtracking to explanatory causal features. Through this case study, we illustrate a new protocol, Enformation Theory, for analyzing and interpreting deep learning models in genomics.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- PandoGen: Generating complete instances of future SARS-CoV-2 sequences using Deep Learning 96%
- Iterative point set registration for aligning scRNA-seq data 96%
- Highly Accurate Cancer Phenotype Prediction with AKLIMATE, a Stacked Kernel Learner Integrating Multimodal Genomic Data and Pathway Knowledge 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.