Hypothesis-driven interpretable neural network for interactions between genes
Wang, S.; Allauzen, A.; Nghe, P.; Opuu, V.
Show abstract
Mechanistic models of genetic interactions are rarely feasible due to a lack of information and computational challenges. Alternatively, machine learning (ML) approaches may predict gene interactions if provided with enough data but they lack interpretability. Here, we propose an ML approach for interpretable genotype-to-fitness mapping, the Direct-Latent Interpretable Model (D-LIM). The neural network is built on a strong hypothesis: mutations in different genes cause independent effects in phenotypes, which then interact via non-linear relationships to determine fitness. D-LIM predicts interpretable genotype-to-fitness maps with state-of-the-art accuracy for gene-to-gene and gene-to-environment perturbations in deep mutational scanning of a metabolic pathway, a protein-protein interaction system, and yeast mutants for environmental adaptation. The hypothesis-driven structure of D-LIM offers interpretable features reminiscent of mechanistic models: the inference of phenotypes, identification of trade-offs, and fitness extrapolation outside of the data domain.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- An extension of the Walsh-Hadamard transform to calculate and model epistasis in genetic landscapes of arbitrary shape and complexity 96%
- Parallel Tempering with Lasso for Model Reduction in Systems Biology 96%
- Inferring gene regulatory networks using transcriptional profiles as dynamical attractors 96%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.