Contrastive modelling of transcription and transcript abundance in legumes using PlanTT
Raymond, N.; Zhang, X.; Sheikh, J.; Daveouis, F.; Verma, R.; Cram, D.; Song, H.; Cao, Y.; Kirzinger, M.; Akaniru, D.; Ubbens, J.; Konkin, D.
Show abstract
Predicting the impacts of sequence variation on gene expression remains a challenging task. Further, in plants, we have a limited understanding of the relative contributions of different gene expression regulatory mechanisms. To address these limitations we generated a comparative multiomic dataset comprising matched 3-RNA-seq and PRO-seq data from matched tissues of reference genotypes of four legumes of the invert repeat lacking clade (Pisum sativum, Vicia faba, Lathyrus sativa and Medicago truncatula). Focused on the challenging task of predicting expression differences between ortholog pairs from unseen orthogroups, we used this dataset and a novel prediction framework to build contrastive models that predict quantitative differences (effect size differences) in transcription and transcript abundance.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- CoVar: A generalizable machine learning approach to identify the coordinated regulators driving variational gene expression 95%
- Application of Modular Response Analysis to Medium- to Large-Size Biological Systems 94%
- Improved Transcriptome Assembly Using a Hybrid of Long and Short Reads with StringTie 94%
Similar papers in this journal
- RWRtoolkit: multi-omic network analysis using random walks on multiplex networks in any species 94%
- Standardized genome-wide function prediction enables comparative functional genomics: a new application area for Gene Ontologies in plants 93%
- EssSubgraph improves performance and generalizability of mammalian essential gene prediction with large networks 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.