PhyGraFT: a network-based method for phylogenetic trait analysis
Matsumoto, H.; Matsui, M.
Show abstract
With the determination of numerous viral and bacterial genome sequences, phylogeny-trait associations are now being studied. In these studies, phylogenetic trees were first reconstructed, and trait data were analyzed based on the reconstructed tree. However, in some cases, such as fast evolution sequences and gene-sharing network data, reconstructing the phylogenetic tree is challenging. In such cases, network-thinking, instead of tree-thinking, is gaining attention. Here, we propose a novel network-thinking approach, PhyGraFT, to analyze trait data from the network. We validated that PhyGraFT can find phylogenetic signals and associations of traits with the simulation dataset. We applied PhyGraFT for influenza type A and virome gene-sharing datasets. As a result, we identified several evolutionary structures and their associated traits. Our approach is expected to provide novel insights into network-thinking not only for typical phylogenetics but also for various biological data, such as antibody evolution.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- GCNCDA: A New Method for Predicting CircRNA-Disease Associations Based on Graph Convolutional Network Algorithm 94%
- DNFE: Directed-network flow entropy for detecting the tipping points during biological processes 94%
- Extended Graphical Lasso for Multiple Interaction Networks for High Dimensional Omics Data 94%
Similar papers in this journal
- LSTM-PHV: Prediction of human-virus protein-protein interactions by LSTM with word2vec 94%
- KGETCDA: an efficient representation learning framework based on knowledge graph encoder from transformer for predicting circRNA-disease associations 94%
- A Reproducibility Analysis-based Statistical Framework for Residue-Residue Evolutionary Coupling Detection 94%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.