New distance measure for comparing protein using cellular automata image
Ferreira de Souza, L.; B. de B. Pereira, H.; M. da Rocha Filho, T.; A. S. Machado, B.; A. Moret, M.
Show abstract
One of the first steps in protein sequence analysis is comparing sequences to look for similarities. We propose an information theoretical distance to compare cellular automata representing protein sequences, and determine similarities. Our approach relies in a stationary Hamming distance for the evolution of the automata according to a properly chosen rule, and to build a pairwise similarity matrix and determine common ancestors among different species in a simpler and less computationally demanding computer codes when compared to other methods.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- SARS-CoV-2 protein structure and sequence mutations: evolutionary analysis and effects on virus variants SARS-CoV-2 protein structure and sequence mutations: 96%
- Alignment of virus-host protein-protein interaction networks by integer linear programming: SARS-CoV-2 96%
- GenomeBits insight into omicron and delta variants of coronavirus pathogen 96%
Similar papers in this journal
- The Potential of the Primitive: a Network Analysis of Early Arthropod Evolution 96%
- Comparing protein-protein interaction networks of SARS-CoV-2 and (H1N1) influenza using topological features 96%
- Role of mitochondrial genetic interactions in determining adaptation to high altitude in human population around the globe 96%
Similar papers in this journal
Similar papers in this journal
- Enrichment analysis on regulatory subspaces: a novel direction for the superior description of cellular responses to SARS-CoV-2 96%
- VICTOR: A visual analytics web application for comparing cluster sets 95%
- AE-LGBM: Sequence-Based Novel Approach To Detect Interacting Protein Pairs via Ensemble of Autoencoder and LightGBM. 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.