TMATCH: A New Algorithm for Protein Alignments using amino-acid hydrophobicities
Cavanaugh, D. P.; Chittur, K.
Show abstract
The identification of proteins of similar structure using sequence alignment is an important problem in bioinformatics. We decribe TMATCH, a basic dynamic programming alignment algorithm which can rapidly identify proteins of similar structure from a database. TMATCH was developed to utilize an optimal hydrophobicity metric for alignments traceable to fundamental properties of amino-acids. Standard alignment algorithms use affine gap penalties as contrasted with the TMATCH algorithm adaptation of local alignment score reinforcement of favorable diagonal paths (transitions) and punishment of unfavorable transitions paired with fixed gap opening penalties. The TMATCH algorithm is especially designed to take advantage of the extra information available within the hydrophobicity scale to detect homologies, as opposed to the probabilities derived from raw percent identities.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- SARS-CoV-2 protein structure and sequence mutations: evolutionary analysis and effects on virus variants SARS-CoV-2 protein structure and sequence mutations: 96%
- GenomeBits insight into omicron and delta variants of coronavirus pathogen 96%
- Classification of protein binding ligands using structural dispersion of binding site atoms from principal axes 95%
Similar papers in this journal
Similar papers in this journal
- Beam search decoder for enhancing sequence decoding speed in single-molecule peptide sequencing data 94%
- MCell4 with BioNetGen: A Monte Carlo Simulator of Rule-Based Reaction-Diffusion Systems with Python Interface 94%
- Paying Attention to Attention: High Attention Sites as Indicators of Protein Family and Function in Language Models 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.