A C++ library for protein sub-structure search
Zhou, J.; Grigoryan, G.
Show abstract
SummaryMASTER is a previously published algorithm for protein sub-structure search. Given a database of protein structures and a query structural motif, composed of multiple disjoint segments, it finds all sub-structures from the database that align onto the query to within a pre-specified backbone root-mean-square deviation. Here, we present an improved version of the algorithm, MASTER v.2, in the form of an open-source C++ Application Program Interface library, thereby providing programmatic access to structure search functionality. An entirely reorganized approach to database representation now enables large structural databases to be stored in memory, further simplifying development of automated search-based methods. Given the increasingly important role of structure-based data mining, our improved implementation should find ample uses in structural biology applications. AvailabilityMASTER is available at https://grigoryanlab.org/master/master-v2.php. Contactgevorg.grigoryan@dartmouth.edu
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- diverse-seq: an application for alignment-free selecting and clustering biological sequences 93%
- dms-viz: Structure-informed visualizations for deep mutational scanning and other mutation-based datasets 92%
- RedOak: a reference-free and alignment-freestructure for indexing a collection of similargenomes 92%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Prop3D: A Flexible, Python-based Platform for Machine Learning with Protein Structural Properties and Biophysical Data 95%
- Global, Highly Specific and Fast Filtering of Alignment Seeds 94%
- CoGAPS 3: Bayesian non-negative matrix factorization for single-cell analysis with asynchronous updates and sparse data structures 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.