k-Nearest Neighbour Adaptive Sampling (kNN-AS), a Simple Tool to Efficiently Explore Conformational Space
Rovers, E. M.; Thudi, A.; Maddison, C.; Schapira, M.
Show abstract
Molecular dynamics (MD) simulations are computationally expensive, a limiting factor when simulating biomolecular systems. Adaptive sampling approaches can accelerate the exploration of conformational space by running repeated short MD simulations from well-chosen starting points. Existing approaches to adaptive sampling have been optimized to either guide sampling in a desired direction or explore well-formed convex spaces. Here, we describe a novel adaptive sampling algorithm that leverages a k-nearest neighbour (k-NN) graph of the sampled conformational space to preferentially launch explorations from boundary states. We term this approach k-Nearest Neighbor Adaptive Sampling (kNN-AS) and show it has state-of-the-art performance on simple and complex artificial energy functions and generalizes well on a protein test case. Implementation of kNN-AS is light, simple and better suited to complex real-world applications where the dimension and shape of the energy landscape is unknown.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A linear response theory based method for prediction of large scale protein conformational changes upon ligand binding 97%
- Enhancing Sampling of Water Rehydration on Ligand Binding: A Comparison of Techniques 97%
- ANUBI: A Platform for Affinity Optimization of Proteins and Peptides in Drug Design 97%
Similar papers in this journal
- Smart Distributed Data Factory: Volunteer Computing Platform for Active Learning-Driven Molecular Data Acquisition 95%
- Telomeric G-quadruplex Intermediates unveiled by Complex Markov Network Analysis. 95%
- High pressure inhibits signaling protein binding to the flagellar motor and bacterial chemotaxis through enhanced hydration 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.