Back

Active subspace learning for coarse-grained molecular dynamics

Wojnar, A.; Pankavich, S.; Pak, A. J.

2025-12-14 biophysics
10.1101/2025.10.13.682174 bioRxiv
Show abstract

We introduce Active Subspace Coarse-Graining (ASCG), an interpretable framework for systematic bottom-up coarse-graining trained from atomistic molecular dynamics simulations that simultaneously defines the coarse-grained mapping, effe ctive interactions, and the equations of motion within one unified mathematical framework. We employ active subspace learning to identify linear projections of atomistic degrees of freedom that maximally describe gradients of the potential energy, yielding a reduced set of coarse-grained variables that capture the dominant collective motions across the potential of mean force. Effective coarse-grained forces and noise terms are obtained directly from the same projection, eliminating the need for separate parameterization schemes. We demonstrate the ASCG method on three biomolecules: dialanine, Trp-cage, and chignolin. We show that free energy surfaces are recapitulated with Jensen-Shannon divergences as low as 0.034 while eliminating all solvent degrees of freedom and reducing solute dimensionality by more than 90%. The ASCG trajectories are integrated with timesteps up to 100 fs, around four to ten times larger than those possible with conventional coarse-graining methods, while ASCG models remain accurate with as little as 100 ns of training data. These results establish ASCG as a robust, data-efficient approach for learning complete coarse-grained representations directly from molecular forces, while representing a departure from traditional particle-based models.

Published in The Journal of Chemical Physics (predicted rank #3) · training set

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.