Pi-Ensemble: Sequence-guided generation of interpolated protein conformational ensembles
Nadeem, H.; Kleiman, D. E.; Zhou, Y.; Leakey, A. D. B.; Shukla, D.
Show abstract
Proteins are critical biomolecular machines that populate ensembles of interconverting conformations. Many biological processes depend on transitions between metastable states. Although molecular dynamics (MD) simulations provide a physically grounded route to characterize these motions, routine sampling of large-scale conformational transitions remains computationally demanding. Recent advances in protein structure prediction have created new opportunities for ensemble generation, but many existing approaches require noising inputs, task-specific training, supervised fitting on extensive MD data, or experimentally-informed restraints. Here, we introduce Pi-Ensemble (Predicting Interpolated Ensemble), a sequence-guided framework for generating protein conformational ensembles interpolating between two structural anchor states. Unlike previous methods, Pi-Ensemble alternately leverages inverse-folding and structure-prediction models to propose intermediate conformations between known protein states, generating diverse ensembles without additional training. We evaluate Pi-Ensemble across diverse protein systems, including enzymes, transporters, receptors, and benchmark cases with reference MD simulations or experimental Double Electron-Electron Resonance (DEER) data. Pi-Ensemble recovers physically plausible intermediate conformations, captures transition pathways observed in large-scale MD simulations, and generates structures consistent with experimental distance distributions. Furthermore, Pi-Ensemble-generated conformations provide effective starting seeds for parallel MD simulations, improving conformational exploration and accelerating convergence relative to simulations initiated only from endpoint structures. These results establish sequence-guided structural interpolation as a practical strategy for probing protein conformational landscapes. By generating diverse and physically reasonable conformational proposals without long-timescale MD or model retraining, Pi-Ensemble provides an extensible framework for studying protein flexibility, guiding adaptive sampling, and accelerating mechanistic investigations of protein function.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Distance-Restraint-Guided Diffusion Models for Sampling Protein Conformational Changes and Ligand Dissociation Pathways 98%
- MELD-Adapt: On-the-Fly Belief Updating in Integrative Molecular Dynamics 95%
- Toward Accurate RNA Folding Thermodynamics: Evaluation of Enhanced Sampling Methods for Force Field Benchmarking 95%
Similar papers in this journal
- ConforFold Recovers Alternative Protein Conformations Beyond MSA Subsampling 97%
- Assessing the relation between protein phosphorylation, AlphaFold3 models and conformational variability 95%
- Neural Network-Derived Potts Models for Structure-Based Protein Design using Backbone Atomic Coordinates and Tertiary Motifs 94%
Similar papers in this journal
Similar papers in this journal
- A general-purpose protein design framework based on mining sequence-structure relationships in known protein structures 96%
- De Novo Protein Fold Design Through Sequence-Independent Fragment Assembly Simulations 95%
- Folding-upon-binding pathways of an intrinsically disordered protein from a deep Markov state model 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.