Extending Conformational Ensemble Prediction to Multidomain Proteins and Protein Complex
Zhu, J.; Bülow, S. v.; Liu, H.; Lindorff-Larsen, K.; Chen, H.
Show abstract
Proteins execute cellular functions through a structural continuum ranging from stable, folded domains to highly dynamic intrinsically disordered regions (IDRs). Conformational ensembles represent the set of three-dimensional structures a protein adopts under a specific set of conditions, and underlie essential processes from catalysis to complex signaling networks. While deep learning has revolutionized structure prediction, capturing the distinct conformational diversity of folded and disordered regions--especially within multidomain proteins and large assemblies--remains a fundamental challenge. Here we introduce IDPFold2, a generative framework that models the heterogenous protein thermodynamics by integrating a Mixture-of-Experts architecture into the flow matching framework. By routing residues from different regions to specialized expert networks, IDPFold2 accurately predicts conformational ensembles for folded domains, IDRs and multidomain proteins. IDPFold2 outperforms state-of-the-art methods in capturing key functional states and fitting the experimental observations across local and global scales. Furthermore, we describe an extension of IDPFold2 to protein assemblies, deciphering the complex binding modes of IDRs within large macromolecular complexes, providing a generalizable tool for exploring the dynamic proteome.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Direct prediction of intrinsically disordered protein conformational properties from sequence 98%
- Predicting structures of large protein assemblies using combinatorial assembly algorithm and AlphaFold2 97%
- US-align: Universal Structure Alignments of Proteins, Nucleic Acids, and Macromolecular Complexes 96%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.