Back

AlphaFold3 and Intrinsically Disordered Proteins: Reliable Monomer Prediction, Unpredictable Multimer Performance

Dao, T. M.; Ghent, S.; Uversky, V. N.; Rahman, T.

2025-12-10 bioinformatics
10.64898/2025.12.05.691730 bioRxiv
Show abstract

AlphaFold3 represents a major advance in protein structure prediction, yet its performance on intrinsically disordered proteins remains uncharacterized. We present the first systematic evaluation of AF3 on disordered systems, revealing a striking dichotomy. For monomers, AF3s pLDDT scores reliably predict disorder (MCC: 0.693), matching AlphaFold2 and rivaling dedicated predictors. This consistency across fundamentally different architectures confirms that disorder prediction emerges from training data, not model design. For multimers, the picture grows complex. Despite comparable aggregate performance (mean DockQ: 0.563 vs 0.571), AF3 and AF2 achieve these results through fundamentally different mechanisms. Conventional structural features explain 58% of AF2s variance but only 42% of AF3s. Users cannot predict when AF3 will succeed or fail from interface properties alone. On disorder-to-order transitions (MFIB benchmark), both models perform equally well, successfully predicting final folded states. Yet seed variance analysis reveals AF3s failures are deterministic: the model converges to identical structures across independent runs, whether correct or incorrect, indicating rigid structural priors override available information. Our findings establish AF3 as reliable for the prediction of monomer disorder but unpredictable for multimers. Architectural innovation alone cannot overcome training data bias. Progress demands disorder-enriched datasets and ensemble sampling, not merely novel architectures.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
Protein Science
246 papers in training set
Top 0.1%
38.0%
2
Bioinformatics
1204 papers in training set
Top 2%
12.0%
3
Bioinformatics Advances
203 papers in training set
Top 0.5%
7.0%
50% of probability mass above
4
Structure
193 papers in training set
Top 0.4%
5.0%
5
Journal of Molecular Biology
232 papers in training set
Top 0.5%
4.7%
6
Computational and Structural Biotechnology Journal
242 papers in training set
Top 1%
3.9%
7
PLOS Computational Biology
1863 papers in training set
Top 10%
3.1%
8
Proteins: Structure, Function, and Bioinformatics
88 papers in training set
Top 0.4%
3.1%
9
Nature Communications
5641 papers in training set
Top 37%
3.1%
10
Briefings in Bioinformatics
354 papers in training set
Top 4%
1.8%
11
Cell Systems
201 papers in training set
Top 3%
1.4%
12
Journal of Structural Biology
64 papers in training set
Top 0.5%
1.4%
13
Journal of Chemical Information and Modeling
238 papers in training set
Top 2%
1.1%
14
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 37%
1.1%
15
eLife
5828 papers in training set
Top 62%
1.0%
16
Molecular Biology and Evolution
542 papers in training set
Top 5%
1.0%
17
Communications Biology
993 papers in training set
Top 27%
1.0%
18
Nature Methods
385 papers in training set
Top 6%
0.9%
19
Scientific Reports
3612 papers in training set
Top 77%
0.8%
20
ACS Omega
105 papers in training set
Top 4%
0.6%
21
Journal of Cheminformatics
29 papers in training set
Top 0.8%
0.6%