Back

Analytical perturbation reveals hidden instability of biological phenotypes

Piorkowska, N. J.; Ostromecki, A.; Franik, G.; Bizon, A.

2026-07-16 endocrinology
10.64898/2026.07.13.26357916 medRxiv
Show abstract

Background Unsupervised machine learning has become a cornerstone of computational phenotyping across clinical medicine, genomics, imaging, and multi-omics research. However, phenotype discovery relies on a sequence of analytical decisions - including missing-data handling, preprocessing, dimensionality reduction, clustering methodology, and stochastic initialization - that are rarely evaluated collectively. Although clustering stability has been extensively investigated, the robustness of complete analytical workflows remains largely unexplored. Results We developed an Analytical Perturbation Framework that systematically quantifies the robustness of phenotype discovery by perturbing complete unsupervised learning workflows rather than individual clustering algorithms. Using a real-world cohort of 1,286 women with polycystic ovary syndrome (PCOS), we generated 116 valid analytical pipelines comprising alternative preprocessing strategies, missing-data handling methods, dimensionality reduction approaches, clustering algorithms, and random initializations. Agreement between independently generated phenotype solutions was consistently low (median Adjusted Rand Index = 0.079), indicating substantial sensitivity of phenotype discovery to routine analytical decisions. Variance decomposition identified preprocessing as the largest contributor to phenotype instability (22.8%), followed by clustering methodology (14.6%), whereas stochastic initialization explained only 3.1% of the observed variability. At the patient level, most individuals exhibited reproducible phenotype assignments (median Patient Robustness Score = 0.719), although a substantial subgroup showed markedly lower assignment stability. Feature perturbation analyses identified follicle-stimulating hormone, anti-thyroglobulin antibodies, anti-thyroid peroxidase antibodies, total testosterone, luteinizing hormone, and androstenedione as the strongest contributors to computational robustness, rather than biological importance. Finally, phenotype solutions demonstrating greater computational robustness also exhibited greater biological coherence during independent validation.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

1
Communications Medicine
113 papers in training set
Top 0.1%
13.1%
2
npj Digital Medicine
118 papers in training set
Top 0.6%
9.9%
3
Translational Psychiatry
260 papers in training set
Top 1.0%
5.5%
4
Scientific Reports
3612 papers in training set
Top 15%
5.5%
5
BMC Medicine
176 papers in training set
Top 0.4%
5.5%
6
BMC Medical Research Methodology
47 papers in training set
Top 0.3%
4.4%
7
The Journal of Clinical Endocrinology & Metabolism
36 papers in training set
Top 0.2%
4.4%
8
Wellcome Open Research
67 papers in training set
Top 0.3%
3.3%
50% of probability mass above
9
eLife
5828 papers in training set
Top 33%
3.3%
10
Molecular Systems Biology
162 papers in training set
Top 0.6%
3.3%
11
Nature Communications
5641 papers in training set
Top 35%
3.3%
12
Computational and Structural Biotechnology Journal
242 papers in training set
Top 1%
3.3%
13
Cell Reports Medicine
153 papers in training set
Top 2%
2.1%
14
Science Advances
1243 papers in training set
Top 16%
2.1%
15
PLOS ONE
5266 papers in training set
Top 44%
2.1%
16
The American Journal of Human Genetics
234 papers in training set
Top 2%
1.1%
17
Cell Genomics
172 papers in training set
Top 3%
1.1%
18
Metabolites
53 papers in training set
Top 0.7%
1.1%
19
Journal of Pathology Informatics
15 papers in training set
Top 0.2%
1.1%
20
Journal of the Endocrine Society
15 papers in training set
Top 0.2%
1.1%
21
Genome Medicine
183 papers in training set
Top 4%
1.1%
22
BMC Bioinformatics
457 papers in training set
Top 5%
1.0%
23
Genetics in Medicine
78 papers in training set
Top 0.9%
0.9%
24
JAMA Network Open
130 papers in training set
Top 4%
0.9%
25
Human Reproduction
20 papers in training set
Top 0.3%
0.9%
26
JAMIA Open
42 papers in training set
Top 1%
0.9%
27
PLOS Medicine
110 papers in training set
Top 4%
0.6%
28
IEEE/ACM Transactions on Computational Biology and Bioinformatics
38 papers in training set
Top 1%
0.6%
29
Sensors
43 papers in training set
Top 1%
0.6%
30
PLOS Computational Biology
1863 papers in training set
Top 21%
0.6%