Back

A new reference-invariant consensus template generation method in ALPACA

Maga, A. M.

2026-05-30 evolutionary biology
10.64898/2026.05.29.728799 bioRxiv
Show abstract

O_LIAutomated landmarking transfers anatomical landmarks from a reference specimen onto many targets, greatly increasing analytical throughput. However, this procedure needs to be bootstrap using an initial sample. An arbitrary or atypical choice imprints a reference-of-origin bias that propagates through the pseudo-landmarks, the resulting morphospace, and the downstream template selection, a risk that is difficult to avoid for large datasets whose variation is not yet understood. C_LIO_LIWe replace the fixed reference with an iterative consensus atlas, warped over a few iterations toward the Procrustes mean shape of all similarity-aligned specimens. We evaluated it on a 62-strain Mus musculus skull panel by running both the original fixed-reference pipeline and the new consensus pipeline 62 times each, using every specimen in turn as the bootstrap. We compared atlas convergence, inter-atlas similarity, morphospace reproducibility, reference-choice variance of pairwise Procrustes distances, downstream k-means selection stability, and leave-one-out out-of-sample fit, and tested generalisation on great-ape datasets of differing sampling balance. C_LIO_LIThe consensus atlas converged within a few iterations and was far less sensitive to the starting specimen than the fixed reference. It produced more reproducible morphospaces (mean RV 0.960 versus 0.944), reduced the reference-of-origin variance of pairwise distances by a median of about 60%, drew downstream template selections from a smaller and more consistent pool of specimens, and fit held-out specimens more closely in all 62 strains. On the great-ape data the atlases agreed closely in surface geometry, but the downstream morphospace became reference-dependent when the sample was taxonomically imbalanced, and a smaller balanced subset outperformed the larger imbalanced one. C_LIO_LIIterative consensus atlas building removes a persistent bias from automated landmarking and yields reference-invariant, reproducible results, with sampling balance mattering more than absolute sample size. Because the atlas stabilises quickly, it can be built from a small balanced subset while the remaining specimens are simply landmarked against it, a practical route to scaling reference-invariant landmarking. The method is implemented in ALPACA within SlicerMorph, with a mock library enabling headless use on HPC. C_LI

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Nature Communications
5641 papers in training set
Top 11%
16.5%
2
eLife
5828 papers in training set
Top 9%
10.7%
3
Communications Biology
993 papers in training set
Top 0.6%
7.6%
4
Science
477 papers in training set
Top 1%
6.5%
5
PLOS ONE
5266 papers in training set
Top 30%
5.0%
6
Science Advances
1243 papers in training set
Top 6%
4.7%
50% of probability mass above
7
Scientific Reports
3612 papers in training set
Top 28%
3.9%
8
Nature
645 papers in training set
Top 4%
3.1%
9
Nature Methods
385 papers in training set
Top 3%
3.1%
10
Systematic Biology
144 papers in training set
Top 0.4%
2.3%
11
Evolutionary Biology
14 papers in training set
Top 0.1%
2.3%
12
PLOS Biology
486 papers in training set
Top 4%
2.1%
13
Bioinformatics
1204 papers in training set
Top 6%
2.1%
14
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 29%
1.7%
15
PLOS Computational Biology
1863 papers in training set
Top 15%
1.7%
16
iScience
1154 papers in training set
Top 21%
1.5%
17
Peer Community Journal
281 papers in training set
Top 3%
1.5%
18
Development
497 papers in training set
Top 3%
1.5%
19
Imaging Neuroscience
282 papers in training set
Top 3%
1.4%
20
Scientific Data
209 papers in training set
Top 2%
1.1%
21
eneuro
439 papers in training set
Top 6%
1.1%
22
GigaScience
212 papers in training set
Top 4%
1.1%
23
Methods in Ecology and Evolution
176 papers in training set
Top 2%
1.0%
24
Genome Biology
637 papers in training set
Top 8%
1.0%
25
Ecology and Evolution
267 papers in training set
Top 5%
1.0%
26
Cell Reports
1498 papers in training set
Top 26%
1.0%
27
PeerJ
308 papers in training set
Top 12%
0.8%
28
American Journal of Biological Anthropology
12 papers in training set
Top 0.1%
0.8%
29
Molecular Biology and Evolution
542 papers in training set
Top 5%
0.8%
30
BMC Ecology and Evolution
51 papers in training set
Top 2%
0.6%