Back

A Systematic Review and Independent Benchmarking of Automated Nerve Morphometry Methods

Chuter, B.; Kim, M. Y.; Stiemke, A. B.; Dave, N.; Zhou, Z. A.; Herrin, J.; Miller, M. C.; White, W.; Hollingsworth, T. J.; Jablonski, M. M.

2026-06-16 bioengineering
10.64898/2026.06.15.731646 bioRxiv
Show abstract

ObjectiveTo systematically review automated nerve morphometry tools and independently benchmark their performance on independent optic nerve datasets. DesignSystematic review and comparative benchmarking study. ControlsBenchmarking was performed using paraphenylenediamine-stained mouse (n = 85) and rat (n = 44) optic nerve images with manually annotated axon counts as ground truth. MethodsPublished studies describing automated or semi-automated neural tissue morphometry tools were identified through systematic searches of PubMed, Embase, and Scopus through January 2026 following PRISMA guidelines. Data extraction covered 70 fields across tool capabilities, imaging modality, species, automation level, and validation approach. Eighteen eligible tools (8 deep learning [DL], 10 classical computer vision [CV]) were benchmarked on both mouse and rat independent datasets. Main Outcome MeasuresPerformance was assessed by mean absolute percentage error (MAPE), Pearson correlation, and median predicted-to-ground-truth ratio. Tools were ranked per image and compared using Friedman tests with Nemenyi post-hoc analysis. ResultsSeventy-one studies met inclusion criteria, spanning from 1999 to 2026. Deep learning methods represented 38% (27/71) of studies, increasing from 0% before 2017 to over 55% of publications after 2020. Axon counting was the most common output (73%, 52/71), while only 35% (25/71) reported g-ratio. Among benchmarked tools, Marina (CV, 2010) achieved the lowest average MAPE (32.9%). The top five tools (MAPE ranging from 32.9 to 44.8%) included both CV and DL methods and were statistically indistinguishable by Friedman-Nemenyi analysis (p > 0.05). Performance varied substantially across datasets: AxonJ (CV) achieved the second best MAPE on rat images (27.7%) but the worst on mouse images (438.6%). ConclusionsNo single tool demonstrated consistently superior performance across both datasets. Classical and deep learning approaches achieved comparable accuracy for axon counting. Tool selection should be guided by target species, tissue preparation protocol, and desired morphometric outputs. This systematic review and independent benchmarking study provide an evidence base for tool selection in optic nerve research.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
PLOS ONE
5266 papers in training set
Top 13%
13.7%
2
Experimental Eye Research
32 papers in training set
Top 0.1%
11.5%
3
Journal of Neural Engineering
221 papers in training set
Top 0.5%
7.6%
4
Scientific Reports
3612 papers in training set
Top 8%
7.6%
5
Translational Vision Science & Technology
39 papers in training set
Top 0.2%
6.5%
6
Bioengineering
29 papers in training set
Top 0.1%
5.1%
50% of probability mass above
7
Frontiers in Neurology
102 papers in training set
Top 0.9%
3.7%
8
Biomedical Optics Express
95 papers in training set
Top 0.4%
3.3%
9
Frontiers in Bioengineering and Biotechnology
98 papers in training set
Top 0.5%
3.3%
10
Annals of Clinical and Translational Neurology
34 papers in training set
Top 0.3%
2.9%
11
British Journal of Ophthalmology
14 papers in training set
Top 0.1%
2.9%
12
Neuroinformatics
46 papers in training set
Top 0.4%
1.8%
13
Investigative Ophthalmology & Visual Science
25 papers in training set
Top 0.3%
1.4%
14
Journal of the Association for Research in Otolaryngology
15 papers in training set
Top 0.2%
1.0%
15
Nature Communications
5641 papers in training set
Top 53%
1.0%
16
eLife
5828 papers in training set
Top 62%
1.0%
17
eneuro
439 papers in training set
Top 7%
0.9%
18
Scientific Data
209 papers in training set
Top 3%
0.9%
19
Computational and Structural Biotechnology Journal
242 papers in training set
Top 6%
0.9%
20
European Radiology
15 papers in training set
Top 0.5%
0.9%
21
GigaScience
212 papers in training set
Top 4%
0.9%
22
Frontiers in Neuroscience
256 papers in training set
Top 6%
0.9%
23
Journal of Neuroscience Methods
122 papers in training set
Top 2%
0.6%
24
Journal of Neuroscience Research
27 papers in training set
Top 0.5%
0.6%
25
Frontiers in Cellular Neuroscience
91 papers in training set
Top 2%
0.6%
26
Ophthalmology Science
22 papers in training set
Top 0.3%
0.6%
27
PeerJ
308 papers in training set
Top 12%
0.6%
28
Journal of Biomechanical Engineering
20 papers in training set
Top 0.7%
0.5%
29
Computer Methods and Programs in Biomedicine
28 papers in training set
Top 1%
0.5%
30
PLOS Digital Health
106 papers in training set
Top 5%
0.5%