Back

Finetuning masking challenges narrow-task evaluation of cell foundation models

Shakeel, M. H.; Shen, M.; Mangiola, S.

2026-06-06 bioinformatics
10.64898/2026.06.04.730272 bioRxiv
Show abstract

Single-cell foundation models are large, self-supervised deep learning networks pretrained on millions of cellular transcriptomes. These models promise to deliver cell representations that are transferable across diverse biological domains and, when used in specific tasks, would outperform narrowly scoped models. A central assumption is that more pretraining data translates to better downstream performance. However, despite its centrality, this assumption remains largely untested. Here, we tested downstream performance on gold-standard benchmarking tasks across massive dataset reductions, showing that performance was largely insensitive to pretraining data size once finetuning was allowed. This trend reveals a finetuning masking effect that offsets differences in representation quality induced by pretraining, making the benefit of additional pretraining scale largely invisible under current benchmark settings. These findings challenge current benchmarking standards, which rely on closed-ended finetuning tasks that are too narrow to expose the full representational value of pretraining. They also challenge the main driving force in single-cell foundation-model development when evaluated through common narrow tasks. We propose that the next generation of foundation models should be assessed less by performance on highly optimised finetuning tasks and more by their ability to support open-ended biological inference, frozen-representation evaluation and zero-shot capability.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

1
Nature Machine Intelligence
70 papers in training set
Top 0.1%
23.0%
2
Nature Communications
5641 papers in training set
Top 10%
17.5%
3
PLOS Computational Biology
1863 papers in training set
Top 4%
8.1%
4
Cell Systems
201 papers in training set
Top 0.9%
5.0%
50% of probability mass above
5
Scientific Reports
3612 papers in training set
Top 31%
3.3%
6
Nature Methods
385 papers in training set
Top 3%
3.3%
7
Briefings in Bioinformatics
354 papers in training set
Top 3%
2.8%
8
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 21%
2.5%
9
Nature
645 papers in training set
Top 5%
2.5%
10
NAR Genomics and Bioinformatics
242 papers in training set
Top 2%
2.5%
11
Genome Biology
637 papers in training set
Top 4%
2.5%
12
Patterns
78 papers in training set
Top 1.0%
2.2%
13
Bioinformatics
1204 papers in training set
Top 6%
2.2%
14
PLOS ONE
5266 papers in training set
Top 52%
1.4%
15
Molecular Systems Biology
162 papers in training set
Top 2%
1.2%
16
New Phytologist
346 papers in training set
Top 4%
1.1%
17
eLife
5828 papers in training set
Top 60%
1.1%
18
Communications Biology
993 papers in training set
Top 26%
1.0%
19
Frontiers in Genetics
230 papers in training set
Top 5%
1.0%
20
Bioinformatics Advances
203 papers in training set
Top 4%
0.9%
21
iScience
1154 papers in training set
Top 33%
0.9%
22
Development
497 papers in training set
Top 5%
0.6%
23
Nucleic Acids Research
1281 papers in training set
Top 14%
0.6%
24
GigaScience
212 papers in training set
Top 5%
0.6%
25
Nature Genetics
286 papers in training set
Top 5%
0.6%