OpenSlideFM: A Computationally Efficient Multi-Scale Foundation Model for Computational Pathology
Ahmad Zafar, S.; Qin, W.; Chengliang, L.; Khan, A. A.; Nazir, A.; Khalid, F.; Faisal, M. S.
Show abstract
BackgroundComputational pathology increasingly relies on foundation models pre-trained on large-scale histopathology datasets, but existing models require substantial computational resources that limit accessibility for resource-constrained institutions. We present OpenSlideFM, a computationally efficient multi-scale foundation model that balances performance with practical deployment requirements. MethodsWe developed a dual-scale transformer architecture that simultaneously processes high-resolution (0.5 m/pixel) and low-resolution (2.0 m/pixel) features to capture both cellular morphology and tissue architecture. The model was pre-trained using self-supervised learning on 20,000 whole-slide images from The Cancer Genome Atlas spanning 31 cancer types and 10,795 patients (Table 1). We validated OpenSlideFM across four independent tasks: pan-cancer classification, lymph node metastasis detection, pathological staging, and prostate cancer grading. O_TBL View this table: org.highwire.dtl.DTLVardef@117e6b6org.highwire.dtl.DTLVardef@2be886org.highwire.dtl.DTLVardef@aeead8org.highwire.dtl.DTLVardef@1bc1897org.highwire.dtl.DTLVardef@1f28b6d_HPS_FORMAT_FIGEXP M_TBL O_FLOATNOTable 1C_FLOATNO O_TABLECAPTIONMulti-Dataset Evaluation Design C_TABLECAPTION C_TBL ResultsOpenSlideFM achieved 81.21% accuracy on 31-class pan-cancer classification, macro-AUROC of 98.65%. Multi-scale architecture significantly outperformed single-scale variants, with 2.35% absolute improvement over high-resolution alone. External validation demonstrated robust generalization: 77.4% AUROC for lymph node metastasis detection, 0.254 quadratic kappa for multi-center pathological staging, and 0.826 quadratic kappa for prostate cancer grading. The model requires only 35 million parameters and trains on a single consumer-grade workstation with NVIDIA GeForce RTX 4090 GPU (24 GB VRAM, 16-core CPU, 384 GB RAM), enabling accessible deployment compared to 300-1850 million parameters and datacenter GPU requirements for existing foundation models. ConclusionsOpenSlideFM demonstrates that computationally efficient foundation models can achieve competitive performance across diverse histopathology tasks while maintaining practical deployment feasibility. The multi-scale architecture provides complementary information from cellular and tissue levels, and the reduced computational requirements democratize access to foundation model capabilities for resource-constrained medical institutions.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- PHARAOH: A collaborative crowdsourcing platform for PHenotyping And Regional Analysis Of Histology 96%
- Segmenting functional tissue units across human organs using community-driven development of generalizable machine learning algorithms 95%
- Generative AI Enables Medical Image Segmentation in Ultra Low-Data Regimes 95%
Similar papers in this journal
- Generalizing AI-driven Assessment of Immunohistochemistry across Immunostains and Cancer Types: A Universal Immunohistochemistry Analyzer 96%
- A Deep Learning Model for Molecular Label Transfer that Enables Cancer Cell Identification from Histopathology Images 95%
- Image-Based Consensus Molecular Subtyping in Rectal Cancer Biopsies and Response to Neoadjuvant Chemoradiotherapy 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- A Crowdsourcing Approach to Develop Machine Learning Models to Quantify Radiographic Joint Damage in Rheumatoid Arthritis 91%
- Real-world effectiveness of Ad26.COV2.S adenoviral vector vaccine for COVID-19 86%
- Diagnostic Codes in AI prediction models and Label Leakage of Same-admission Clinical Outcomes 86%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.