Unifying Multiple Foundation Models for Advanced Computational Pathology

Lei, Wenhui; Tan, Yusheng; Li, Anqi; Chen, Hanyu; Tian, Hengrui; Li, Ruiying; Jiang, Zhengqun; Yan, Fang; Zhang, Xiaofan; Zhang, Shaoting

Computer Science > Computer Vision and Pattern Recognition

arXiv:2503.00736v3 (cs)

[Submitted on 2 Mar 2025 (v1), last revised 11 Dec 2025 (this version, v3)]

Title:Unifying Multiple Foundation Models for Advanced Computational Pathology

Authors:Wenhui Lei, Yusheng Tan, Anqi Li, Hanyu Chen, Hengrui Tian, Ruiying Li, Zhengqun Jiang, Fang Yan, Xiaofan Zhang, Shaoting Zhang

View PDF HTML (experimental)

Abstract:Foundation models have advanced computational pathology by learning transferable visual representations from large histological datasets, yet recent evaluations reveal substantial variability in their performance across tasks. This inconsistency arises from differences in training data diversity and is further constrained by the reliance of many high-performing models on proprietary datasets that cannot be shared or expanded. Offline distillation offers a partial remedy but depends heavily on the size and heterogeneity of the distillation corpus and requires full retraining to incorporate new models. To address these limitations, we propose Shazam, a task-specific online integration framework that unifies multiple pretrained pathology foundation models within a single flexible inference system. Shazam fuses multi-level representations through adaptive expert weighting and learns task-aligned features via online distillation. Across spatial transcriptomics prediction, survival prognosis, tile classification, and visual question answering, Shazam consistently outperforms strong individual models, highlighting its promise as a scalable approach for harnessing the rapid evolution of pathology foundation models in a unified and adaptable manner.

Comments:	37 pages, 5 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2503.00736 [cs.CV]
	(or arXiv:2503.00736v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2503.00736

Submission history

From: WenHui Lei [view email]
[v1] Sun, 2 Mar 2025 05:20:41 UTC (2,466 KB)
[v2] Thu, 6 Mar 2025 03:35:09 UTC (2,466 KB)
[v3] Thu, 11 Dec 2025 04:35:11 UTC (3,052 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Unifying Multiple Foundation Models for Advanced Computational Pathology

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Unifying Multiple Foundation Models for Advanced Computational Pathology

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators