Morphology signal in whole slide image foundation models can automatically triage slides

๐Ÿ“… 2026-09-01
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
ๆœฌๆ–‡ไฝฟ็”จๅ…ฌๅผ€็š„ๅ…จๅˆ‡็‰‡ๅ›พๅƒๅŸบ็ก€ๆจกๅž‹่‡ชๅŠจ็ญ›้€‰ๅซๆœ‰่‚ฟ็˜ค็š„ๅˆ‡็‰‡๏ผŒ้€š่ฟ‡้›ถๆ ทๆœฌๅˆ†็ฑป้ข„ๆต‹ๆŽ’ๅๆœ‰ๆ•ˆ่ฏ†ๅˆซๅ…ณ้”ฎไฟกๆฏใ€‚
๐Ÿ“ Abstract
Patient exams in the cancer diagnosis and staging process typically generate several whole slide images (WSIs). One of the initial steps in training models on WSI data is identifying one or a few slides containing tumor or other diagnostic biomarkers necessary for downstream prediction tasks such as estimating recurrence risk or progression-free survival. This step requires tedious manual curation by experienced pathologists. Many published datasets make the artificial assumption of 1 slide per patient. Alternatively, all slides per patient may be used for model training, which may dilute the signal from the few slides containing tumor or other relevant information. In this paper, we present a pipeline to overcome these challenges using publicly available WSI foundation models (FMs). Our evaluations show that ranking WSIs based on predictions from zero-shot classification using WSI FMs accurately identifies slides with the most tumor, indicating that WSI FMs contain sufficient morphology signal to automatically triage slides. We also present a formulation for ranked evaluation to benchmark FM performance in slide triage. We show, on multiple datasets, that tumor slides are identified in the top-2 ranked slides for patients with up to 43 slides.
Problem

Research questions and friction points this paper is trying to address.

whole slide images
tumor identification
manual curation
foundation models
cancer diagnosis
Innovation

Methods, ideas, or system contributions that make the work stand out.

Whole Slide Image Foundation Models
Automatic Triage
Zero-shot Classification
Ranked Evaluation
๐Ÿ”Ž Similar Papers
No similar papers found.
A
Ayushi Sinha
Department of Radiology, Mayo Clinic, Rochester MN
Shashank Yadav
Shashank Yadav
University of Arizona
Biomedical InformaticsRepresentation Learning
B
Benjamin Holmes
AI Program, Mayo Clinic, Rochester MN
P
Pravat Das
AI Program, Mayo Clinic, Rochester MN
A
Aaron W. Bogan
Department of Quantitative Health Sciences, Mayo Clinic, Scottsdale AZ
J
James S. Lewis Jr.
Department of Laboratory Medicine and Pathology, Mayo Clinic, Scottsdale AZ
Santiago Romero-Brufau
Santiago Romero-Brufau
Assistant Professor, Mayo Clinic
early warning scoresmachine learningclinical implementationclinical informatics
Andrew Y. K. Foong
Andrew Y. K. Foong
AI Scientist, Mayo Clinic
Deep learningAI for healthcareAI for scienceBayesian Deep LearningGenerative models
S
Scott H. Kaufmann
Department of Oncology, Mayo Clinic, Rochester MN
K
Kathryn M. Van Abel
Department of Otolaryngology-Head and Neck Surgery, Mayo Clinic, Rochester MN
D
David M. Routman
Department of Radiation Oncology, Mayo Clinic, Rochester MN
M
Michael R. Lucas
AI Program, Mayo Clinic, Rochester MN