REVIEW 3 cited by
Do Histopathological Foundation Models Eliminate Batch Effects? A Comparative Study
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Deep learning has led to remarkable advancements in computational histopathology, e.g., in diagnostics, biomarker prediction, and outcome prognosis. Yet, the lack of annotated data and the impact of batch effects, e.g., systematic technical data differences across hospitals, hamper model robustness and generalization. Recent histopathological foundation models -- pretrained on millions to billions of images -- have been reported to improve generalization performances on various downstream tasks. However, it has not been systematically assessed whether they fully eliminate batch effects. In this study, we empirically show that the feature embeddings of the foundation models still contain distinct hospital signatures that can lead to biased predictions and misclassifications. We further find that the signatures are not removed by stain normalization methods, dominate distances in feature space, and are evident across various principal components. Our work provides a novel perspective on the evaluation of medical foundation models, paving the way for more robust pretraining strategies and downstream predictors.
Forward citations
Cited by 3 Pith papers
-
Harnessing Adversarial Distillation to Customise Debiased, Disease-Specific Pathology Foundation Models for Breast Cancer
SmartStu distills multiple teacher pathology models into compact breast-cancer encoders with an adversarial noise model and self-supervision, matching or improving external-cohort accuracy at over 30x smaller size.
-
Beyond Counts: A Distributional Robustness Margin For Pathology Foundation Models
CRoMa scores each pathology image embedding by the margin between cross-site biological matches and same-site biological distractors, revealing distributional lower tails that pooled robustness scores hide.
-
MeDi: Metadata-Guided Diffusion Models for Mitigating Biases in Tumor Classification
Conditioning a histopathology diffusion model on tissue source site metadata improves synthetic image fidelity and enhances downstream tumor classification under subpopulation shift.
Discussion (0). Sign in to comment.