Pith. sign in

REVIEW 37 cited by

Virchow2: Scaling Self-Supervised Mixed Magnification Models in Pathology

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.00738 v3 pith:4YAKZJD5 submitted 2024-08-01 cs.CV

Virchow2: Scaling Self-Supervised Mixed Magnification Models in Pathology

classification cs.CV
keywords modelsdatascalemillionmodelparameterpathologyperformance
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Foundation models are rapidly being developed for computational pathology applications. However, it remains an open question which factors are most important for downstream performance with data scale and diversity, model size, and training algorithm all playing a role. In this work, we propose algorithmic modifications, tailored for pathology, and we present the result of scaling both data and model size, surpassing previous studies in both dimensions. We introduce three new models: Virchow2, a 632 million parameter vision transformer, Virchow2G, a 1.9 billion parameter vision transformer, and Virchow2G Mini, a 22 million parameter distillation of Virchow2G, each trained with 3.1 million histopathology whole slide images, with diverse tissues, originating institutions, and stains. We achieve state of the art performance on 12 tile-level tasks, as compared to the top performing competing models. Our results suggest that data diversity and domain-specific methods can outperform models that only scale in the number of parameters, but, on average, performance benefits from the combination of domain-specific methods, data scale, and model scale.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 37 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

    cs.CV 2026-06 unverdicted novelty 7.0

    STREAM applies stochastic Riemannian flow matching on VFM-derived unit hypersphere latents with a novel anisotropic decoder to achieve SOTA reconstruction and generation on breast and colorectal cancer histopathology ...

  2. Geometry-Aware State Space Model: A New Paradigm for Whole-Slide Image Representation

    cs.CV 2026-05 unverdicted novelty 7.0

    BatMIL uses hybrid hyperbolic-Euclidean geometry, an S4 state-space backbone, and chunk-level mixture-of-experts to outperform prior multiple-instance learning methods on seven whole-slide image datasets across six cancers.

  3. A Generative Foundation Model for Multimodal Histopathology

    cs.CV 2026-04 unverdicted novelty 7.0

    MuPD is a pretrained generative foundation model using a diffusion transformer with cross-modal attention that synthesizes histopathology images from text or RNA data and outperforms task-specific models on generation...

  4. LaGuadia: Language-Guided Adaptive Distillation from Pathology Foundation Models

    cs.CV 2026-07 conditional novelty 6.0

    Language-guided adaptive multi-teacher distillation yields an 87M pathology encoder that matches or exceeds GigaPath and UNI on WSI captioning, VQA, and MIL tasks.

  5. ALICE: Learning a General-Purpose Pathology Foundation Model from Vision, Vision-Language, and Slide-Level Experts

    cs.CV 2026-07 accept novelty 6.0

    Multi-stage agglomerative distillation consolidates eight vision, vision-language, and slide-level pathology teachers into one backbone that ranks first on average across 96 downstream tasks.

  6. Multi-Teacher Contrastive Distillation for Edge-Efficient Pathology Foundation Models

    cs.CV 2026-07 conditional novelty 6.0

    Multi-teacher contrastive distillation (MuCoDi) compresses Virchow2/UNI2/H-Optimus-1 into edge encoders that reach 71.0% external AUROC (vs 71.8% best teacher) and up to 605× Raspberry Pi speedup.

  7. The Good, the Bad, and the Brittle: Benchmarking Robustness and Generalisation of Histopathology Foundation Models

    cs.CV 2026-07 conditional novelty 6.0

    Mid-sized pathology foundation models match or beat billion-parameter ones on clinically realistic perturbations and distribution-shift tests, so scaling alone has largely saturated for robustness.

  8. JASPR: Joint Spatial Representation learning of histology and spatial genomics for improved virtual genomic screening and clinical prognostication

    cs.CV 2026-06 unverdicted novelty 6.0

    JASPR integrates HE images and ST data through self-supervised cross-modal reconstruction with shared and modality-specific modules to improve virtual gene prediction and breast cancer prognostication.

  9. A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology

    cs.AI 2026-06 unverdicted novelty 6.0

    PathPocket constructs a 4.55M-entity pathology hypergraph from 110k graded documents and deploys a multi-agent framework that outperforms prior systems on 200k cases while raising pathologist accuracy in user studies.

  10. DaX: Learning General Pathology Representations Across Scales

    eess.IV 2026-06 unverdicted novelty 6.0

    DaX is a pathology vision foundation model that extends DINOv3 with continuous magnification training and cross-scale consistency, achieving top average performance on a benchmark of 161 tasks from 44 datasets coverin...

  11. LRMIL: Efficient Low-Resolution Multiple Instance Learning via High-Resolution Knowledge Distillation for Whole Slide Image Classification

    cs.CV 2026-06 unverdicted novelty 6.0

    LRMIL employs a two-stage high-to-low resolution knowledge distillation strategy to train an efficient low-resolution MIL model for WSI classification that outperforms existing methods with lower computational cost.

  12. Symb-xMIL: Symbolic Explanations for Multiple Instance Learning in Digital Pathology

    cs.CV 2026-06 unverdicted novelty 6.0

    Symb-xMIL is a post-hoc explanation framework that quantifies MIL model alignment with logical decision rules in histopathology to enable rule-based interpretability.

  13. A Pathology Foundation Model for Gastric Cancer with Real-World Validation

    cs.CV 2026-06 unverdicted novelty 6.0

    GRACE, a gastric-specific pathology foundation model trained on multicenter HE-stained slides, outperforms pancancer models on 28 tasks and improves pathologist accuracy, speed, and agreement in a reader study while e...

  14. Do Foundation Models See Biology? Evaluating Attention Coherence with Spatial Transcriptomics in Glioblastoma

    cs.CV 2026-06 unverdicted novelty 6.0

    A new evaluation framework aligns attention maps from five pathology foundation models with co-registered Visium spatial transcriptomics in glioblastoma, revealing five-fold stronger coherence with multi-gene pathways...

  15. Spatial Transcriptomics-Guided Alignment Enhances Molecular Profiling in Pathology Foundation Model

    cs.LG 2026-05 unverdicted novelty 6.0

    STAMP uses a curated 1.8M-pair spatial transcriptomics atlas and pathway-informed alignment to augment pathology foundation models for molecular phenotype inference from H&E WSIs.

  16. A Clinically Validated Foundation Model for Comprehensive Lung Pathology Interpretation

    eess.IV 2026-05 unverdicted novelty 6.0

    PulmoFoundation achieves 92.3% average AUC on 32 lung pathology tasks in prospective validation and raises pathologist accuracy from 83.8% to 91.7% in a crossover RCT.

  17. CRISP -- Clustering-Based Redundancy-Reduced Instance Sampling for Pathology Case Representation and Retrieval

    cs.CV 2026-05 unverdicted novelty 6.0

    CRISP is a clustering-based sampling framework that builds case-level representations from multiple whole-slide images for improved pathology retrieval, matching or exceeding single-slide selection on two breast cance...

  18. Unified Multi-Foundation-Model Slide Representation for Pan-Cancer Recognition and Text-Guided Tumor Localization

    cs.CV 2026-04 unverdicted novelty 6.0

    ASTRA unifies heterogeneous pathology foundation-model representations for pan-cancer classification and weakly supervised tumor localization using only slide-level structured annotations.

  19. PC-MIL: Decoupling Feature Resolution from Supervision Scale in Whole-Slide Learning

    cs.CV 2026-04 unverdicted novelty 6.0

    PC-MIL shows that anchoring supervision at a 2 mm scale and progressively mixing slide- and region-level labels improves cross-context accuracy in WSI cancer detection without reducing global performance.

  20. MOOZY: A Patient-First Foundation Model for Computational Pathology

    cs.CV 2026-03 conditional novelty 6.0

    Patient-level pretraining with a case transformer and multi-task public supervision yields transferable WSI embeddings that beat larger slide-centric models on held-out pathology tasks.

  21. Enabling clinical use of foundation models for computational pathology

    cs.CV 2026-02 conditional novelty 6.0

    Novel robustness losses added during downstream training on foundation-model features from pathology slides improve both robustness to technical variation and classification accuracy.

  22. Uncertainty Estimation in Pathology Foundation Models via Deep Mutual Learning

    cs.CV 2026-06 unverdicted novelty 5.0

    DICE ensembles frozen pathology foundation models, aligns them with deep mutual learning to make disagreement a reliable uncertainty proxy, and shows consensus-based localization on WSI tasks.

  23. Mitigating Batch Effects in Histopathology via Language-Mediated Robust Embedding Generation

    cs.CV 2026-06 unverdicted novelty 5.0

    GLMP generates robust pathology embeddings by routing histology images through an intermediate textual representation produced by general-purpose MLLMs to mitigate batch effects.

  24. Multi-FRuGaL: Multimodal Flexible Redundancy-aware Decomposed Gated Learning for Cancer Diagnosis and Prognosis

    cs.CV 2026-06 unverdicted novelty 5.0

    Multi-FRuGaL is a decomposition-aware gated fusion framework for multimodal cancer data that maintains performance under missing modalities and reports AUC gains on two head-and-neck cancer cohorts.

  25. SlideCheck: Guiding Self-Supervised Pretraining of Pathology Foundation Models via Dataset Distributions

    cs.CV 2026-05 unverdicted novelty 5.0

    SlideCheck uses a dual-head MLP on frozen features plus MIL attention to score patches and filter pretraining subsets that approach full-data performance in self-supervised pathology ViT models.

  26. Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction

    cs.CV 2026-05 unverdicted novelty 5.0

    A masked-diffusion pretrained convolutional model outperforms ViT pathology foundation models on cell-level dense prediction tasks in histology.

  27. Retrieval-Guided Generation for Safer Histopathology Image Captioning

    cs.CV 2026-04 unverdicted novelty 5.0

    Retrieval-guided captioning from similar cases achieves higher semantic alignment (cosine similarity ~0.60 vs ~0.47) and fewer unsupported diagnoses than MedGemma on the ARCH dataset.

  28. Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

    cs.CV 2026-04 unverdicted novelty 5.0

    Weakly supervised MIL with foundation models enables robust five-grade Nancy index prediction and neutrophilic activity assessment from slide-level labels in multicenter UC biopsies.

  29. SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification

    cs.CV 2026-04 unverdicted novelty 5.0

    SSMamba uses a two-stage self-supervised pretraining and fine-tuning pipeline with Mamba-based components to outperform prior pathological foundation models on ROI and WSI classification tasks.

  30. Evaluating Computational Pathology Foundation Models for Prostate Cancer Grading under Distribution Shifts

    eess.IV 2024-10 unverdicted novelty 5.0

    Pathology foundation models deliver strong in-distribution prostate cancer grading performance but exhibit large drops under cross-site image appearance shifts while remaining relatively robust to label distribution shifts.

  31. CellPrior-Net: Prior-Guided Nuclei Detection and Classification for H&E Whole-Slide Images

    cs.MM 2026-07 unverdicted novelty 4.0

    CellPrior-Net integrates hematoxylin channel prior into a lightweight CNN for nuclei detection and classification in H&E WSIs, claiming comparable accuracy to SOTA with significantly reduced inference time across 10.4...

  32. Mitosis Detection in the Wild: Multi-Tumor and Context-Aware Generalization in the MIDOG 2025 Challenge

    cs.CV 2026-06 unverdicted novelty 4.0

    MIDOG 2025 challenge shows top mitosis detection F1 of 0.740 and atypical figure balanced accuracy of 0.908 across diverse tumors, with clear drops in challenging regions and tumor-type variation.

  33. Validation of an AI-based end-to-end model for prostate pathology using long-term archived routine samples

    cs.CV 2026-05 unverdicted novelty 4.0

    GleasonAI achieves quadratic-weighted kappa of 0.86 on ISUP grading of 10,366 long-term archived prostate biopsy cores, with performance stable over 17 years and a clear prognostic gradient for cancer-specific mortality.

  34. Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction

    cs.CV 2026-04 unverdicted novelty 4.0

    H-optimus-1 achieves the strongest externally validated survival prediction from histopathology images, with second-generation PFMs outperforming first-generation counterparts and a compact distilled model offering ef...

  35. OpenTME: An Open Dataset of AI-powered H&E Tumor Microenvironment Profiles from TCGA

    cs.CV 2026-04 unverdicted novelty 4.0

    OpenTME provides pre-computed TME profiles with over 4,500 quantitative readouts per slide from 3,634 TCGA H&E images using an AI pipeline based on pathology foundation models.

  36. MMAP: A Multi-Magnification and Prototype-Aware Architecture for Predicting Spatial Gene Expression

    cs.CV 2025-10 unverdicted novelty 4.0

    MMAP uses multi-magnification patch features and slide-level prototype embeddings to predict spatial gene expression from H&E images and reports better MAE, MSE, and PCC than prior methods.

  37. Transformer-Based Hematological Malignancy Prediction from Peripheral Blood Smears in a Real-World Cohort

    q-bio.QM 2025-09 unverdicted novelty 4.0

    cAItomorph applies a transformer aggregator on DinoBloom embeddings to predict eight coarse hematological malignancy classes from peripheral blood single-cell images, achieving 0.72 accuracy and reducing false discove...