REVIEW 12 cited by
Hibou: A Family of Foundational Vision Transformers for Pathology
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Hibou: A Family of Foundational Vision Transformers for Pathology
read the original abstract
Pathology, the microscopic examination of diseased tissue, is critical for diagnosing various medical conditions, particularly cancers. Traditional methods are labor-intensive and prone to human error. Digital pathology, which converts glass slides into high-resolution digital images for analysis by computer algorithms, revolutionizes the field by enhancing diagnostic accuracy, consistency, and efficiency through automated image analysis and large-scale data processing. Foundational transformer pretraining is crucial for developing robust, generalizable models as it enables learning from vast amounts of unannotated data. This paper introduces the Hibou family of foundational vision transformers for pathology, leveraging the DINOv2 framework to pretrain two model variants, Hibou-B and Hibou-L, on a proprietary dataset of over 1 million whole slide images (WSIs) representing diverse tissue types and staining techniques. Our pretrained models demonstrate superior performance on both patch-level and slide-level benchmarks, surpassing existing state-of-the-art methods. Notably, Hibou-L achieves the highest average accuracy across multiple benchmark datasets. To support further research and application in the field, we have open-sourced the Hibou models, which can be accessed at https://github.com/HistAI/hibou.
Forward citations
Cited by 12 Pith papers
-
Benchmarking Pathology Foundation Models for Spatial Domain Understanding
SpaPath-Bench evaluates spatial representation in 19 pathology foundation models via spatial domain identification on 42 paired WSI-ST slides using three agreement criteria across 83K runs.
-
The Good, the Bad, and the Brittle: Benchmarking Robustness and Generalisation of Histopathology Foundation Models
Mid-sized pathology foundation models match or beat billion-parameter ones on clinically realistic perturbations and distribution-shift tests, so scaling alone has largely saturated for robustness.
-
DaX: Learning General Pathology Representations Across Scales
DaX is a pathology vision foundation model that extends DINOv3 with continuous magnification training and cross-scale consistency, achieving top average performance on a benchmark of 161 tasks from 44 datasets coverin...
-
CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers
CellDX AI Autopilot lets users train pathology classifiers via AI agent skills on a large pre-extracted whole-slide image dataset with automated hyperparameter tuning that claims over 30x cost reduction.
-
MOOZY: A Patient-First Foundation Model for Computational Pathology
Patient-level pretraining with a case transformer and multi-task public supervision yields transferable WSI embeddings that beat larger slide-centric models on held-out pathology tasks.
-
Enabling clinical use of foundation models for computational pathology
Novel robustness losses added during downstream training on foundation-model features from pathology slides improve both robustness to technical variation and classification accuracy.
-
Uncertainty Estimation in Pathology Foundation Models via Deep Mutual Learning
DICE ensembles frozen pathology foundation models, aligns them with deep mutual learning to make disagreement a reliable uncertainty proxy, and shows consensus-based localization on WSI tasks.
-
Mitigating Batch Effects in Histopathology via Language-Mediated Robust Embedding Generation
GLMP generates robust pathology embeddings by routing histology images through an intermediate textual representation produced by general-purpose MLLMs to mitigate batch effects.
-
Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction
A masked-diffusion pretrained convolutional model outperforms ViT pathology foundation models on cell-level dense prediction tasks in histology.
-
Evaluating Computational Pathology Foundation Models for Prostate Cancer Grading under Distribution Shifts
Pathology foundation models deliver strong in-distribution prostate cancer grading performance but exhibit large drops under cross-site image appearance shifts while remaining relatively robust to label distribution shifts.
-
CellPrior-Net: Prior-Guided Nuclei Detection and Classification for H&E Whole-Slide Images
CellPrior-Net integrates hematoxylin channel prior into a lightweight CNN for nuclei detection and classification in H&E WSIs, claiming comparable accuracy to SOTA with significantly reduced inference time across 10.4...
-
Mitosis Detection in the Wild: Multi-Tumor and Context-Aware Generalization in the MIDOG 2025 Challenge
MIDOG 2025 challenge shows top mitosis detection F1 of 0.740 and atypical figure balanced accuracy of 0.908 across diverse tumors, with clear drops in challenging regions and tumor-type variation.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.