REVIEW 2 cited by
GazeSAM: What You See is What You Segment
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This study investigates the potential of eye-tracking technology and the Segment Anything Model (SAM) to design a collaborative human-computer interaction system that automates medical image segmentation. We present the \textbf{GazeSAM} system to enable radiologists to collect segmentation masks by simply looking at the region of interest during image diagnosis. The proposed system tracks radiologists' eye movement and utilizes the eye-gaze data as the input prompt for SAM, which automatically generates the segmentation mask in real time. This study is the first work to leverage the power of eye-tracking technology and SAM to enhance the efficiency of daily clinical practice. Moreover, eye-gaze data coupled with image and corresponding segmentation labels can be easily recorded for further advanced eye-tracking research. The code is available in \url{https://github.com/ukaukaaaa/GazeSAM}.
Forward citations
Cited by 2 Pith papers
-
GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs
GazeLT classifies long-tailed chest X-ray diseases by training a student model with a teacher that learns time-windowed radiologist gaze attention, improving tail-class accuracy on the NIH-CXR-LT and MIMIC-CXR-LT benchmarks.
-
Zero-Shot Gaze-based Volumetric Medical Image Segmentation
Gaze-based prompts can drive zero-shot SAM-2 and MedSAM-2 segmentation of 3D abdominal CT organs, trading a roughly 0.08 Dice drop for an approximately 26 second per-organ speedup over manual bounding boxes.
Discussion (0). Continue with ORCID to comment.