Pith. sign in

REVIEW 2 cited by

GazeSAM: What You See is What You Segment

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.13844 v1 pith:2FFIYD4B submitted 2023-04-26 cs.CV

classification cs.CV
keywords segmentationeye-trackinggazesamimagesystemdataeye-gazeradiologists
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This study investigates the potential of eye-tracking technology and the Segment Anything Model (SAM) to design a collaborative human-computer interaction system that automates medical image segmentation. We present the \textbf{GazeSAM} system to enable radiologists to collect segmentation masks by simply looking at the region of interest during image diagnosis. The proposed system tracks radiologists' eye movement and utilizes the eye-gaze data as the input prompt for SAM, which automatically generates the segmentation mask in real time. This study is the first work to leverage the power of eye-tracking technology and SAM to enhance the efficiency of daily clinical practice. Moreover, eye-gaze data coupled with image and corresponding segmentation labels can be easily recorded for further advanced eye-tracking research. The code is available in \url{https://github.com/ukaukaaaa/GazeSAM}.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs

    cs.CV 2025-08 conditional novelty 5.0 of 10

    GazeLT classifies long-tailed chest X-ray diseases by training a student model with a teacher that learns time-windowed radiologist gaze attention, improving tail-class accuracy on the NIH-CXR-LT and MIMIC-CXR-LT benchmarks.

  2. Zero-Shot Gaze-based Volumetric Medical Image Segmentation

    cs.CV 2025-05 conditional novelty 4.0 of 10

    Gaze-based prompts can drive zero-shot SAM-2 and MedSAM-2 segmentation of 3D abdominal CT organs, trading a roughly 0.08 Dice drop for an approximately 26 second per-organ speedup over manual bounding boxes.

Pith tools