Pith. sign in

Paper Citation Record · LEDGER

Segment and Track Anything

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2305.06558.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.06558 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:11:06.887432Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:13:49.304240Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 99305b33-8589-4c07-9df4-c17ffe475cc8 · inbound

DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory cites this paper.

DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory Segment and Track Anything

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:03:58.147228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T13:03:57.828598Z digest=sha256:d2695b6c2c398f9fa36c55f70e1cacb6c619aa5d2fa423d79b60146a26084220

Observation 0ff62b57-678d-45e7-b2a5-9c5b361a04b4 · inbound

SAM 2: Segment Anything in Images and Videos cites this paper.

SAM 2: Segment Anything in Images and Videos Segment and Track Anything

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:56:25.371685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:56:25.331304Z digest=sha256:334dda93270c91affeed352a1827af304647696052f53b819b6a12f91faf9449

Observation aad9540f-c6dd-42f1-a3a4-4ce68d8109d5 · inbound

On Efficient Variants of Segment Anything Model: A Survey cites this paper.

On Efficient Variants of Segment Anything Model: A Survey Segment and Track Anything

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:43:23.490908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T19:42:24.122342Z digest=sha256:f7aab1c1aea62dd1647af05fe8457f9d57ba61bd173aca8fa7eec593854daa10

Observation dbfacbdb-dc75-4c8f-93df-277c442ea72b · inbound

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement cites this paper.

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement Segment and Track Anything

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T08:25:29.380265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T08:25:01.468957Z digest=sha256:9d5fbe021dc6475561274176bc5acb1a6cb13b017664c09a9813e98bfaee835a

Observation bca115e4-1bfd-4dd2-9908-fdd7d2fa1fb5 · inbound

SAM-guided Pseudo Label Enhancement for Multi-modal 3D Semantic Segmentation cites this paper.

SAM-guided Pseudo Label Enhancement for Multi-modal 3D Semantic Segmentation Segment and Track Anything

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T17:11:06.887432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:11:06.887432Z digest=sha256:2f63d3e706c81f01c2e73ea3da3e29b7cebb6ace4ff9d44bf0917b3f86cb8bbc

Observation ed34e449-ddcb-432b-8bfb-19ae0bc3e577 · inbound

ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models cites this paper.

ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models Segment and Track Anything

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T05:23:20.253205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:23:20.253205Z digest=sha256:92afe7fcb30806f209d0af75376de4f2550899230318d35541137d7107480630

Observation aada70c0-8b73-4e17-80ac-3f01202fb14d · inbound

SAMRefiner: Taming Segment Anything Model for Universal Mask Refinement cites this paper.

SAMRefiner: Taming Segment Anything Model for Universal Mask Refinement Segment and Track Anything

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:29:00.503655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:29:00.503655Z digest=sha256:53ddf263dbe67dedc77cf77c78e82005223970ec7cf14defebb0292a1014ad64

Observation 3326b1ed-f614-48c3-a81e-97ed8d14275c · inbound

COMBO-Grasp: Learning Constraint-Based Manipulation for Bimanual Occluded Grasping cites this paper.

COMBO-Grasp: Learning Constraint-Based Manipulation for Bimanual Occluded Grasping Segment and Track Anything

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T11:04:52.457436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:04:52.457436Z digest=sha256:b0bbb1281a2fa52a88a6deb2b5bc336e012ab9378511d5f64344feede87ff3aa

Observation f0496b03-a06d-45e7-a7f5-e8825b1c0182 · inbound

Hierarchical Instruction-aware Embodied Visual Tracking cites this paper.

Hierarchical Instruction-aware Embodied Visual Tracking Segment and Track Anything

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:24.826217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:24.826217Z digest=sha256:f236cfb4dc16651c416e312a1d0f29eb15c3191f3583fd03d47d9ba7636a053f

Observation 9b1cc4fd-526c-497a-b21e-a3484ac88c49 · inbound

SAM-I2V: Upgrading SAM to Support Promptable Video Segmentation with Less than 0.2% Training Cost cites this paper.

SAM-I2V: Upgrading SAM to Support Promptable Video Segmentation with Less than 0.2% Training Cost Segment and Track Anything

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:52:38.984260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:52:38.984260Z digest=sha256:9ec361fab1db6269028439963c0965d85a3f6a81d412beb33c57b7cb2e49166c

Observation 29f7552e-8ec6-4694-ac33-60863cf5c5c7 · inbound

LayerFlow: A Unified Model for Layer-aware Video Generation cites this paper.

LayerFlow: A Unified Model for Layer-aware Video Generation Segment and Track Anything

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:50.093556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:50.093556Z digest=sha256:707df3e67a3ce96b06b6bf456039d5ea8a7b5e60476444bb31bc1a5a4b7692af

Observation 917015bf-08bc-43be-87e3-d8b9bc2d7aa5 · inbound

Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos cites this paper.

Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos Segment and Track Anything

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:09.287312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:09.287312Z digest=sha256:d525b9c0ebb8215446ed1b931451d4d8f98cb3cd15b03d1af9e61fa06d70a650

Observation 736f90f5-e8ac-4755-8c8e-9cca77784074 · inbound

A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects cites this paper.

A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects Segment and Track Anything

Reference 164

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:20.234594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:34:20.234594Z digest=sha256:0b6292a351727e6225304cce2d797f24f65ac13193e52293e8e8546cf20879a0

Observation 9fb9494b-f604-4bec-bb6a-b100b8452b68 · inbound

STR-Match: Matching SpatioTemporal Relevance Score for Training-Free Video Editing cites this paper.

STR-Match: Matching SpatioTemporal Relevance Score for Training-Free Video Editing Segment and Track Anything

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:58:20.179134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:58:20.179134Z digest=sha256:6d0a33b4690854c8b8d8891eee60d4e8a1153f8255f72f2c402b133ad261ec57

Observation ee9369df-9b15-406f-94ae-1f7bc5d66af0 · inbound

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation cites this paper.

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation Segment and Track Anything

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:53:32.123934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:53:32.123934Z digest=sha256:9d978bf7e3b36831287d3fa5634bdbe44bcb5cf6679944d981f9c63737e64384

Observation fe661157-e11b-41fb-83bf-d5599062b8c6 · inbound

ViRefSAM: Visual Reference-Guided Segment Anything Model for Remote Sensing Segmentation cites this paper.

ViRefSAM: Visual Reference-Guided Segment Anything Model for Remote Sensing Segmentation Segment and Track Anything

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:40:06.711380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:40:06.711380Z digest=sha256:896457f1359327483f1d0fbe71efbf46bd7e5664071350f8208547a914effc6d

Observation 0dc380bd-5ec3-4a86-9b19-408a2abb5427 · inbound

CrowdTrack: A Benchmark for Difficult Multiple Pedestrian Tracking in Real Scenarios cites this paper.

CrowdTrack: A Benchmark for Difficult Multiple Pedestrian Tracking in Real Scenarios Segment and Track Anything

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:50.844918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:31:50.844918Z digest=sha256:f4a3f7ade5fc6aac11dac8f3396cbd7fae147196d14ad21aeb0957f65cb2e51f

Observation 9cdf3218-26e9-44f6-b31c-a370dfcdc5a4 · inbound

High-fidelity 3D Gaussian Inpainting: preserving multi-view consistency and photorealistic details cites this paper.

High-fidelity 3D Gaussian Inpainting: preserving multi-view consistency and photorealistic details Segment and Track Anything

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T14:44:58.782053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:44:58.782053Z digest=sha256:9a7d76ff2f1725adfac2225afd2836354595ea2948be64a3e1765a33ab5b0724

Observation e23316d3-68b9-4291-b465-fdcda0340cb3 · inbound

Grouped Speculative Decoding for Autoregressive Image Generation cites this paper.

Grouped Speculative Decoding for Autoregressive Image Generation Segment and Track Anything

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:57:09.363139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:57:09.363139Z digest=sha256:bca835d0facff783b420458613d748afe2cf1deb1b7986a0583e4912226730ee

Observation 538a1d52-85c3-4e25-be6c-d8fa0183b243 · inbound

Representative Volume Element: Existence and Extent in Cracked Heterogeneous Medium cites this paper.

Representative Volume Element: Existence and Extent in Cracked Heterogeneous Medium Segment and Track Anything

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:29:52.721427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:29:52.721427Z digest=sha256:d57fbbf1ce9c186f7596aa88fef80e79cb2355b22a59fe0953f347c7907b6f8d

Observation f9e0e54b-bce0-41ea-a5b5-a435ff99900e · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Segment and Track Anything

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.664360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:91435c602c6b940ae06b5aef2f09ff4029af748c55ba654164d5e691d9b5f9ef

Observation bb55ce6e-db0c-41a4-9ca2-2c958af88378 · inbound

DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing cites this paper.

DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing Segment and Track Anything

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T18:34:23.512742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:34:23.512742Z digest=sha256:d1b00848453a5f944f6a6b9ba64955ae49ccaf8e02e16468553a3865924c6cea

Observation d3d34e96-2a6a-4207-a531-14fd3f85fee9 · inbound

VoCap: Video Object Captioning and Segmentation from Any Prompt cites this paper.

VoCap: Video Object Captioning and Segmentation from Any Prompt Segment and Track Anything

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T14:01:16.517409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:01:16.517409Z digest=sha256:54eb55962222612325058545679f8b6f2652570f26005eb456dc0620d975e8ea

Observation 22ef0701-514c-41be-a049-6a883b793cc2 · inbound

Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control cites this paper.

Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control Segment and Track Anything

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T00:05:11.934372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:05:11.934372Z digest=sha256:8ca2d6827456c49f74a0879628be08abab72c36bc4f7db4d88fa5618f2554ea4

Observation 31a06d06-832b-4879-b307-bd7cd3992c96 · inbound

Grasp Like Humans: Learning Generalizable Multi-Fingered Grasping from Human Proprioceptive Sensorimotor Integration cites this paper.

Grasp Like Humans: Learning Generalizable Multi-Fingered Grasping from Human Proprioceptive Sensorimotor Integration Segment and Track Anything

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:44:40.124432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:44:40.124432Z digest=sha256:bf999e315cbdfa7dc4d04930cde3dc013ad968505dad012ce411708f38e076d2

Observation 51f54a4b-81d8-4b67-b2c5-cbcaa0599749 · inbound

ViSTR-GP: Online Cyberattack Detection via Vision-to-State Tensor Regression and Gaussian Processes in Automated Robotic Operations cites this paper.

ViSTR-GP: Online Cyberattack Detection via Vision-to-State Tensor Regression and Gaussian Processes in Automated Robotic Operations Segment and Track Anything

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T17:25:30.813136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:25:30.813136Z digest=sha256:087a345004b49a648f463d0a1be511af39ca49d34a67222c949710b57e9795de

Observation 8f735f20-1c5b-47cb-8360-7803ec16e0e9 · inbound

Reinforcement Learning for Unsupervised Domain Adaptation in Spatio-Temporal Echocardiography Segmentation cites this paper.

Reinforcement Learning for Unsupervised Domain Adaptation in Spatio-Temporal Echocardiography Segmentation Segment and Track Anything

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:56:01.667572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T06:53:21.438159Z digest=sha256:2e6743339411feb7831d21f66038f0e91b8125bb93708a56e15021d459aa71d8

Observation 638f907f-54be-4905-87a0-ba068f0665b0 · inbound

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents cites this paper.

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents Segment and Track Anything

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:58:31.814788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T20:56:58.770875Z digest=sha256:00b7869dfeee5f119531158e33e7762a180645902c244c8a2f236c32a9d0a110

Observation c67fe298-7b32-47f0-aff3-30da96baf5c1 · inbound

Efficient Segment Anything with Depth-Aware Fusion and Limited Training Data cites this paper.

Efficient Segment Anything with Depth-Aware Fusion and Limited Training Data Segment and Track Anything

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T00:02:11.238949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:02:11.238949Z digest=sha256:a45a97654faa98593db13864dc07783bd05ecc768681816694035ada8913cc7a

Observation 8ec5f0ab-ef26-47f4-8ee7-153c2e331ded · inbound

ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis cites this paper.

ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis Segment and Track Anything

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T18:24:58.428701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:24:58.428701Z digest=sha256:bd771d5cd6b7946848e8b1fef7665e6f4ca96bff75dc31822f09c556025b992f

Observation 521aac84-4e08-43cf-ad2c-a24e5a4ea3b5 · inbound

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models cites this paper.

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models Segment and Track Anything

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:10:53.007612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:39:26.420463Z digest=sha256:76e7edd4ee8ab9450dcf6f38316a8a7d7ccb3725b8f86280b06115f174fe71a7

Observation b755e582-44c9-4a1b-85f9-40414dcdad08 · inbound

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking cites this paper.

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking Segment and Track Anything

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:34:47.308672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:34:26.106672Z digest=sha256:bb9ec957f9a15e5fef5b290a164dbb2ef675e97df91c27fa3054f6c8b1413154

Observation 119424ae-d507-4a02-a73c-9eb699106c1d · inbound

One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition cites this paper.

One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition Segment and Track Anything

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.807561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T08:50:27.871886Z digest=sha256:eb75fd2701035f328d6555b1750f696cd341e4d4171003de7a975d777aeef25a

Observation 456f231d-bb91-41ed-b7cc-04ac797e0193 · inbound

Temporal-Emerged Prompting for Segment Anything in Multiframe Infrared Small Target Detection cites this paper.

Temporal-Emerged Prompting for Segment Anything in Multiframe Infrared Small Target Detection Segment and Track Anything

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:49.306052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T05:13:51.469018Z digest=sha256:7fdbfae584efe664e6eb541e1429191151f6c7d0d65da75d3771be4a618b794f

Observation bfbb65b8-fc5d-4102-8d79-1dc80c1f94fd · inbound

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation cites this paper.

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation Segment and Track Anything

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T00:48:38.659879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:48:38.659879Z digest=sha256:1075bb42d5ec8c1596137fec7e5e65001faf535df3eb2c26267348ac983eed49

Observation 48fc112f-e9ba-41a2-acfd-e23884ca5a27 · inbound

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation cites this paper.

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation Segment and Track Anything

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:49:54.922688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:49:54.922688Z digest=sha256:858b2b0d87b60e2a8641466ab0c0a523bda4430010935c65efb9d805f9429c08