Pith. sign in

Paper Citation Record · LEDGER

VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.13860.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13860 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:12:19.231195Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:59:20.317233Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation edcfabd1-07f9-4cd5-b156-2823a5aa0bc2 · inbound

Zero-Shot 3D Visual Grounding from Vision-Language Models cites this paper.

Zero-Shot 3D Visual Grounding from Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:19.231195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:19.231195Z digest=sha256:2c6b06490a861ea8cf42ac706fb9a60241345dbd45e7284dbe2313b50f8f437c

Observation 6b33626f-3e6d-49ac-a012-b6b4d3b5ab12 · inbound

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models cites this paper.

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:08.430220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:08.430220Z digest=sha256:acb544ac81b4182489b30222a49a91884e273e835d67d7cf297428b81edd1aa8

Observation e338d33b-6e96-42b4-afde-065288427dff · inbound

SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding cites this paper.

SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T14:52:56.561970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:52:56.561970Z digest=sha256:220efaf4d817769821022943a47b694cac2746b7ad280528128bdb0730fc9a8c

Observation f4b65a09-a4b9-479e-9008-e80826c52478 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:05:10.065051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T13:04:30.544504Z digest=sha256:40c2b720608833e0e25d1d1a946233dd6a5dbbb30fba9b0dfee6d7683d3ae2af

Observation 3354e406-cad3-44e4-a168-aa6be5881022 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T21:37:15.172508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:37:15.172508Z digest=sha256:b618b34781fa488e38ea35c6e8e6613a592ffc1d888e4090c23021ef59b5dd60

Observation 85c84e12-c257-4e9a-9a62-ff8967061ac8 · inbound

QuadAgent: A Responsive Agent System for Vision-Language Guided Quadrotor Agile Flight cites this paper.

QuadAgent: A Responsive Agent System for Vision-Language Guided Quadrotor Agile Flight VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.144096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:16:36.843013Z digest=sha256:228873229991d7f1e4a06077be9baeeae03661155a30d103a3a448547e4f8a88

Observation a173309e-a90b-4a2b-81a9-2be06ae79344 · inbound

Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding cites this paper.

Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:26.479613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T13:48:28.695094Z digest=sha256:0a88a9d2b80467d0c2a0893030e47490eaf2947fae32eadc147c30d16394f609

Observation a42499ad-265c-4a59-a797-7260600de9e4 · inbound

SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching cites this paper.

SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:51:18.439609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T08:47:25.712575Z digest=sha256:8bcc91d6bb24cf93e0e1ceb2f3471d37a824276347d0a956f8b06a4834a08d8f

Observation 2d930105-f2a8-4667-b03d-01ac17814941 · inbound

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning cites this paper.

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.672837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:47:52.739735Z digest=sha256:820d05f426fd1e3269c2b980318d96fc1e9203e14325fe367e7ce9df5367373c

Observation c656486a-9678-4d86-8d69-55b8f97fdd94 · inbound

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning cites this paper.

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:59:20.320423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T20:49:01.269513Z digest=sha256:fd5d114d901615bfedd2738b1c74c6f8f763ec0884b75a18ace89297d0697fca

Observation 75455b46-ea1f-4277-91c1-37e3a05e8699 · inbound

G$^2$TAM: Geometry Grounded Track Anything Model cites this paper.

G$^2$TAM: Geometry Grounded Track Anything Model VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T23:56:52.009530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:56:52.009530Z digest=sha256:12ffa6000324d3b4ef8780a49ebf91a628abaea5b90cc3132ffed2b4a47dcb8e

Observation 3adfee7b-12a5-4b11-9ef3-89a8da7c8d6c · inbound

TDVR: Joint Text Disambiguation and Viewpoint Reasoning for Zero-Shot 3D Visual Grounding cites this paper.

TDVR: Joint Text Disambiguation and Viewpoint Reasoning for Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T13:09:12.444281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:09:12.444281Z digest=sha256:1f72084e404bbb4b3cafea95ee324886cfa9955e49b0ce7020526ad9fc3fb445

Observation 174dfef7-f6eb-4c34-9be9-a3e11ae3c7dd · inbound

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching cites this paper.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.060750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.060750Z digest=sha256:77729d64aba3cdd902e5f43dfc1c73c601188050ba44b5b3b324e620889115b0