Pith. sign in

Paper Citation Record · LEDGER

Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2311.17842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.17842 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:44:50.818749Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:07:47.839412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 39f06107-907f-4f92-88c2-74165d30b77e · inbound

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own cites this paper.

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-24T06:44:02.595306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T06:40:00.328012Z digest=sha256:fc7bc9c0036b2215b16261e21a2b653fe16ad38405a9c890d2f8811dfd814024

Observation 65853186-2c21-4a4d-9e2c-a1f6adcd845e · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.448860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:dc794690fe88e200f308af63802361617ddf7d343a025e1477cad36db32c0d7c

Observation 7d67daff-9c66-42ce-927d-03d170e48deb · inbound

ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation cites this paper.

ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:25:17.951151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T08:25:17.847571Z digest=sha256:bcc96e4e9465d7b7689b1dae2046fee569fbf17922816c11936a82cc4d2bd2e7

Observation 2b7ad28e-5358-40e8-8d1d-576be0bb51dc · inbound

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning cites this paper.

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:13:13.838808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T17:12:31.645346Z digest=sha256:a0cde845b81a4e25412c6b070a2557f5b2a1ce37c897e5a42b2e5752ec5533c9

Observation eb7e01d9-409a-4a48-b9ef-a81b19f2ea95 · inbound

Integrating LMM Planners and 3D Skill Policies for Generalizable Manipulation cites this paper.

Integrating LMM Planners and 3D Skill Policies for Generalizable Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T22:44:50.818749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:44:50.818749Z digest=sha256:de5bcbb758f61b925e5645135fbfe556cbe851b74403813e7bb75904da9ca01a

Observation 11d633e9-7b8f-4858-8594-c2712d140d47 · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.957012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:9616fe3431d33d0fdf943b9bb1706b02529e9eb0983225a65c62f6a3f03a815b

Observation 5406b303-0436-4e37-9c41-6f9c14b71bb2 · inbound

Bilevel Learning for Bilevel Planning cites this paper.

Bilevel Learning for Bilevel Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:02.506842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:02.506842Z digest=sha256:d835d3a1d50bd21d8d0a10d1c4028a0697ce10d465a4b771082877129385b746

Observation 4d482576-0e4f-46c0-b471-50f0e5da1b27 · inbound

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models cites this paper.

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:53:37.357899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T22:53:37.120692Z digest=sha256:df7be58e52fcd4e72d4a2e83bd526747b3fd67e23571e7ad45190a61625f96c4

Observation 2f1cd2c7-f65e-4034-bc60-c64512065632 · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:05:00.897341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:0839e2319d8a6947f9d1ffc7b6ae6b49a646bc985296146f57d9d4ebc515ce11

Observation 6d24e350-4701-49cd-a325-55a6b8d48022 · inbound

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead cites this paper.

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:08.979271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:08.979271Z digest=sha256:747830b77dfbd902dfcab39632f65520cc3101e46d9278625245e60ec5326b9f

Observation f5ed7ac3-7079-4303-a82c-676d6c11c2f8 · inbound

RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback cites this paper.

RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:12.043431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:10:12.043431Z digest=sha256:953cc167d081e888d69e09328f6995ec7fae98f23ad3c78b440f26e9fbe0bc6a

Observation 7020a0e8-4ae1-415f-b44e-3c122ea15438 · inbound

Learning Compositional Behaviors from Demonstration and Language cites this paper.

Learning Compositional Behaviors from Demonstration and Language Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:45.175493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:23:45.175493Z digest=sha256:8b1bb7a6265c74ebf193eb54c42c41770a3646343860bd8ecfc21a3d34ea4c5d

Observation 306cb825-6fc5-49d0-b7ac-c2cf156d7404 · inbound

Reinforced Reasoning for Embodied Planning cites this paper.

Reinforced Reasoning for Embodied Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:22.543422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:22:22.543422Z digest=sha256:0c26a510b448edcddb54b4d0059067531598d7f619805f623e16ac5b72c2818c

Observation 5d3a4d0f-abc1-438d-8dbb-3e89c20b1864 · inbound

LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks cites this paper.

LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:51.749872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:51.749872Z digest=sha256:d89cef6f9a7a2429a2918ea8844b62d3d350dd4b04fc8de0e443a0331ab069fa

Observation 07e3eb45-38f2-4ad7-8cd8-c095b1e52707 · inbound

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models cites this paper.

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:39.572928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:39.572928Z digest=sha256:eb05a88d800e5ee472876f0146f80e3cd170218b84403b9cac034613c66c6afa

Observation 5cd0f8ce-e481-46fb-9773-16ebf030187e · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:15.017442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:48410cc44f92893cc775831859cb1383f840c7ce10045dca134f8e1019987b2b

Observation ba8792ae-56ba-418b-951d-0ee3d2da5641 · inbound

Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search cites this paper.

Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:02.032044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:02.032044Z digest=sha256:ca37774c8a333a6e78731bac13ce76208db70fec1eca1aae4b94f21b78afcc50

Observation f7a13579-681a-48dd-9bce-fc3b01b1feea · inbound

UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation cites this paper.

UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:53.946964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:53.946964Z digest=sha256:e7fdc056348288cf23cece05a5606867bb6b680111c92bd67d5b91224182725f

Observation ade70b3d-d5c4-4b51-8e03-abe6fbd8aba5 · inbound

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation cites this paper.

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:08.821795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:08.821795Z digest=sha256:362da234c2203ca389a8b777312ac645d0acdea72df6ead1326e482a0dfd5ac4

Observation 19db234a-03e3-41df-b5c5-02f98bd13859 · inbound

Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation cites this paper.

Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:20.537219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:20.537219Z digest=sha256:bdcb76a3c40fc279e32a38b07a66e1efff23bc2da07232df5daaf5fed9b6820b

Observation 42bb8f5a-c3de-46c7-b076-5a94690a86f0 · inbound

CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity cites this paper.

CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T23:43:16.734915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:43:16.734915Z digest=sha256:62fd88725bcc88fb746c940d3af6883945f31b9e7cda279f8e246276fa7234c7

Observation b8cd82c8-14e0-4c5a-9b39-57dd03ce7e62 · inbound

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models cites this paper.

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:03.288372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:03.288372Z digest=sha256:fdc46bf90f3b42383379d93f44558bc0b126a1ba8fdb84abcfa0fd019640de66

Observation 0a8f2789-7dd0-4c41-8128-deecbcadba68 · inbound

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration cites this paper.

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:46:00.599774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:46:00.599774Z digest=sha256:571481c750ad633c05e44d1119041c32561dc152fa33e6bf836f47cc472ccdd2

Observation 16401e03-e3f8-45d2-9f34-24e1d49b3917 · inbound

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning cites this paper.

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.811731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.811731Z digest=sha256:2c2b84f3ad350589e0c4e242fadbb7690e5335a3dc219e1e5afe6e5f5cf0d449

Observation b9d28fc8-2aa9-4429-a214-c699ac2c1e70 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 154

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:15.975373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:fc3f2272c7f471b31c1257a6fa1ed16c19580a02f7c19d79a3b2bfb02999e6d4

Observation 88f04f39-f3ec-466f-96fc-c4487bb4d52b · inbound

Robix: A Unified Model for Robot Interaction, Reasoning and Planning cites this paper.

Robix: A Unified Model for Robot Interaction, Reasoning and Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:59:04.852667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:59:04.852667Z digest=sha256:58820947bd4f39f618ced6f75fd80531c42c8c6a8b03e1cc8030b6864bba6f56

Observation eef6cf35-1900-491b-af48-265d54f6a449 · inbound

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning cites this paper.

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:44:07.172848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T05:43:56.774051Z digest=sha256:5cd70b1736ac629c995db69b5d92a3886c4dac087a0501a88723d134cf605be6

Observation 82207850-e5fc-433f-a79a-024d82bb5538 · inbound

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning cites this paper.

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T20:53:36.101266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T20:53:36.101266Z digest=sha256:d5713dba2a9dfffc0e38efa5726b9facb7bce15692d46e08ec6bd6f928e32e10

Observation 2444650f-add4-449c-b56b-1596fd0d5779 · inbound

Visual-Language-Guided Task Planning for Horticultural Robots cites this paper.

Visual-Language-Guided Task Planning for Horticultural Robots Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T09:58:55.434426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:58:55.434426Z digest=sha256:4600b8ded3b9495ebfb8f6353b035d2821efc78fb0a30996db181ecf08429ef9

Observation 3439fae1-6b7d-4e96-9eb4-d23c88f6cb61 · inbound

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making cites this paper.

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.850538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T00:46:03.360770Z digest=sha256:87b9888e8cc25809e9f6f1e884790934ed7529e4f8a028ec357143f4a860875f

Observation c7dd11fd-762d-4521-9bdf-7601353feacc · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:15:56.948870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:2ef5beba0ddd11fd821a9ab8257961705fb2f83c620d6518617d9deb666d130f

Observation 51e22515-4b97-4443-9023-dd82c958fdbe · inbound

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model cites this paper.

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:09.735840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T11:45:18.081248Z digest=sha256:43ef776807c3cffd12b6d34599847ea154398c758d9b2a382ba07e877206269c

Observation d8a1348e-396f-4c0f-83c8-50a3adf419fe · inbound

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning cites this paper.

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:16:20.158919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T15:39:17.505374Z digest=sha256:a04880e0d6835306de92d7ca1ed1a6772c3c6894568d1fa967e6a46d5d61da21

Observation ed4c4206-4600-4f69-b64f-27217592cb54 · inbound

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning cites this paper.

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:38:52.784788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T18:34:53.530617Z digest=sha256:e9851e478c2825023d61341498e61c226a9bb459d4a34e280b3242a29858da38

Observation 751a4216-31ba-480e-a9f3-885af3d075eb · inbound

Make Your VLA More Robust Without More Data By Interleaving Motion Planning cites this paper.

Make Your VLA More Robust Without More Data By Interleaving Motion Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.528833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T17:31:55.850609Z digest=sha256:a5fa7f150615d90cc5c21644c6028c7ab82e430b848aa5f0d2d57fbb6e7a84c3

Observation e786e925-720b-4a77-9c11-7b5e4e7a59f0 · inbound

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning cites this paper.

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.901522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T09:59:02.899488Z digest=sha256:dc44283cb50df69af5aec3a2749859c2762907ec7fefeaf1619041335f5bd834

Observation 7fb0d65e-c19a-4357-929e-f455fc3a5962 · inbound

SCOPE: Evolving Symbolic World for Planning in Open-Ended Environments cites this paper.

SCOPE: Evolving Symbolic World for Planning in Open-Ended Environments Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:49:42.289160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T10:57:26.511652Z digest=sha256:37ff001de0b4345ad4ec4e945ba69cd3189fe32d1b95a0b8cdc7c87fb5f26dde

Observation d031e5f0-debe-4c7a-a968-1138a3798eba · inbound

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation cites this paper.

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:56.864505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T01:25:21.796778Z digest=sha256:8d19a7a8ef317b4aeddb3e51062b67adb3661f8ea16858ab7b56be0712c0834f

Observation fee1f2eb-73c2-4d15-a7ff-2403307937aa · inbound

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation cites this paper.

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:05:49.683377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T00:54:50.393828Z digest=sha256:7738175b236a9b2d2ce54b5dd3ded8d87c7222b19ac7a681fcfcafb4aeed31ae

Observation ce46f28f-3d72-4c2c-919a-5a38905c2fd7 · inbound

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review cites this paper.

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 113

Resolution
verified exact
local_arxiv, observed 2026-07-10T23:07:47.855321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T23:01:01.563768Z digest=sha256:722683ed1afba149bf10c01b5404c904e6d5e7bb4bd994428382729074ea69d2

Observation 40037c77-c19f-44a3-bbfa-d4c098a147bb · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 181

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:52.251951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:52.251951Z digest=sha256:d711098ce7147974220ba7a9cba2e36337b40cb6cbc6c47d01e4da5423f38bed

Observation 93c6a9b7-3b43-4137-8a71-094363a72d39 · inbound

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop cites this paper.

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T11:52:36.240433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:52:36.240433Z digest=sha256:46c0fb17d64d4d038f9ebdc56e8b2a18bf9060953cfcb47a09f1b254a58be3d6

Observation 980bd74d-6450-418b-b157-57ab0fa39523 · inbound

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models cites this paper.

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T04:59:32.314653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:59:32.314653Z digest=sha256:70bc225b9524f2aa038b6907c28fb840e9b9669128df7ba46b4dda619dac9020