Pith. sign in

Paper Citation Record · LEDGER

Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2311.17842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.17842 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:02:02.506842Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:07:47.839412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 39f06107-907f-4f92-88c2-74165d30b77e · inbound

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own cites this paper.

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-24T06:44:02.595306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T06:40:00.328012Z digest=sha256:7035fda19b9fd610c486720d93fafe907554c3f04daeacfd09b4c21d65cf1926

Observation 65853186-2c21-4a4d-9e2c-a1f6adcd845e · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.448860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:de7ebbb2e4a205284b2ee657fc0a18c64734e6e4182d1d47b1a401b6501c18cf

Observation 7d67daff-9c66-42ce-927d-03d170e48deb · inbound

ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation cites this paper.

ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:25:17.951151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:25:17.847571Z digest=sha256:b07ddea7797f5e191050c7bd4eecace3124ba606ca5afa11e867d094c8c11f8f

Observation 2b7ad28e-5358-40e8-8d1d-576be0bb51dc · inbound

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning cites this paper.

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:13:13.838808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T17:12:31.645346Z digest=sha256:80338c5f589b93e58f690280ae8bce8ef8bd437e128fcbeb9bad94fa9ea38aa5

Observation 11d633e9-7b8f-4858-8594-c2712d140d47 · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.957012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:a2d1d2ba04b0d0b1f440a4c59dc2e6967845f9c8988538d49394576520c7d741

Observation 5406b303-0436-4e37-9c41-6f9c14b71bb2 · inbound

Bilevel Learning for Bilevel Planning cites this paper.

Bilevel Learning for Bilevel Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:02.506842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:02.506842Z digest=sha256:d7ef07495e503f915294357b886477eef5e02e353191f5b838de32474e4e0114

Observation 4d482576-0e4f-46c0-b471-50f0e5da1b27 · inbound

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models cites this paper.

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:53:37.357899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T22:53:37.120692Z digest=sha256:9ae2644a7c9d140a4c2a77c3599ad18c755f236e2caf0273602bb50b2b6707cd

Observation 2f1cd2c7-f65e-4034-bc60-c64512065632 · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:05:00.897341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:9f1f4915306aa565e2a4c43245845a34e6cf79f9a2b40e1f0c54c7eb33547b0b

Observation 6d24e350-4701-49cd-a325-55a6b8d48022 · inbound

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead cites this paper.

Toward Embodied AGI: A Review of Embodied AI and the Road Ahead Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:08.979271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:08.979271Z digest=sha256:747830b77dfbd902dfcab39632f65520cc3101e46d9278625245e60ec5326b9f

Observation f5ed7ac3-7079-4303-a82c-676d6c11c2f8 · inbound

RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback cites this paper.

RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:12.043431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:10:12.043431Z digest=sha256:953cc167d081e888d69e09328f6995ec7fae98f23ad3c78b440f26e9fbe0bc6a

Observation 7020a0e8-4ae1-415f-b44e-3c122ea15438 · inbound

Learning Compositional Behaviors from Demonstration and Language cites this paper.

Learning Compositional Behaviors from Demonstration and Language Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:45.175493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:23:45.175493Z digest=sha256:25e910b6e6d4b154cee1489025952dd909ef8d7c72580227642152132867d4d3

Observation 306cb825-6fc5-49d0-b7ac-c2cf156d7404 · inbound

Reinforced Reasoning for Embodied Planning cites this paper.

Reinforced Reasoning for Embodied Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:22.543422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:22:22.543422Z digest=sha256:7b16775a9820590f8772d28ade64cbee8088637ff01f4a4d346b13e92326a4f9

Observation 5d3a4d0f-abc1-438d-8dbb-3e89c20b1864 · inbound

LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks cites this paper.

LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:51.749872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:51.749872Z digest=sha256:d89cef6f9a7a2429a2918ea8844b62d3d350dd4b04fc8de0e443a0331ab069fa

Observation 07e3eb45-38f2-4ad7-8cd8-c095b1e52707 · inbound

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models cites this paper.

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:39.572928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:39.572928Z digest=sha256:a940662f6caaf3f86fb79b9e505834feaea73a8e25e7b07b3eed2bc53e2cd3dc

Observation 5cd0f8ce-e481-46fb-9773-16ebf030187e · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:15.017442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:7e6975f6d2a6ddbd9d938f0ede504c253c9fec10af8d59a9eb0f24790a6bf5c3

Observation ba8792ae-56ba-418b-951d-0ee3d2da5641 · inbound

Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search cites this paper.

Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:02.032044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:02.032044Z digest=sha256:38128e70957e77ada81a51c0cdef7923031834a8e6b456938768d66e812dc00f

Observation f7a13579-681a-48dd-9bce-fc3b01b1feea · inbound

UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation cites this paper.

UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:53.946964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:53.946964Z digest=sha256:e7fdc056348288cf23cece05a5606867bb6b680111c92bd67d5b91224182725f

Observation ade70b3d-d5c4-4b51-8e03-abe6fbd8aba5 · inbound

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation cites this paper.

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:08.821795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:08.821795Z digest=sha256:362da234c2203ca389a8b777312ac645d0acdea72df6ead1326e482a0dfd5ac4

Observation 19db234a-03e3-41df-b5c5-02f98bd13859 · inbound

Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation cites this paper.

Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:20.537219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:20.537219Z digest=sha256:bdcb76a3c40fc279e32a38b07a66e1efff23bc2da07232df5daaf5fed9b6820b

Observation 42bb8f5a-c3de-46c7-b076-5a94690a86f0 · inbound

CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity cites this paper.

CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T23:43:16.734915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:43:16.734915Z digest=sha256:62fd88725bcc88fb746c940d3af6883945f31b9e7cda279f8e246276fa7234c7

Observation b8cd82c8-14e0-4c5a-9b39-57dd03ce7e62 · inbound

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models cites this paper.

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:03.288372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:03.288372Z digest=sha256:3a14ea9c63f01a8ea19b8d15fdf0567a4b3c92eb3fdfcbf959786f3e7423c161

Observation 0a8f2789-7dd0-4c41-8128-deecbcadba68 · inbound

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration cites this paper.

Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:46:00.599774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:46:00.599774Z digest=sha256:571481c750ad633c05e44d1119041c32561dc152fa33e6bf836f47cc472ccdd2

Observation 16401e03-e3f8-45d2-9f34-24e1d49b3917 · inbound

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning cites this paper.

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.811731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.811731Z digest=sha256:77d15fd447402d061d333babe596cf28aed49e6a5dfdf326823ff50385bb312d

Observation b9d28fc8-2aa9-4429-a214-c699ac2c1e70 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 154

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:15.975373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:dd4175b6b07efdde759d75ea755b2d5f8ad8b0eb4502a60ecfcba88abce3e1bc

Observation 88f04f39-f3ec-466f-96fc-c4487bb4d52b · inbound

Robix: A Unified Model for Robot Interaction, Reasoning and Planning cites this paper.

Robix: A Unified Model for Robot Interaction, Reasoning and Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:59:04.852667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:59:04.852667Z digest=sha256:58820947bd4f39f618ced6f75fd80531c42c8c6a8b03e1cc8030b6864bba6f56

Observation eef6cf35-1900-491b-af48-265d54f6a449 · inbound

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning cites this paper.

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:44:07.172848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T05:43:56.774051Z digest=sha256:67b7095a01ee597bdfb1ddb539a94920a83929aa0e2f4318750769d722acade6

Observation 82207850-e5fc-433f-a79a-024d82bb5538 · inbound

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning cites this paper.

SkillWrapper: Generative Predicate Invention for Task-level Robot Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T20:53:36.101266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T20:53:36.101266Z digest=sha256:d5713dba2a9dfffc0e38efa5726b9facb7bce15692d46e08ec6bd6f928e32e10

Observation 2444650f-add4-449c-b56b-1596fd0d5779 · inbound

Visual-Language-Guided Task Planning for Horticultural Robots cites this paper.

Visual-Language-Guided Task Planning for Horticultural Robots Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T09:58:55.434426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:58:55.434426Z digest=sha256:4600b8ded3b9495ebfb8f6353b035d2821efc78fb0a30996db181ecf08429ef9

Observation 3439fae1-6b7d-4e96-9eb4-d23c88f6cb61 · inbound

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making cites this paper.

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.850538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:46:03.360770Z digest=sha256:06257d3308a2c4b8a1b2a26254fa76cf7301271111d172afed86f657a6443f93

Observation c7dd11fd-762d-4521-9bdf-7601353feacc · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:15:56.948870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:ef0f06a2f68c8e8d85a90a2d3ea718b877b2bc784f5267ce7e712fddb30a71ea

Observation 51e22515-4b97-4443-9023-dd82c958fdbe · inbound

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model cites this paper.

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:09.735840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T11:45:18.081248Z digest=sha256:6b25bcd87d6cb49e45740685a44eddacfb1d3bc5a7f413ca4a690cbc91a1995c

Observation d8a1348e-396f-4c0f-83c8-50a3adf419fe · inbound

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning cites this paper.

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:16:20.158919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T15:39:17.505374Z digest=sha256:fbc45bb9dfa0445c88b69bbb3b08a9cc0a2f038b70c2b54840d994a26aa6970c

Observation ed4c4206-4600-4f69-b64f-27217592cb54 · inbound

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning cites this paper.

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:38:52.784788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T18:34:53.530617Z digest=sha256:211fe4392147f9708bf9fb9a6f6e9d1d1290d81a377caeff9498419ba3ea1e9a

Observation 751a4216-31ba-480e-a9f3-885af3d075eb · inbound

Make Your VLA More Robust Without More Data By Interleaving Motion Planning cites this paper.

Make Your VLA More Robust Without More Data By Interleaving Motion Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.528833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T17:31:55.850609Z digest=sha256:cc758f29539307fea5ff09cd484030dd0c899a5d917a5d7c11fdefe1c6f481e4

Observation e786e925-720b-4a77-9c11-7b5e4e7a59f0 · inbound

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning cites this paper.

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.901522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T09:59:02.899488Z digest=sha256:9ac8bc04f6d4ff159c1e7b18e2f668febc0fa1474166ac5e56dc171172fe08e1

Observation 7fb0d65e-c19a-4357-929e-f455fc3a5962 · inbound

SCOPE: Evolving Symbolic World for Planning in Open-Ended Environments cites this paper.

SCOPE: Evolving Symbolic World for Planning in Open-Ended Environments Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:49:42.289160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T10:57:26.511652Z digest=sha256:af3e2591c8b4e654fbc0a4b7eb6e7ee9ef391dc802c8d6715977c37d50269ad4

Observation d031e5f0-debe-4c7a-a968-1138a3798eba · inbound

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation cites this paper.

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:56.864505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T01:25:21.796778Z digest=sha256:5657bae2e9ca2c175a3f6f99b0d8b8fd9a1d125762fa4dbd90318f199a5ab0b7

Observation fee1f2eb-73c2-4d15-a7ff-2403307937aa · inbound

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation cites this paper.

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:05:49.683377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T00:54:50.393828Z digest=sha256:d7406487e426d8715930f3f3bc1220f3adc1a29069044b7a64e85a9c441a0ffc

Observation ce46f28f-3d72-4c2c-919a-5a38905c2fd7 · inbound

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review cites this paper.

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 113

Resolution
verified exact
local_arxiv, observed 2026-07-10T23:07:47.855321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T23:01:01.563768Z digest=sha256:2a3d7fa123b7c715576a083482a60fd03a3a03af585d6f85b2cfb3e0394bf117

Observation 40037c77-c19f-44a3-bbfa-d4c098a147bb · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 181

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:52.251951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:52.251951Z digest=sha256:d711098ce7147974220ba7a9cba2e36337b40cb6cbc6c47d01e4da5423f38bed

Observation 93c6a9b7-3b43-4137-8a71-094363a72d39 · inbound

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop cites this paper.

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T11:52:36.240433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:52:36.240433Z digest=sha256:46c0fb17d64d4d038f9ebdc56e8b2a18bf9060953cfcb47a09f1b254a58be3d6

Observation 980bd74d-6450-418b-b157-57ab0fa39523 · inbound

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models cites this paper.

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T04:59:32.314653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:59:32.314653Z digest=sha256:70bc225b9524f2aa038b6907c28fb840e9b9669128df7ba46b4dda619dac9020