Pith. sign in

Paper Citation Record · LEDGER

Evaluating Vision-Language Models as Evaluators in Path Planning

As of 17 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 3 inbound Pith citation observations for arXiv:2411.18711.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.18711 v4

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:04:44.887611Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:49:07.509682Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T22:49:11.112158Z

Reference resolution

100 of 100 outbound references displayed

  • verified exact1
  • verified fuzzy59
  • unresolved39
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ade9c557-2325-44ab-8d0f-3a47a0aadaec · outbound

This paper cites Faithfulness vs.

Evaluating Vision-Language Models as Evaluators in Path Planning Faithfulness vs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.468322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.468322Z digest=sha256:89bbe94658c677944ab3818d207f1749141f201b34e1b1e2c011525bbc7d5bd8

Observation 01e605a1-7603-43b4-a035-4aa5f33450ed · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

Evaluating Vision-Language Models as Evaluators in Path Planning Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.473526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.473526Z digest=sha256:08255729498170c17eca5b5ed27f33c4b3974578747ba5981fba6efd106e2939

Observation 3761979d-6e9a-4f2d-a81c-6dd3208b7f74 · outbound

This paper cites Look Further Ahead: Testing the Limits of GPT-4 in Path Planning.

Evaluating Vision-Language Models as Evaluators in Path Planning Look Further Ahead: Testing the Limits of GPT-4 in Path Planning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.477902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.477902Z digest=sha256:1485517ced74456837bee4e871dc036dd145234748ed0bc64d93ebbe8af91bf1

Observation 8092616c-44de-4283-abe3-3c698262016e · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

Evaluating Vision-Language Models as Evaluators in Path Planning Lawrence Zitnick, and Devi Parikh

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.482092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.482092Z digest=sha256:a958fad6b635f8d588b3e3277eef02f1b48fce70c71da528266a14ff873bf30e

Observation 233da0ae-c88f-4c80-8869-17a93b546bb7 · outbound

This paper cites Vision- Language Models as a Source of Rewards, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Vision- Language Models as a Source of Rewards, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.486867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.486867Z digest=sha256:7d45b17d69e76a54469f262f33cb46f6545d5bf13ccc27262fbe97805f42edba

Observation 0c1a60d7-f548-4277-a792-8c5e1d28bd18 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.491249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.491249Z digest=sha256:c53323676d09c7e4befd430ed088c675739e3ffc4dda9baeec9a0b8ae7976bc6

Observation f4f330c3-da50-4450-ae7f-513c72441a5d · outbound

This paper cites Rehg, and Chao Zheng.

Evaluating Vision-Language Models as Evaluators in Path Planning Rehg, and Chao Zheng

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.496325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.496325Z digest=sha256:efecfbbf230652c22ff2a4e0a38a1a59eb706de662fd6ba3e5dc926967627210

Observation 32f0e98a-2ad7-40d1-932b-2028269b1a11 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Evaluating Vision-Language Models as Evaluators in Path Planning Emerg- ing properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.500853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.500853Z digest=sha256:40becf74521ea3a31b78fb2a5ba380b62e275fc3c2bb265048fbd5ed2425f0fc

Observation 7b5b0e98-dc65-4bf4-a698-34cf5f8e319d · outbound

This paper cites MapGPT: Map- guided prompting with adaptive path planning for vision- and-language navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning MapGPT: Map- guided prompting with adaptive path planning for vision- and-language navigation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.505242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.505242Z digest=sha256:2ca78113f7e4e774dcfd12157dabdbaea7403d4df12141a7de67ad17e39a599c

Observation ca82eb8c-1164-4336-9a24-8d3c71d6786f · outbound

This paper cites Multi-object hallucination in vision language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Multi-object hallucination in vision language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.509313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.509313Z digest=sha256:e335f094829cc5d30c5e522210c103a64f1f61d8300b4b0a8651d708cc1b3d86

Observation f65a7ef2-b024-492b-afa2-d58e5cfa6dc7 · outbound

This paper cites Autotamp: Autoregressive task and motion planning with llms as translators and check- ers.

Evaluating Vision-Language Models as Evaluators in Path Planning Autotamp: Autoregressive task and motion planning with llms as translators and check- ers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.513378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.513378Z digest=sha256:a5f9159052c29957d3ab218f4f153462bc8af39da740f4af3087c09a7ee0c69c

Observation 5b82d06f-ced8-47e9-bdd9-159ceb70355f · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Evaluating Vision-Language Models as Evaluators in Path Planning InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.517807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.517807Z digest=sha256:52ef8f5fe3798f107e6c069747a4706921711b1d0c063e2366119ca85db34fad

Observation d99cc4f6-441a-4bd6-8385-0c5c2d47e72c · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

Evaluating Vision-Language Models as Evaluators in Path Planning Gonzalez, Ion Stoica, and Eric P

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.522611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.522611Z digest=sha256:aee767324e8dcd421a2f6fc21a8a8600316402b7af593602ff1385e732407318

Observation 587a1600-5f7f-49e0-b3d3-e84220ad8e7d · outbound

This paper cites Task and motion planning with large language models for ob- ject rearrangement.

Evaluating Vision-Language Models as Evaluators in Path Planning Task and motion planning with large language models for ob- ject rearrangement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.526554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.526554Z digest=sha256:9c64c7394f4861fc88144611f6f1c6b0c8ea60170ff0b9f5dcf1bae5d8c9bd0d

Observation 181dd57e-4680-4619-b038-02ec2e330096 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.530668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.530668Z digest=sha256:4983c08909f5b624095f56b6e98ac7d353502859a7c9c075d7f7daae18cb802b

Observation 453e7994-d40c-41bf-a3f0-180d6044a2f3 · outbound

This paper cites Tenenbaum, Leslie Pack Kaelbling, Andy Zeng, and Jonathan Tompson.

Evaluating Vision-Language Models as Evaluators in Path Planning Tenenbaum, Leslie Pack Kaelbling, Andy Zeng, and Jonathan Tompson

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.534874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.534874Z digest=sha256:a5d674e673f85442b66c3bb7d5ce9e453867ac2afb868560c647a1c5b5b405d1

Observation 58d9f6c3-7cd2-4d09-b1e9-c2e0f2cfcac8 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.538679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.538679Z digest=sha256:f80de624f72c7e6a45e715146bbb78f3c4e80c12a61724bcce1a95a5d2506bfe

Observation 36974c1e-8970-4e9f-89ad-4bac5a9e97a2 · outbound

This paper cites Cric: A vqa dataset for compositional reasoning on vision 9 and commonsense.

Evaluating Vision-Language Models as Evaluators in Path Planning Cric: A vqa dataset for compositional reasoning on vision 9 and commonsense

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.542945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.542945Z digest=sha256:1955e6672889759151c73db3bc70653e26fc47825a31f3e5f9ad9f5d22a88142

Observation fc792dc8-6b55-4f65-89b0-bfdb299ccfe9 · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing.

Evaluating Vision-Language Models as Evaluators in Path Planning Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.547164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.547164Z digest=sha256:b529f6f56bd4ee5c0484e43528188397b31c4cf325f4fc6faf729d15efed535a

Observation a3cf708c-c8df-4f20-ab5b-fc86919d52ee · outbound

This paper cites Vision-and-language navigation: A survey of tasks, methods, and future directions.

Evaluating Vision-Language Models as Evaluators in Path Planning Vision-and-language navigation: A survey of tasks, methods, and future directions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.551234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.551234Z digest=sha256:2a584f34042ce706bd0d88f0a75af58e522a475fe08184c1fe7fa36eacc7b180

Observation 9d340550-7e46-4412-ab55-45dd5cddecc9 · outbound

This paper cites Task success is not enough: Investigating the use of video-language models as behavior critics for catching undesirable agent behaviors.

Evaluating Vision-Language Models as Evaluators in Path Planning Task success is not enough: Investigating the use of video-language models as behavior critics for catching undesirable agent behaviors

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.554854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.554854Z digest=sha256:5ae7e99190b7dea1976d3124caef4d1faf4c17805b3e6384abd7d48e9b75b745

Observation 6fd427a0-91c6-442d-804a-4ff8ed4cb593 · outbound

This paper cites Hal- lusionbench: An advanced diagnostic suite for entangled language hallucination & visual illusion in large vision- language models, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Hal- lusionbench: An advanced diagnostic suite for entangled language hallucination & visual illusion in large vision- language models, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.178934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.558971Z digest=sha256:dcea35073ff95cf07f68852c876cb3dd8dafa8c7e7baf160afc9c91b46c43b30

Observation 16a0fdbb-a7fb-440a-b308-e0cd982e0b36 · outbound

This paper cites Generating and evolving reward functions for highway driving with large language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Generating and evolving reward functions for highway driving with large language models, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.164228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.563253Z digest=sha256:784af2955a00b9750dea074791c3b12d16da5be2aa2ed18e6bc89f87ce3b478d

Observation da16eae5-a81d-4993-9316-ebe64dc9fb09 · outbound

This paper cites Hudson and Christopher D.

Evaluating Vision-Language Models as Evaluators in Path Planning Hudson and Christopher D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.149923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.567498Z digest=sha256:1ccc27e6701a056ae8b516c68639c915e3c4830b3bed74bb8b362751e92d70d3

Observation 751f3513-7d07-40c2-a2ec-6a1309b9c47e · outbound

This paper cites Open- clip, 2021.

Evaluating Vision-Language Models as Evaluators in Path Planning Open- clip, 2021

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.134842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.571709Z digest=sha256:7a77a442215de191a8753d2a5867e1d99201a8f233260d2de44ed586eff26908

Observation 104ad04a-0b97-4ab9-97f6-ae2a612b739d · outbound

This paper cites What’s ”up” with vision-language models? Investigating their strug- gle with spatial reasoning.

Evaluating Vision-Language Models as Evaluators in Path Planning What’s ”up” with vision-language models? Investigating their strug- gle with spatial reasoning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.121589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.575741Z digest=sha256:ab977f3de7013840b43b6c317d049992d4253c2b134776037616593b953ab61b

Observation 7f3b778d-2e92-45b4-b4a3-c618f13bdf27 · outbound

This paper cites Position: LLMs Can’t Plan, But Can Help Planning in LLM-Modulo Frameworks.

Evaluating Vision-Language Models as Evaluators in Path Planning Position: LLMs Can’t Plan, But Can Help Planning in LLM-Modulo Frameworks

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.107670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.579754Z digest=sha256:30caca9e1b0d030a690c0fc8e21e111eb69c7026cc777e4687990c7f5c7d6745

Observation f746ba24-f58f-4904-a87b-75630d9f239d · outbound

This paper cites Kavraki, P.

Evaluating Vision-Language Models as Evaluators in Path Planning Kavraki, P

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.093809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.584135Z digest=sha256:830ef8dc113183e40a347a8d03ff5e51c060d9c1ea1ca05272ae7fd7067c2800

Observation c5920d40-9111-482c-8d24-0b7ee755258c · outbound

This paper cites Kuffner and S.M.

Evaluating Vision-Language Models as Evaluators in Path Planning Kuffner and S.M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.080404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.588280Z digest=sha256:5aa70bd2b1a8048acb98b5f2c5b48eb24da1489dff84461f534f553aba429419

Observation 8a08d27a-0f30-4d6b-89e9-d0d04968ecaa · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.592439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.592439Z digest=sha256:12f1538f958e32eff44969c9cd986859648e6041e83d9d98e5936d98ccb11483

Observation 72860de6-2c21-4c95-bc23-a7976209dd30 · outbound

This paper cites Why i’m optimistic about our align- ment approach: Evaluation is easier than generation.

Evaluating Vision-Language Models as Evaluators in Path Planning Why i’m optimistic about our align- ment approach: Evaluation is easier than generation

Reference 31

Resolution
verified exact
raw_fallback, observed 2026-08-12T11:04:45.048905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.596527Z digest=sha256:f340960d5ea93dc3d6b292798b9f557b04b1c75c2c9d7a5019313e321acdce4a

Observation d63578c3-077e-4918-a307-b20b62015e44 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Evaluating Vision-Language Models as Evaluators in Path Planning LLaVA-OneVision: Easy Visual Task Transfer

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.605211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.605211Z digest=sha256:daa04d4db287db272d58a48c6132929aee7488095e83e11f3b05f1240359fb10

Observation ed81fadc-07f5-49df-801b-2f5cd7737962 · outbound

This paper cites Auto mc-reward: Automated dense reward design with large language models for minecraft.

Evaluating Vision-Language Models as Evaluators in Path Planning Auto mc-reward: Automated dense reward design with large language models for minecraft

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.045353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.610072Z digest=sha256:461ba652ed8a3b7ad60db2e80683129b03af942a1a3b29cc06f255031c68ae8a

Observation 28305202-3f6b-485e-9d9c-342077825d35 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Evaluating object hallucination in large vision-language models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.614434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.614434Z digest=sha256:ca7c3b86407b9d864bd9041e3510c2b979c476ddf425a45cc4752b3e6f78e069

Observation 7c5da2e7-6986-4049-aac8-8976bd59f1b2 · outbound

This paper cites OmniBench: Towards The Future of Uni- versal Omni-Language Models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning OmniBench: Towards The Future of Uni- versal Omni-Language Models, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.021161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.618780Z digest=sha256:bd176012933b218a2c11c0548c1cd74ee2b483696ba2c773c329c85cae1ca1db

Observation 69284036-e248-469c-85fc-d829989cd295 · outbound

This paper cites Chin, Shuhong Chai, Neil Bose, and Eonjoo Kim.

Evaluating Vision-Language Models as Evaluators in Path Planning Chin, Shuhong Chai, Neil Bose, and Eonjoo Kim

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.007403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.622940Z digest=sha256:10948a8425568aed02d00dd998eeb6c0be2099e913e04a201f0774cc4beb302d

Observation 0c3e868d-fd36-49f4-8691-3b309339f336 · outbound

This paper cites Visual instruction tuning.

Evaluating Vision-Language Models as Evaluators in Path Planning Visual instruction tuning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.626964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.626964Z digest=sha256:3de980b652ea0ee645e9e4b250e43349047c61d634781b425cec8bffc94ab6c4

Observation 4a300cf0-fe05-4baf-8e0e-f722440854cb · outbound

This paper cites Improved Baselines with Visual Instruction Tuning, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Improved Baselines with Visual Instruction Tuning, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.984335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.631016Z digest=sha256:b9a02f1571d9064c7c66192e4bc7a52cee10f97c64ac1751e95dd853cbd9f9de

Observation c5870301-fb45-40dc-bf24-19ac27975806 · outbound

This paper cites LLaV A-NeXT: Im- proved reasoning, OCR, and world knowledge, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning LLaV A-NeXT: Im- proved reasoning, OCR, and world knowledge, 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.970141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.635085Z digest=sha256:97d155342f84807b6cb62c32c4036c69c9c8bddbda8d73d6e9f1b55498b1c8d7

Observation 3534786f-a71e-428d-8782-29cc05eb01c3 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? In Computer Vi- sion – ECCV 2024 , pages 216–233, Cham, 2025.

Evaluating Vision-Language Models as Evaluators in Path Planning Mmbench: Is your multi-modal model an all-around player? In Computer Vi- sion – ECCV 2024 , pages 216–233, Cham, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.954628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.639816Z digest=sha256:3b46e36c7fac1c315dc5e2d6e71d8846fa5ffaf2dacf6468f038db06b462ab66

Observation d1788933-3460-4a41-a68d-02d1fdcca032 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.941485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.644548Z digest=sha256:ab81c0014b075c48b7bcb4bf4a8b3716c31e367a305c217f9a9ce1518bc6ed8f

Observation 02298d95-bdd7-4554-879d-6161e2687468 · outbound

This paper cites Mathvista: Evaluating mathe- matical reasoning of foundation models in visual contexts.

Evaluating Vision-Language Models as Evaluators in Path Planning Mathvista: Evaluating mathe- matical reasoning of foundation models in visual contexts

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.928484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.649163Z digest=sha256:f6142194f62d280eda261e5bcd9dc72382c756efa7d2fcbdf66d2d37dada855c

Observation 9545c317-4529-403b-a694-402e8e14dd65 · outbound

This paper cites Ok-vqa: A visual question answering benchmark requiring external knowledge.

Evaluating Vision-Language Models as Evaluators in Path Planning Ok-vqa: A visual question answering benchmark requiring external knowledge

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.914836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.653818Z digest=sha256:08285f5eb5ba07b3ee71237c087ef08dd06b050780ac3f956c7d8bada242ae4c

Observation fc42eb96-b56d-4b35-92c0-c43a37e5218d · outbound

This paper cites LLM-a*: Large language model en- hanced incremental heuristic search on path planning.

Evaluating Vision-Language Models as Evaluators in Path Planning LLM-a*: Large language model en- hanced incremental heuristic search on path planning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.900899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.657811Z digest=sha256:e8913473284573dd2ed8f0114188b86dc8d1f8e59b0f060287da01467b23c23a

Observation f2778f7b-f303-43de-a0d0-90cfa4fef962 · outbound

This paper cites Jmmmu: A japanese massive multi- discipline multimodal understanding benchmark for culture- aware evaluation, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Jmmmu: A japanese massive multi- discipline multimodal understanding benchmark for culture- aware evaluation, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.877713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.665547Z digest=sha256:3b0fdf1aad95e4103a8b378aa53df0fa778d247ddcc4f7d7ddd05794197b0b2b

Observation df49d972-79f8-4305-a284-f22d0b8bb3c8 · outbound

This paper cites GPT-4V(ision) System Card.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4V(ision) System Card

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.864318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.669617Z digest=sha256:f8f7cef39cb3472c27ac6780366a5b1b8fc680811ad5a2c1e14618f700db9f8c

Observation 3a46277f-1688-468b-ae25-4576b0b7248c · outbound

This paper cites GPT-4o System Card, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4o System Card, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.852234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.674070Z digest=sha256:1418ee57a21aa55f9679672711f1ad1b7323ff105915820b6af9fba78ae06f0c

Observation 49af364b-1ec3-43b6-a1ab-1c1960625206 · outbound

This paper cites GPT-4 Technical Report, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4 Technical Report, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.839972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.677668Z digest=sha256:f03a662894acc1fbcd5100fefd465788c480ab949bbad552f6696e1312fd86e3

Observation 2be7e609-e614-41a6-8640-3b242c9a3142 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.826918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.681380Z digest=sha256:6cefa970804b9b6ce024030b2c34fd7c4f3c71978a35b98a5f62d658647655dd

Observation c0808b7a-de1c-4e46-a2b1-b2ba824cbcb5 · outbound

This paper cites LangNav: Lan- guage as a perceptual representation for navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning LangNav: Lan- guage as a perceptual representation for navigation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.813938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.685785Z digest=sha256:fb1032d625e65f3d7ebfaaf8e4cc2acea61d214e4b326fe27f86954db73534ef

Observation fd4834ce-28e1-4ef1-bf36-d618db0fb3bb · outbound

This paper cites VLP: Vision Language Planning for Autonomous Driving.

Evaluating Vision-Language Models as Evaluators in Path Planning VLP: Vision Language Planning for Autonomous Driving

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.800745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.690397Z digest=sha256:97b56b42a48316bafd7b79b08d8940d12d59f8bc049b8bcd540cae19c8da6025

Observation 99c8b953-1c46-4ad5-b9d8-5260e87b8c2e · outbound

This paper cites Path planning for autonomous underwater vehicles.

Evaluating Vision-Language Models as Evaluators in Path Planning Path planning for autonomous underwater vehicles

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.786723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.694719Z digest=sha256:6ea8e80e2d5686e44290c08ab60c628a1f2ce2332b1213992ab092b2cfead9d5

Observation c1538b26-955b-4445-88a4-12e2ce8e465c · outbound

This paper cites Clearance- driven motion planning for mobile robots with differential constraints.

Evaluating Vision-Language Models as Evaluators in Path Planning Clearance- driven motion planning for mobile robots with differential constraints

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.774346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.699192Z digest=sha256:d01ae82d30d6bb2c6fc547075fc6345c52e856fe7fd6b85b878c70ce1657b8a9

Observation 7005ad3b-978d-46f4-863b-f689888b76a5 · outbound

This paper cites Plonski, Pratap Tokekar, and V olkan Isler.Energy- Efficient Path Planning for Solar-Powered Mobile Robots , pages 717–731.

Evaluating Vision-Language Models as Evaluators in Path Planning Plonski, Pratap Tokekar, and V olkan Isler.Energy- Efficient Path Planning for Solar-Powered Mobile Robots , pages 717–731

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.761555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.703585Z digest=sha256:3413afcda50a0e6438a6ee0ad424c2eee5a7c33db85adbdda8f2dabf330c1ae0

Observation 97acd0c8-db5d-457b-835e-ca4a44c14580 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Evaluating Vision-Language Models as Evaluators in Path Planning Learning transferable visual models from natural language supervision

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.748543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.708095Z digest=sha256:347ab0f3f964900b99d1d8a83cdb2afc5d4a3df46533be38a7865a986d0b82ab

Observation 208c4b37-928b-43b2-9bb2-3b32290f0cb2 · outbound

This paper cites LAION-5b: An open large-scale dataset for train- ing next generation image-text models.

Evaluating Vision-Language Models as Evaluators in Path Planning LAION-5b: An open large-scale dataset for train- ing next generation image-text models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.734649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.712367Z digest=sha256:b6ad6e11a72687dcb841d5b3f0c565493a9d9036d2265cdfbade59a02294f500

Observation c8c31025-8e85-4d14-93d6-13da16e0cf44 · outbound

This paper cites Investigating the Limitation of CLIP Mod- els: The Worst-Performing Categories, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Investigating the Limitation of CLIP Mod- els: The Worst-Performing Categories, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.719336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.716271Z digest=sha256:e117a2782bab0b3e2300cc99f1d197ce1d610d22d2b7535a93898745861e0d8f

Observation 1e0c459e-dbdb-4c73-b732-7123298430c9 · outbound

This paper cites Tenen- baum, Leslie Pack Kaelbling, and Michael Katz.

Evaluating Vision-Language Models as Evaluators in Path Planning Tenen- baum, Leslie Pack Kaelbling, and Michael Katz

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.705388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.720569Z digest=sha256:84f4cd4d2941465423f45c1fe9715f0582e8ffd85744532dfb135baf8765f916

Observation 13b5b4a4-0daa-4350-ad8e-11a710779da2 · outbound

This paper cites Sucan, Mark Moll, and Lydia E.

Evaluating Vision-Language Models as Evaluators in Path Planning Sucan, Mark Moll, and Lydia E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.579131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.724836Z digest=sha256:11f6da5b221a176ed7d2c7f01075da2d0b563cab32578f825b784991b4a3c37a

Observation 51a45461-d8de-48fc-8e17-dbde915b2d2c · outbound

This paper cites Explore the hallucination on low-level perception for mllms, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Explore the hallucination on low-level perception for mllms, 2024

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.565304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.729106Z digest=sha256:32f3c44ff57d8a75ca5aaa77d34c24cc4e5cfe24b50d9f49dd88c0881ce4b3ff

Observation 0c054946-0e83-4f26-b2a6-214da8be18a7 · outbound

This paper cites Learning to nav- igate unseen environments: Back translation with environ- mental dropout.

Evaluating Vision-Language Models as Evaluators in Path Planning Learning to nav- igate unseen environments: Back translation with environ- mental dropout

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.550962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.734506Z digest=sha256:31639b5467a5663a3cd541148a14c59ea9a67185c177c06de59161dfa4916cda

Observation 338e3d43-c4cd-4d4d-814a-ce43c9e89ffe · outbound

This paper cites Gemini: A family of highly capable multi- modal models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Gemini: A family of highly capable multi- modal models, 2024

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.738697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.738697Z digest=sha256:56fe35d53e2fc20800cf2703e6af2aad894c31825a47090851853dbf09744959

Observation 8d596672-66fd-4258-8223-56030596b5dd · outbound

This paper cites Mass- producing failures of multimodal systems with language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Mass- producing failures of multimodal systems with language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.526822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.743101Z digest=sha256:5779f5c18e9a847b07fcd263b71a16c35554b8004ce239b4aba79f172df1a193

Observation 38fbf9ff-1a3a-4dee-87a6-5bdf75093e3f · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.513386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.747472Z digest=sha256:b96b6c02d8c4d59ecbc23ec47dc628e791dbf07d1e2ebb2d9742781b9d9dfe5a

Observation 2e9a1d4a-dc7f-44f4-8186-41d648aa540e · outbound

This paper cites Llama 2: Open foundation and fine- tuned chat models, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Llama 2: Open foundation and fine- tuned chat models, 2023

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.500389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.751680Z digest=sha256:aa5dfb4d2f59ae73e2d8f085e3b963ff90b234081861d3922a3dc12643bb4193

Observation e3c0274e-1378-4ba9-8692-4305b46a184d · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.487491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.756200Z digest=sha256:38268f0da4dfe533a31f0a4fa3c45fe8b0a5ce4bd7e6965cd1e658d4229a72a1

Observation 8218cfe2-8240-4893-beb4-ca100ef9707a · outbound

This paper cites Large Language Models Still Can’t Plan (A Benchmark for LLMs on Planning and Reasoning about Change).

Evaluating Vision-Language Models as Evaluators in Path Planning Large Language Models Still Can’t Plan (A Benchmark for LLMs on Planning and Reasoning about Change)

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.474491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.760423Z digest=sha256:c3bd1b97c5bb78d8c83d98ed5434ec6b01618b0cfff4a6b49f6cf4ad0dde80fd

Observation a5caeb81-5348-4536-9ab7-fec37bbde83b · outbound

This paper cites LLMs Still Can’t Plan; Can LRMs? A Preliminary Evaluation of OpenAI’s o1 on PlanBench, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning LLMs Still Can’t Plan; Can LRMs? A Preliminary Evaluation of OpenAI’s o1 on PlanBench, 2024

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.461224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.764554Z digest=sha256:7f558b68dc74bda60f10f0541c1cdabc4dd2e240bf0a261ccc65eb1b032475a8

Observation a70f9c6a-7463-475c-9399-92e807d2b61c · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Evaluating Vision-Language Models as Evaluators in Path Planning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.768564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.768564Z digest=sha256:75d616b63d1a9aaf01e5043e9690657dc135c8c97d670638da9966f7fcf09b58

Observation e6ef1e45-00f8-4e60-bb4d-7d12c52c19e6 · outbound

This paper cites Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.446679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.773169Z digest=sha256:135ba37cd09c016be4dbeb4dd9152fa4cf66832b3c6604ec93bdae469e675359

Observation a92fed30-c6c7-4975-9a0d-0e9b45721379 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

Evaluating Vision-Language Models as Evaluators in Path Planning Chi, Quoc V Le, and Denny Zhou

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.430518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.777491Z digest=sha256:2b2254d8f21e3b1b963dabb0f48c6f83f4dce327ba64e4f1204c60974066d151

Observation 61302c78-1f1e-406c-8c9b-caa0d484fed0 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation.

Evaluating Vision-Language Models as Evaluators in Path Planning Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.416581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.781477Z digest=sha256:84c993e5311da8bc4c9515cc91ed54187cbac488ea08348faa340516be078f18

Observation 6545e40f-d0ed-4e12-9072-8eb767b47fe9 · outbound

This paper cites Text2Reward: Reward Shaping with Language Models for Reinforcement Learning.

Evaluating Vision-Language Models as Evaluators in Path Planning Text2Reward: Reward Shaping with Language Models for Reinforcement Learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.402513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.785562Z digest=sha256:6c53dd9075368a414330c7f631a8eff010c8cf8fb907988bf66941f755a7572e

Observation 34eddb6d-510e-41dd-9b97-90abb6b76ff3 · outbound

This paper cites Evaluating spatial understanding of large language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Evaluating spatial understanding of large language models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.387997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.789884Z digest=sha256:66966de10f51b5f5b0e6cbcccb4667e1ba3d05bc0f068d7e4ee7988bb192f2a5

Observation bd391b2f-403e-4ef6-8dd1-074afec3a009 · outbound

This paper cites Guiding long-horizon task and motion planning with vision language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Guiding long-horizon task and motion planning with vision language models, 2024

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.373817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.793901Z digest=sha256:5a32a91b2b56822aaa853424a2d46425a61269bbd5df9143c484b1c5ded2f2e2

Observation 606d6a23-387b-4cc5-b586-367b13321856 · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models.

Evaluating Vision-Language Models as Evaluators in Path Planning Coca: Contrastive captioners are image-text foundation models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.797911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.797911Z digest=sha256:274bf83740e251b57e65e9e492bb45a6c2ea2f29f01631f45b0328c7eef72abd

Observation b60f5fd1-affd-40a6-9a94-6ddabd629ecd · outbound

This paper cites Mm-vet: Evaluating large multimodal models for integrated capabilities.

Evaluating Vision-Language Models as Evaluators in Path Planning Mm-vet: Evaluating large multimodal models for integrated capabilities

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.349494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.801895Z digest=sha256:8be3c55072b4de389687572e99e8fbf1f2f9732c1d7c13ae90263b4d4342c6b4

Observation 1faf5b72-fe6e-4f91-8347-7cf89c35e0b0 · outbound

This paper cites MMMU: A Massive Multi-discipline Multimodal Under- standing and Reasoning Benchmark for Expert AGI.

Evaluating Vision-Language Models as Evaluators in Path Planning MMMU: A Massive Multi-discipline Multimodal Under- standing and Reasoning Benchmark for Expert AGI

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.335634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.805741Z digest=sha256:9d9dc6e04c57cbe43692f34ed0f8e53bba290a03e9f662c9fbccc04629457235

Observation 913e0fea-1979-4c77-8c4f-94073a936305 · outbound

This paper cites MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark.

Evaluating Vision-Language Models as Evaluators in Path Planning MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.809641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.809641Z digest=sha256:6415803e75d3cc7e91c836b763fe62ca6d54a985a018a4cda16dfa00fe8c350b

Observation 42b58dfc-dacc-4079-8477-7d891a1bbed2 · outbound

This paper cites Sigmoid loss for language image pre-training.

Evaluating Vision-Language Models as Evaluators in Path Planning Sigmoid loss for language image pre-training

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.321121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.813434Z digest=sha256:fd8c38f45798117e719f82f4e530593d0b548f70b2a51e90fb0414561c5e6b71

Observation 92df6538-c39a-4af8-b197-da54b87e7d31 · outbound

This paper cites CMMMU: A Chinese Massive Multi-discipline Multi- modal Understanding Benchmark, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning CMMMU: A Chinese Massive Multi-discipline Multi- modal Understanding Benchmark, 2024

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.307576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.817313Z digest=sha256:b079913e2643ff7d2206747812e622419f4ebd33f1dcfd124a7e7de2aaf97aa5

Observation ee29249f-ac69-4f5b-9198-d65cce5efafb · outbound

This paper cites Multimodal chain-of-thought rea- soning in language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Multimodal chain-of-thought rea- soning in language models, 2024

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.295221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.821187Z digest=sha256:382f16391489b31e86c0f5d01682530ba6ed796a5b91d7d928a963ad206114c8

Observation b9ed1ede-e31e-49eb-95b6-cc626a4cc5cf · outbound

This paper cites Policy Improvement using Language Feed- back Models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Policy Improvement using Language Feed- back Models, 2024

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.282719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.824729Z digest=sha256:b1f4dcb606366a576f36f39fdda8a034ddb1b7d03e04e229ffed28ff4e0f745e

Observation 3898e54f-21ac-40e0-97ea-31d3d2feb1a5 · outbound

This paper cites Visual7W: Grounded Question Answering in Images.

Evaluating Vision-Language Models as Evaluators in Path Planning Visual7W: Grounded Question Answering in Images

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.269706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.828765Z digest=sha256:f1c53a0ad3be92ca39f1e4f8cde2a4314b35bf8b42d81dcb9d0f1a87667d96f6

Observation 30afc4de-d523-4a54-938e-65a6753c89cb · outbound

This paper cites Min Clearance = min pj ∈P min Oi∈O D(pj, Oi).

Evaluating Vision-Language Models as Evaluators in Path Planning Min Clearance = min pj ∈P min Oi∈O D(pj, Oi)

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.832912Z digest=sha256:4eda7d2ffbfd51acf838aac8903e21ba9a891864a2e89838c885c58fa591af24

Observation bba82f7d-5a91-44ef-a3e9-71d7261c0db1 · outbound

This paper cites Max Clearance = max pj ∈P min Oi∈O D(pj, Oi).

Evaluating Vision-Language Models as Evaluators in Path Planning Max Clearance = max pj ∈P min Oi∈O D(pj, Oi)

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.243179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.837254Z digest=sha256:6befb3c48237cd194d1b8ac6704e3748ad945c7c599a3ec72b5e9c884fb1641e

Observation 7ba76492-d207-4345-a47e-c5540837a059 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.229295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.841771Z digest=sha256:14eea1254490e95656d18a74d650cc9bae678c5a8c91b14e629f8b701a22b7da

Observation 71a11df0-45e8-480c-b18d-79f385857564 · outbound

This paper cites Path Length = nX j=2 D(pj−1, pj).

Evaluating Vision-Language Models as Evaluators in Path Planning Path Length = nX j=2 D(pj−1, pj)

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.215081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.845673Z digest=sha256:d01d89144ad6c127c8183ab70aed53c161999a080ef789da049df8e04b0cf81a

Observation 1482156a-5a18-4662-bd23-dfddcca7b0f6 · outbound

This paper cites Smoothness = nX j=3 θj where θj is the angle between the vectors − − − − − − →pj−2pj−1 and− − − − →pj−1pj.

Evaluating Vision-Language Models as Evaluators in Path Planning Smoothness = nX j=3 θj where θj is the angle between the vectors − − − − − − →pj−2pj−1 and− − − − →pj−1pj

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.201026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.849758Z digest=sha256:8c5fa82a1c7484133318cac91588c66d6c52695e111dcf88b98f92f903a8d82b

Observation 585fbad9-0e2f-4a7f-a1d7-0bafe47bd97c · outbound

This paper cites Sharp Turns = nX j=3 δj, where δj = ( 1 if θj > 90◦ 0 otherwise.

Evaluating Vision-Language Models as Evaluators in Path Planning Sharp Turns = nX j=3 δj, where δj = ( 1 if θj > 90◦ 0 otherwise

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.187140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.854193Z digest=sha256:2744546e465d1d24e679394df03bab718626aad8de4eef6719b4994e1a758e8d

Observation a07297e6-002d-4573-b2b1-06743bbd9168 · outbound

This paper cites Maximum angle = n max j=3 θj C.

Evaluating Vision-Language Models as Evaluators in Path Planning Maximum angle = n max j=3 θj C

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.173624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.858557Z digest=sha256:7c19c7d286e37b262ef2723a02605e80e2ea4cf49e8ee6bb79e767231c432464

Observation f1989607-864e-4dc0-b6d3-9f660d93f832 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.160650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.863056Z digest=sha256:5ee9a7c6ff3fdc1874455c8671370388efb1f854f6dff1697e4f277d93f9123b

Observation 8d3f074f-d807-4773-a14e-7c25a4875a3e · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.146507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.867141Z digest=sha256:1073251a64c36a36f96bf546e39b1238705f71007f8d35e45e1c4a29b2f67aae

Observation 1a7f99d8-acec-4c38-ae02-b07a8bf41e79 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.130946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.871269Z digest=sha256:bbce96251e709f4ff6d87a8c55fe939f8d62c3742a49dcf5f9c7290f81b376c9

Observation 9c553200-9c59-4854-a2ec-4cd40bffa397 · outbound

This paper cites Smoother paths have a lower smoothness value.

Evaluating Vision-Language Models as Evaluators in Path Planning Smoother paths have a lower smoothness value

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.117915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.875172Z digest=sha256:8147347dee9e7858561dc001dc9e7da929dd5671b79d58e4c212461841133f29

Observation 90985cf5-5f42-43b0-988f-9f1e9e4d1cb7 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.104257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.879270Z digest=sha256:0a808e50089ece990ee08c62b0f00ff2af4b6a7523059062b31064e3e6558ee2

Observation 252f6b04-5df5-4c9a-8607-d1f62ad1f11d · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.091181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.883284Z digest=sha256:d6cd064acc235b0ef42f8d1da4c9fc2ae25cf066108cecda7b81ab5eb8113611

Observation d23d5a32-8b22-4c8a-bb54-0cfab950f2be · outbound

This paper cites Path 1 has a smaller value for the given met- ric.

Evaluating Vision-Language Models as Evaluators in Path Planning Path 1 has a smaller value for the given met- ric

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.077702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.887611Z digest=sha256:7ff6c2322caf77d0dbe29e3114bb9c20fedf146ed75d5d2fe8d547736742e176

Observation a058d5c1-a1b5-4368-b40c-1362551359d5 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 2022

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T11:04:46.058456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:04:44.600983Z digest=sha256:ce4b48b9c4ca028ebc59ae3e4058d6a21d0197350056ac6856f2e34967e14eed

Observation da70871a-eec8-4526-9040-f859f5465cf5 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.661729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.661729Z digest=sha256:4c294eb0a38fc52dd9567105133871c035d424889cb49d70e937e8d179f59bac

Pith citing papers

Observation 1a90adff-d994-4835-914a-c1ef00a2389d · inbound

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction cites this paper.

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:49:11.200446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:49:07.509682Z digest=sha256:d37c21f854b5adaee98b0a1321e2893f0f9359d44926073f8d49d49091151abe

Observation 78728900-4e2e-484e-b0cf-c68cb9e86c6b · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:49.914971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:49.914971Z digest=sha256:ee7b104650aef241122c3c7245d6f3e8867fad4711167e0b34a126baa55da69b

Observation 26b5b30f-c90a-4d2c-aaad-dd4e026f029c · inbound

Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies cites this paper.

Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T07:49:44.998466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:49:44.998466Z digest=sha256:57fb9ce2d97fc3b1d6c4928cdddd6afa2627a1fb79d35217b02d9486cd0c31b6