Pith. sign in

Paper Citation Record · LEDGER

Visual Grounding in Zero-Shot Vision-Language Control

As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.06154.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06154 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:12.082156Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a603f8c3-022b-4147-8169-63fe57346953 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Visual Grounding in Zero-Shot Vision-Language Control Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.790543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:08.270889Z digest=sha256:a133e77808e1b849b7cc107138ab6b2ba3ce281b931c0a7fedfec9386336888c

Observation 0e43c405-639a-4c5b-b059-ff170912dad6 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Visual Grounding in Zero-Shot Vision-Language Control Improved Baselines with Visual Instruction Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.301524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.301524Z digest=sha256:b228b3cd5a08b883c80f2966b8145598c5097c216f2c1b3a1163ef4cdcd697a9

Observation 4603868a-d4e8-4c73-89f7-1d53b5515177 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.373123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.373123Z digest=sha256:6669d2c41321f4832c7e2a2626664d23ea802cbfefda34c40d60c00510fe6616

Observation a1a210d9-32bb-4b99-8083-5b762cb6b2ef · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.442796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.442796Z digest=sha256:619944f9db59921f90e4e9a9c5b6e31be1ae34836336b3b5680ecec8053992a2

Observation dbb8b77a-3727-4987-a03a-e8a06282aa5f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.553499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.553499Z digest=sha256:37692cdf166b7964af5eeff1f9086a9f99ef4329c4114079a89ed1464b90ea60

Observation fb1a59d1-ed3d-4bad-9ccb-6af984254bd5 · outbound

This paper cites Qwen3-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.643309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.643309Z digest=sha256:57ff3992bcb4281b6945577897c653f8eca923842a2b57835db3c2ad230f6024

Observation 4d55ec77-a452-4a5a-becb-390200596896 · outbound

This paper cites Gemma 4 Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Gemma 4 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.729258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.729258Z digest=sha256:906b51d93cfdc586511470db7210b071611aca53654760381e539c556861502d

Observation ec3159d7-5976-4240-90d2-7389098270cd · outbound

This paper cites Qwen3.5: Towards native multimodal agents,.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3.5: Towards native multimodal agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.579700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:08.814585Z digest=sha256:0088fbe4623352e62a6f4ccf6bac2a25abc93f6389a629581a70fb2a7d7553c7

Observation f585939b-9dd9-4c7b-88d2-52e812ad10b1 · outbound

This paper cites Introducing Mistral 3,.

Visual Grounding in Zero-Shot Vision-Language Control Introducing Mistral 3,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.388390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:08.880443Z digest=sha256:6d637e7fe059ab40d022f633e18baa2dc5e45207b7a5cdbeca7f140cd9ddc129

Observation 06420827-4300-4bed-81da-c6f12310a09a · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.961126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.961126Z digest=sha256:f2a6852d013fc6a88c263ee8a4e7f88ab85b4cc08776c26acbdc6092fa19f989

Observation 9ffb7a4b-5bc3-433f-b84e-657efa32534c · outbound

This paper cites SmolVLM: Redefining small and efficient multimodal models.

Visual Grounding in Zero-Shot Vision-Language Control SmolVLM: Redefining small and efficient multimodal models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.055707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.055707Z digest=sha256:7c8560f5603420853dba054a0af61e12d7056420208ddf1a7897afb86a61f3c5

Observation 2e96abb8-7e75-49cf-bad1-44b987d96f15 · outbound

This paper cites NavGPT: Explicit reasoning in vision- and-language navigation with large language models,.

Visual Grounding in Zero-Shot Vision-Language Control NavGPT: Explicit reasoning in vision- and-language navigation with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.185974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.150370Z digest=sha256:956ef8c7e1be3c370ad9cba98c2a7ca4b1390c9060f6d843014ca4f15e0ed886

Observation 2ac4a8d0-ddbe-4845-83ab-753e13619f30 · outbound

This paper cites VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.988861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.191891Z digest=sha256:47523d66ddacd8d5048489163b22225663f1a095256d3b4d71a577b6efd0b09d

Observation 1617f2b0-e8b3-4d04-a32d-24e9620e142a · outbound

This paper cites VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.794638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.271864Z digest=sha256:413b265e7403010674bd53273cacbf9d6c0e56cf283952763dc3eacfdddf4eef

Observation f31e68ee-2539-4c67-bb19-0ef0a3340012 · outbound

This paper cites MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,.

Visual Grounding in Zero-Shot Vision-Language Control MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.624658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.376035Z digest=sha256:6714e3a78484b372ee5fd6400d9172b2172495108b7dd02766001b0320785f3b

Observation 70fc3451-2ab0-43bf-bfdb-0fc46a2249cf · outbound

This paper cites HazardVLM: A video language model for real-time hazard description in automated driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control HazardVLM: A video language model for real-time hazard description in automated driving systems,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.405609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.469350Z digest=sha256:b5f774680053db598e3d83fbbdcafd1064907891354545540c752d858a91fc7c

Observation 25951574-ea08-42cf-87f1-209bb847b839 · outbound

This paper cites LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,.

Visual Grounding in Zero-Shot Vision-Language Control LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.178944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.543585Z digest=sha256:34d43d990f37950eb82c6fb0607e02fd9309f445fb3ef82d84d9af2f6ab80ca0

Observation c5278006-8a5e-4c69-99f3-2bcd668bf024 · outbound

This paper cites Semantic scene understand- ing with large language models on unmanned aerial vehicles,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic scene understand- ing with large language models on unmanned aerial vehicles,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.962984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.642706Z digest=sha256:bb60f2051f01b67f3b71febfbd0b390ae622f4267f334f9d56a30bbc370c3f1a

Observation 91026761-6ef7-4e28-b7b5-ce3798e118e0 · outbound

This paper cites Shortcut learning in deep neural networks,.

Visual Grounding in Zero-Shot Vision-Language Control Shortcut learning in deep neural networks,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.715223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.715223Z digest=sha256:e9180adf5aec4d9e04601b23e12c44d28923ae1f67505c2854de8c52a61d3f39

Observation f29b433f-6864-482b-8456-418ec2180865 · outbound

This paper cites Beyond accuracy: Behavioral testing of NLP models with CheckList,.

Visual Grounding in Zero-Shot Vision-Language Control Beyond accuracy: Behavioral testing of NLP models with CheckList,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.751966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:09.835083Z digest=sha256:722c37eec636e6978ab55bf4413445b8622d378e85fa6f00f1aee36db387bfac

Observation ca7c9e58-04a9-4a16-9f56-56cadde12161 · outbound

This paper cites Holistic Evaluation of Language Models.

Visual Grounding in Zero-Shot Vision-Language Control Holistic Evaluation of Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.972585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.972585Z digest=sha256:1c911df5b33ccefd5dd97565e6a1d8e119ce652611ae635b930ad128640ab5d7

Observation e00c28b0-ce6a-4c24-a854-1d9bf318dca4 · outbound

This paper cites Metamorphic testing: A review of challenges and opportunities,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing: A review of challenges and opportunities,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:10.087742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:10.087742Z digest=sha256:7fe919f90438f7e388631664e63a3945ea3c4127f94d74e1d059e3f9737377c2

Observation c77f0cb6-c974-420b-b245-66d362744f41 · outbound

This paper cites DeepTest: Automated testing of DNN-driven autonomous cars,.

Visual Grounding in Zero-Shot Vision-Language Control DeepTest: Automated testing of DNN-driven autonomous cars,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.544924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.174166Z digest=sha256:40e49117ec5a2babb0663d7176b3b0b5100fd7a1894be0d6090cc95cab7e37fb

Observation 7e56556c-58b4-4bdd-8297-6dc57e25b612 · outbound

This paper cites DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.289522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.263345Z digest=sha256:67feb5fe49ce23813138e42a6d70ecd3e805596c3d2a223b3c0c989ca036fdbd

Observation 3d13e243-58ed-4600-a09d-910ea348d10e · outbound

This paper cites Metamorphic testing for semantic invariance in large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing for semantic invariance in large language models,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.128045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.360725Z digest=sha256:3be6d747d75a9a7e656e9ca5e1a1e10b73a71a3bae11724ef61f2d1e3e44dc8e

Observation b0ec6d09-d4c1-48cf-8a38-2302de9e0bad · outbound

This paper cites Semantic invariance in agentic AI,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic invariance in agentic AI,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.966883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.489405Z digest=sha256:4366583ffff769270c8c02e155767577cf782682abc27266ace0812f982f73a7

Observation 30cb1895-c93d-4bb4-a201-3ef5de9543b3 · outbound

This paper cites Energy-aware multilingual evaluation of large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Energy-aware multilingual evaluation of large language models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.818647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.601401Z digest=sha256:c7b28106c0bf5557d04cbbde2ccaf26b0f8cb7a63f41a3def22852944e3f4ec8

Observation d6c3c2ca-c79c-444a-b2a7-7605c6b4ab47 · outbound

This paper cites EdgeShard: Efficient LLM inference via collaborative edge computing,.

Visual Grounding in Zero-Shot Vision-Language Control EdgeShard: Efficient LLM inference via collaborative edge computing,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.611251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.719094Z digest=sha256:12846b936909e11fb42ea8b8678e3f39451a9731cc2aaeb3b620d58d4431dfbb

Observation 225b71da-b3d6-4df4-9682-930d474951cf · outbound

This paper cites Power hungry processing: Watts driving the cost of AI deployment?.

Visual Grounding in Zero-Shot Vision-Language Control Power hungry processing: Watts driving the cost of AI deployment?

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.454238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.811054Z digest=sha256:4a1e2a431338ed187051237b3ce36b03927557d0f930cfff40fe3eed2bf0a550

Observation 87e4c2cd-8ee4-431e-a0bb-8a28e650ca2b · outbound

This paper cites An environment for autonomous driving decision- making,.

Visual Grounding in Zero-Shot Vision-Language Control An environment for autonomous driving decision- making,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.313176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:10.906404Z digest=sha256:78d41f3d73aaa7910a32425623081391509a273e6ff78c5bd0eb2bd5fa985f88

Observation ff57a3ee-4a1a-4b0c-a83c-a2dc768af9cd · outbound

This paper cites Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,.

Visual Grounding in Zero-Shot Vision-Language Control Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.146672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:11.001892Z digest=sha256:83c88ab0177dfc62551e744ef56513d89068deb4daa921649f4d952b136d7031

Observation 74bdaa22-0965-492e-9f57-c6556a94b078 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Visual Grounding in Zero-Shot Vision-Language Control Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.105161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.105161Z digest=sha256:09942a8da06c03bf76140fda6ad6f4a52059c7805d7db3c82aa3f9c113ab4945

Observation 7b990891-95c6-4224-a76f-fa9af2c2d912 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models,.

Visual Grounding in Zero-Shot Vision-Language Control Chain-of-thought prompting elicits reasoning in large lan- guage models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.192482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.192482Z digest=sha256:f70e880095e147460bd0dda1478e8271036cc7ed3840fcb14b932e4473cca725

Observation e81779ee-264f-462c-bdf0-47724d25e710 · outbound

This paper cites Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,.

Visual Grounding in Zero-Shot Vision-Language Control Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.001615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:11.303725Z digest=sha256:f142b8740bb67e2bf8da3f03e99dd621109b37df7c8ba9646eb0e0d078bf1964

Observation 4ec7dd12-c123-451e-b0c2-20d72d6cdf58 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Visual Grounding in Zero-Shot Vision-Language Control Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.374361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.374361Z digest=sha256:2d119025f4ca1975b12cdf14996486f28e0064e1bb38ac536aa79da063af38d9

Observation db19dba4-69c2-4470-8fa9-744f8721c355 · outbound

This paper cites Do Prompt-Based Models Really Understand the Meaning of their Prompts?.

Visual Grounding in Zero-Shot Vision-Language Control Do Prompt-Based Models Really Understand the Meaning of their Prompts?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.478970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.478970Z digest=sha256:caad86f8318a187e4a7fdc1854f43cac0d7555a3019889d9545e55ec87446a17

Observation d7353651-c3b0-43ac-b1a7-9db0d51c80dc · outbound

This paper cites PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts.

Visual Grounding in Zero-Shot Vision-Language Control PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.549507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.549507Z digest=sha256:cf73c82e8e0dc31933d4b7e883d77b6819021d3520da19c6aa5d830812e51eb9

Observation 8ebd7ffe-3c10-49de-9ef6-83dd799032d7 · outbound

This paper cites Measuring and improving consistency in pretrained language models,.

Visual Grounding in Zero-Shot Vision-Language Control Measuring and improving consistency in pretrained language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.854441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:11.628585Z digest=sha256:4f4950600db3aaba2700d9369bb93836e34b26738da01ddbb6a98d05a984149e

Observation 0cfbf370-8663-4870-bbf6-caf7a772d722 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Visual Grounding in Zero-Shot Vision-Language Control Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.698424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.698424Z digest=sha256:b44f9fffbaed2133e7171c4005ae01356681e0f501ba0f73744e73188949341c

Observation 57408b9b-37d8-41a9-ba60-92e3229aad1c · outbound

This paper cites Probing classifiers: Promises, shortcomings, and ad- vances,.

Visual Grounding in Zero-Shot Vision-Language Control Probing classifiers: Promises, shortcomings, and ad- vances,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.691304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:11.782731Z digest=sha256:7d146ebf852a9f571892324ea138b56e7f0e71ef7c0caf0df912e35d8e7bb5b6

Observation 3e1d8560-31a1-4c9a-b891-4f3e30f6441a · outbound

This paper cites Con- strained model predictive control: Stability and optimality,.

Visual Grounding in Zero-Shot Vision-Language Control Con- strained model predictive control: Stability and optimality,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.866001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.866001Z digest=sha256:48feac42aac506a9c24b2a0236bc56b5187a85ef26612cf83660f5e1e0d18aa8

Observation bc9e8cc5-c59c-41d8-881c-0040b5fd8e28 · outbound

This paper cites A simple sequentially rejective multiple test procedure,.

Visual Grounding in Zero-Shot Vision-Language Control A simple sequentially rejective multiple test procedure,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.927436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.927436Z digest=sha256:ea47279352d0f2ff6e4ab2ae2a32d7b8432aa1f59d58d5c6f7d69105a4c768e5

Observation 0a71f3cc-ccf6-462e-aee2-b81059cc820c · outbound

This paper cites SciPy 1.0: Fundamental algorithms for scientific computing in Python,.

Visual Grounding in Zero-Shot Vision-Language Control SciPy 1.0: Fundamental algorithms for scientific computing in Python,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.541497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:12.001667Z digest=sha256:063fe827ea84edd47d0c07bff7b6f435af93c0f7400dc2f76e0d76978a93784c

Observation f6350d25-8935-4296-a0fa-a42e0757c57c · outbound

This paper cites Transformers: State-of-the-art natural language pro- cessing,.

Visual Grounding in Zero-Shot Vision-Language Control Transformers: State-of-the-art natural language pro- cessing,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.424487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:57:12.082156Z digest=sha256:4a153576240bc03e262e1984b3688be0b8ce5cf0466b2e73a58aeec0268c2fd6

Pith citing papers

No inbound Pith citation observations are available.