Pith. sign in

Paper Citation Record · LEDGER

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models

As of 7 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2608.02197.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.02197 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T11:57:40.870137Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ab1be2c-cd79-483a-91d3-f9bb8504dd95 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.689174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.689174Z digest=sha256:884249f6c6ad67d84d9fa74ad82c6c96ed18e38a45c21189f33cddfa0896a685

Observation 4fe066ca-2b8f-4ae4-b95a-3f65aad88066 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.694039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.694039Z digest=sha256:bbd6ccbd066d8eef6cfa04bde324a1b5601f53d79990df96d3711de32b5b0c08

Observation 5805bee0-890a-462f-9798-ac30c8781e44 · outbound

This paper cites AAAI Conference on Artificial Intelligence , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models AAAI Conference on Artificial Intelligence , year=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.697519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.697519Z digest=sha256:86c9edc62ab053cca80edd0402d6b486b9ede9ca0913ef88122f5012f5d7e0ae

Observation 30de7cf8-ed26-4e8c-ac34-27595c947fb1 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.700483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.700483Z digest=sha256:1a39d18b318e175d5e8ef2ebc9cd6d820c749bd7fc1681317392160f02786484

Observation 2507c561-7a95-473f-a2fe-8ff0d29cc462 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.703298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.703298Z digest=sha256:55ed6bbcfe90a5a8d453128314ad898ef15de5b7b00bb654bd5a8cf8dcbe7ecd

Observation 88c2854d-376e-400a-ae16-95fd975b3dbd · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.706639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.706639Z digest=sha256:ee48cb83910496dae9246bf742c9d9cfa56f4e4b8f7e3f2cea7e80b82bc27646

Observation 2717589e-0931-4fb7-b91c-646170a32397 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.709550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.709550Z digest=sha256:6ecaa81b0f775c5aefe7ea23638134326be8b93d835465a45e23ad7c19de4e47

Observation 4ff15f73-5f29-4fe9-aca2-22e132ff9a35 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.712625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.712625Z digest=sha256:f6794cac94ff3a4002d6bee77bc16dd2f04ab731ce2c41c4a1078e86ee94d702

Observation 866ea810-160a-4a2e-8f33-ce8dcdb215b2 · outbound

This paper cites Proceedings of The 9th Conference on Robot Learning , pages =.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Proceedings of The 9th Conference on Robot Learning , pages =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.715883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.715883Z digest=sha256:8e81da226c0cb7a117ed9e82dc6ea1b60f059dcdbbf00dba9d7c1fdd3331fe37

Observation 255a1815-1fbf-4a6f-83e5-fc9295b8d346 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.718769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.718769Z digest=sha256:836f524bbdb3bdc75062889c2328c4a91030a10e1cc3d567ceba372d4cb057ca

Observation 581d0b33-eae5-428a-ad83-9ab2c3004cb8 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.721475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.721475Z digest=sha256:3803cf596141f3ac604fa21fb6e75946475d242d9044b8aca276b79c1b74a4c8

Observation 374f0fa4-f483-4b19-8433-8cab600aa0bd · outbound

This paper cites Conference on Empirical Methods in Natural Language Processing , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Conference on Empirical Methods in Natural Language Processing , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.724285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.724285Z digest=sha256:21b9ddcf21cbef61c80521bb3d7e1ad1ed7babc308ad02187d4e11e92a0c62a3

Observation 27903a81-128f-49a2-bf0b-a8025d141597 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.726856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.726856Z digest=sha256:807111c8bfa8465de91c6f17c7dff03dadc0ba48cd2d28127282c1d2e2f23c87

Observation b078b554-02f4-44d4-94b7-4c491869dae2 · outbound

This paper cites 2024 , eprint=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2024 , eprint=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.729288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.729288Z digest=sha256:05165fc813af77effea14f8ababbabbe6a2e50c38f6ffd2b1a1a195fddaa1cb1

Observation 084ba25a-67c4-4e45-9ce0-962cc7dbf93c · outbound

This paper cites 2025 , eprint=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2025 , eprint=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.731816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.731816Z digest=sha256:6ff3f3606463062fd649c8b87805f5dd2e4ac71956daf2c006c8485fd4d261c2

Observation b0e48609-837a-4fcd-a0fa-a14a2f12fb47 · outbound

This paper cites 2026 , eprint=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2026 , eprint=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.734882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.734882Z digest=sha256:8fd978dba83a9f8ad49912d12159fda2210f90084215fdd8c38b4893dd19ba0b

Observation 0b735300-60f4-4a61-ab68-fafa9a99904b · outbound

This paper cites 2024 , url=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2024 , url=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.738103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.738103Z digest=sha256:9e17e2f20880ebfcc66a26c1e70f046383e2f6589cca317cbba77f760facb411

Observation 92ef1562-f681-40ee-bac2-6210b4908096 · outbound

This paper cites Conference on Robot Learning , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Conference on Robot Learning , year=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.741422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.741422Z digest=sha256:236d29c19d583e4ddf70ab420106f8e3ab01ae66e0110f31d288a46977a6f9db

Observation e6678af5-d7b7-42d2-8623-42c577c440e7 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.744239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.744239Z digest=sha256:e2bdd888a6c4d15a8dd965c23fe597b31a8edd04259f3d2fd4f7aace17200f7a

Observation 37aa1a86-1b50-47d7-882e-25649b8ad648 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.747693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.747693Z digest=sha256:d4a5f2ebc0de024db2e6ea6aad0730e7e4e7c62e9abbd83b30266339dc5cb89c

Observation b68d45c1-03e4-4570-b42b-cbde06f42852 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.751079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.751079Z digest=sha256:fe5c956f14cc8017fe3b1d822f8f47629d2c267fc6d1f37d8da072252c381672

Observation 1f050c24-a7ef-41d7-8c7f-79ddcfaab3e7 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.754155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.754155Z digest=sha256:dfb1cc7f8337687751ef284bab3d0a42177329a27a56b193e6a8125d626716aa

Observation bdc3bcbb-391d-43da-82d5-ab93209d195e · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.756778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.756778Z digest=sha256:0155bd3ca12fadda453e8713ea9e9bf7cb7753d07a4cdcf68869bd9467615ef0

Observation 2e57f597-462f-4df3-8ece-19ae4e1cc5c4 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.759385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.759385Z digest=sha256:6cb07efba3b17d57cbc30f390bbaba6258bb141c8b9d45ef5b0736ea6c65dc00

Observation 768c7528-7671-4a50-9adc-94e87fd033fc · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.762276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.762276Z digest=sha256:8e5dec5f3e5507c9b6464d8f1dd3d88d42bfeae0d9bd3e8bc309136f388cac6d

Observation f9a6f5d7-a80b-4ae9-8af6-28a95f7343b2 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.765237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.765237Z digest=sha256:e588c36a81bf9f573a299b43282a8352fe793345f4244f6cb7c1d5693f13d709

Observation 02721ad6-9c6d-483b-ad73-c6104078a2e1 · outbound

This paper cites 2023 , eprint=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models 2023 , eprint=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.767922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.767922Z digest=sha256:e69cae9401e5ee5d37820db496d36a329db08c246aa62c2b25e937196277f04e

Observation c60642a5-185b-494a-8359-5ebd894b26e2 · outbound

This paper cites Nature Machine Intelligence , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Nature Machine Intelligence , year=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.770348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.770348Z digest=sha256:bb7834da206f23628a7b04cf7d959bbb2f8bbb10aa9e662d86ba3891fb049eae

Observation f66a8724-c043-4a7f-bb80-8149b8fca07d · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.773148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.773148Z digest=sha256:4d5fc9d5636f816dc301f2fa1be382fb1131c7d127463fda239f9d26f68963c0

Observation ade07406-9ac6-4c63-89a1-d0c7a5fa2612 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.775685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.775685Z digest=sha256:056f06eb6278886c0da242e1c028ff962813d7c48bc816ae2e6f327b8445878a

Observation 0137c711-66bf-4c63-9b7b-c94fbc4b3353 · outbound

This paper cites ArXiv , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models ArXiv , year=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.778473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.778473Z digest=sha256:b83744f25fcb7a10b93da53334cd43f95adc15f0930b2bc4bd714202b11fb78a

Observation 3e72db6d-bf3e-41c4-93b4-423dec735f44 · outbound

This paper cites Conference on Robot Learning , year=.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Conference on Robot Learning , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.781005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.781005Z digest=sha256:ad23d7d22e64c9dff9db934346ae0d27fdf7b5af0807064953c8827238badc0f

Observation 35543a8b-f4f0-4511-8f86-f75b66d7d1e7 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.784142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.784142Z digest=sha256:a22f0d6688f10fd0dfbc663572157f208726f20fb47e20aeb7b2007579fdcf5e

Observation 3af9e1b6-5df7-4817-aa25-eaa11ba2d3b0 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.786692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.786692Z digest=sha256:edede055095050f7d08ed2ff18f69c4e2349ff9c061be6cc79cca69d8c438298

Observation 9b565bce-0e98-4b76-b7d9-d789a85a2ff7 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.789668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.789668Z digest=sha256:8ad9472e1b1b852bba1d591381cfab4ce698db625e1a9acd15c89b49d9435ffe

Observation aab439bf-3d7f-4eaf-8bcd-c1c15c9fb142 · outbound

This paper cites CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.792418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.792418Z digest=sha256:e1fdac3716ea787432b4893789b6e0b9931c765cacdcb8f4a16aea88b8d0b14a

Observation 042860c8-07ce-4426-953b-2fdaf8438876 · outbound

This paper cites Vision Transformers Need Registers.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Vision Transformers Need Registers

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.795888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.795888Z digest=sha256:39f18d5b1270e11546a93963850d27e0e26590ab941ee80bb5bb913c0ec2e083

Observation d68759b8-9061-47a7-9452-aa4e2fafab7e · outbound

This paper cites GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.798816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.798816Z digest=sha256:5853ea77ebe6959dfb282cd35573d50aef21e796dd5874e992fe8c8b8185fdbf

Observation 7e8b6d5d-f9c6-4618-b353-6e6bd8cadc5c · outbound

This paper cites SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.802406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.802406Z digest=sha256:2df0681eec34aed85cda07c03dae795905baf6aae82dd24c1ba82e6c8d76616e

Observation 7fde7fcd-0580-4755-abfd-b35902c57c5c · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.805359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.805359Z digest=sha256:0e9ebc7e298dbd215f5dc9f330ad7ab448a1117af19f0b98500ace0522ccca43

Observation 3e4d4107-1c0f-4075-9c47-3ff034baeaa3 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.808907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.808907Z digest=sha256:13bbf01936d111b4ae1f7d9bf7e57adf71126c6e33970d63f38d3f45c733d00e

Observation 35eeda9f-01e8-40be-b2f7-bc412d9ef95a · outbound

This paper cites K.; and Panov, A.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models K.; and Panov, A

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.811540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.811540Z digest=sha256:80082b05b279ffb6a491650a600bb60f24e43f58f2402c0605d1e3661cb3ea50

Observation 1f067dcc-ead8-4fcc-9df7-e7e40125fd61 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.813969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.813969Z digest=sha256:cbafff5fcbbc778aaeab32ace2974b785d5a495bbf67736ff64b42a69e150dbd

Observation 72df7c2d-feb4-4843-a9de-af139903f0ba · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.816402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.816402Z digest=sha256:ac24ea3277b6961bc7ffa6c2f8d568d6dd060ca49af2b07cd8fa8bb25f59536e

Observation c90fcc36-a54f-4fe0-83fc-782de7589e73 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.819560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.819560Z digest=sha256:2b856e9199a8530dce6d9852fac9b6eb0db075ef1a95af782dbac51100599295

Observation 866f949d-1f4b-448d-9bf3-77f4d8a5c3ea · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.822184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.822184Z digest=sha256:5d063121d9061797bf5ac9552cdd9258ef645b886e1b5dab507bd7ac00eee2ce

Observation f07a0489-153d-4201-9b12-c8a3337bb571 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.825212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.825212Z digest=sha256:e8c1b769fa9ab10fce67d8055d6478f88aba9438a261f85a374f18e453de3539

Observation 2c3aed0a-44ba-44b7-a45e-eebad6d181e5 · outbound

This paper cites R.; Fu, C.; Lunawat, I.; Sieh, I.; Kirmani, S.; Levine, S.; Wu, J.; Finn, C.; Su, H.; Vuong, Q.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models R.; Fu, C.; Lunawat, I.; Sieh, I.; Kirmani, S.; Levine, S.; Wu, J.; Finn, C.; Su, H.; Vuong, Q

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.828202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.828202Z digest=sha256:fd814771bebc00a9d48ad3929d933c3ed435fc2d9bc719dd8191f5a28a79c5ed

Observation 512359e9-af7f-4560-87a0-aa7bb75e3abf · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.830995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.830995Z digest=sha256:0dd092e611b1058566416e7c671cda0ab8344d5af7609690dfd96faffb24f490

Observation 47eb7604-1d83-4922-b15c-94dbf7b68e9d · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.833537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.833537Z digest=sha256:32b8ec750831a56609807c3be3741067b04925ebd5d1cdd529e332aaf408df1a

Observation e1a76801-a6a3-45a8-8d7f-fc2b135b2b13 · outbound

This paper cites LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.836041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.836041Z digest=sha256:bd2aabe797e26c223aca90e2a5e23019c5585f7e75499c0314701aa9c427f4fa

Observation 43170df6-d614-4c33-94f2-bfce8ce52198 · outbound

This paper cites Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.838682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.838682Z digest=sha256:bda175bbff16a535005ca5aa071b54ce72be801555c75c2858625f064e3701de

Observation b3be2380-7ad6-4c35-a6ef-c977b09ebb8f · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.841228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.841228Z digest=sha256:1d7ad36373638d418f63745f27325841648e8988b4e7189655f2b75c9475abbb

Observation 59473aee-1b0d-41e1-9ffd-fb11c1a168d4 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.844058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.844058Z digest=sha256:2a25a244a7a2e660503c28ac556df70fd388908cdd1635c6650bdafb2bc06448

Observation 2a658a44-9ef1-4a9b-a27d-6ff5db8fe674 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.847244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.847244Z digest=sha256:22699a0d93d2c15999a78063ce69f3913e84344e25002bfe76b7a7314164262b

Observation 7b54e302-5223-4fbf-b14d-d5e40bd80eb2 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.849735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.849735Z digest=sha256:3ee0cd29ec481ab74a125ce63536408fa282b448f00c60f42ef62dfdc6645054

Observation f0f812ab-e057-4896-a429-b6a98d1b7a4c · outbound

This paper cites VLANeXt: Recipes for Building Strong VLA Models.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models VLANeXt: Recipes for Building Strong VLA Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.853069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.853069Z digest=sha256:4c403aa66080a59ead17ce0da22eb6129c72997604c02350341695fc6c15ef59

Observation 62fd6580-26e2-4760-ad5a-8f06427f3bf9 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.856209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.856209Z digest=sha256:d45db7a6b874ad8c626778180f2e4e7eb0f5fc12477bc6eced3ef04490f213e1

Observation deb79fa8-4ee0-4c21-b1bf-5230b348c4dd · outbound

This paper cites Towards Perceiving Small Visual Details in Zero-shot Visual Question Answering with Multimodal LLMs.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Towards Perceiving Small Visual Details in Zero-shot Visual Question Answering with Multimodal LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.858753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.858753Z digest=sha256:e7c44005afb285364fb35aa5724ce7f556883bfa88c8d513431f6c2d20465164

Observation 6658d97c-8203-4d30-ae54-8f84c085a47f · outbound

This paper cites MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.861348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.861348Z digest=sha256:5a8c520f5190b0e218c17b5c5bc10a08796cd6514afbfbae66f264ff9e285c58

Observation 344253a1-31c9-4b33-baf3-fa0c6e6fd0d0 · outbound

This paper cites J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; Handa, A.; Liu, M.-Y.; Xiang, D.; Wetzstein, G.; and Lin, T.-Y.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; Handa, A.; Liu, M.-Y.; Xiang, D.; Wetzstein, G.; and Lin, T.-Y

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.864011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.864011Z digest=sha256:31282f5752f0cfb63d1cc4b628b9ed4c31788b8d7ba37f2210fe07be38d77676

Observation b8e99a9f-0024-44ea-a67a-91feec2e19a4 · outbound

This paper cites TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.867391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.867391Z digest=sha256:3dcb79d9023b079a55d06f184a7195ac4bc8359e3d774d4cd5bb7d7590f427b4

Observation 71a6628d-66ee-4752-8587-d1e5f167dcd2 · outbound

This paper cites an unresolved cited work.

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T11:57:40.870137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:57:40.870137Z digest=sha256:08653c659daf6a873cee673e70cb896452e2bd6e56316dc99491ea7372980dc5

Pith citing papers

No inbound Pith citation observations are available.