Pith. sign in

Paper Citation Record · LEDGER

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

As of 15 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 8 inbound Pith citation observations for arXiv:2508.07650.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07650 v2

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:59:04.663968Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T17:38:20.785896Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 0a654059-9e67-4146-9782-73063cb25d2a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.281025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.281025Z digest=sha256:5ac738ac6d630a601498a4cadaeb80e4c96a020c292077dcb62fbbf35dd9dd25

Observation 633a45ab-0d9e-4ef9-b84a-796c397622bc · outbound

This paper cites write newline.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.374741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.374741Z digest=sha256:d05808cf4e0bc9e1ad18d02c285c96c0821dcb80634b9b0ef6f136cfda7e5570

Observation d5704008-07ae-4823-999f-98e7addf7ddd · outbound

This paper cites Qwen2.5-VL Technical Report.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.513385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.513385Z digest=sha256:6b6ad28c09c5d44bd22942f4986dd8bcf0e1ec2ddfe3a7b5b9f7730896ff3d23

Observation 4b160908-3609-46e3-bba5-8edf69cc572f · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.642456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.642456Z digest=sha256:ab5e0b5e2b23a6890269f0ca0a4e41368c2b57f875454d6d27d927554bc90535

Observation 1559dab7-3594-454f-bd9f-302cac670c54 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.765491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.765491Z digest=sha256:0f841b011c07a77deb7c1320b8fb4e854299fa5e77823d553f7b92ebf10af876

Observation 2ab685a4-4b7f-4a09-b1be-3b9db0277c27 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RT-1: Robotics Transformer for Real-World Control at Scale

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.915445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.915445Z digest=sha256:347ef8e2e04e30a307d52263dc0ecd3ead9fca9c56b854d746c047374c4eda43

Observation c99be82a-8dec-4b14-9ead-6d31a1e9b659 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.047936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.047936Z digest=sha256:5832c457321e27222dfd6ca3542a0232b39a84b826ff629c03f925e49ccb90bb

Observation 20ff8100-4ca7-47bb-9441-8c965fb14dc3 · outbound

This paper cites Training Strategies for Efficient Embodied Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Training Strategies for Efficient Embodied Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.209286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.209286Z digest=sha256:966c72fdfbbf9f3a73d1b18c0198bb6e99b7bc687945d97d8fd6b8468ffe3f87

Observation 6dd3773a-3739-4d0c-ac3d-15d3faa5cb56 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.316143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.316143Z digest=sha256:3e351b687396a0d4481aaff577bb808c99c2e8e775a7b6a04450444b3725211b

Observation c0535c25-b455-4093-be57-bbe23e6ed593 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.465144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.465144Z digest=sha256:022273bb30766aab1b152a1c12afb11ba3493907e7f48b9fa0503ba1f057ded7

Observation 8112ce35-afb6-4377-be70-593bac236238 · outbound

This paper cites U.; Akram, W.; Saoud, L.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions U.; Akram, W.; Saoud, L

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.586486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.586486Z digest=sha256:b4110cd55c915d7a9985c695acf57e0c60d53ed670dc85b55abb75a7b31841d4

Observation 9c7419bf-4469-49a7-aeb0-36cf0ae16ef8 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.753969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.753969Z digest=sha256:663cf3e86d24aaddf305fea600548bc0cc6502526c447e3615eb21aec934640e

Observation cbf66c8f-7dff-4bcf-af3d-5e294652bcc7 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.879712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.879712Z digest=sha256:6e941191b4629a415fab0ae79cef77c36c174c5ce1a2d5d32c4fce813c5a840d

Observation f22f2f3f-3975-4f85-88c5-a93078de177f · outbound

This paper cites Flow Matching for Generative Modeling.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Flow Matching for Generative Modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.984771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.984771Z digest=sha256:abccf67209e0ddfe0f65d233d3d278b1b9f9d0a704004a554344e9d5a0f64f4a

Observation 21b237a7-900b-4669-83fe-c4bb2e4577ba · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.125016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.125016Z digest=sha256:14096a96b26988c470be0b14170c3d7e6421feec03e7d58343eb6305c3cd9c90

Observation d2a5271a-649c-47e2-89a7-f093f194ff20 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:06.090482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:59:02.219607Z digest=sha256:029cf785b635ea3363bd366f6270bc1cbfa7d64ddb7248b247388fa05dc25d72

Observation 2d951652-97cf-4845-b9e3-30536abadda5 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.395778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.395778Z digest=sha256:a96fcf0f723e4f2694de25d4217ac8c519c7967f5b1c9a6c94d09bf723661c8f

Observation 3972e8e3-e061-4e2c-a015-369cad400f00 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.521887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.521887Z digest=sha256:3311f1f386cf556ecd56e8e4f67c15bb4a0ff399ff3da740004ccf7f44ab8a0f

Observation 1d9a072c-0eda-4c4d-b4ff-bd99563f9f45 · outbound

This paper cites C.; Hagenbuchner, M.; and Monfardini, G.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions C.; Hagenbuchner, M.; and Monfardini, G

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.669629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.669629Z digest=sha256:9fd01c77a719a9c4dd6de89635ebe571548cad12e9eeaaf0966672a69b442ca1

Observation 387a2d33-f234-4fb8-8dc5-7b234f572307 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Octo: An Open-Source Generalist Robot Policy

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.850855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.850855Z digest=sha256:ec6c086b22e2ef673ed3ee5f77afbe955d3cb986a68a434a88e7b461e1a8e71d

Observation 1eb08441-3332-416c-85bb-0788f87d9060 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions N.; Kaiser, .; and Polosukhin, I

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.028529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.028529Z digest=sha256:77f8b25518dbd305234f424f08814f3144ce0b84c116f374b1e3e967016a755d

Observation 46b3d570-56a7-4917-9861-7454eaa9e4de · outbound

This paper cites RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.238165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.238165Z digest=sha256:5516282b356bf945eab0ae43d6065713b035196eb6596d90e138365bad328e7a

Observation ed56506c-f6b4-4bdb-9e27-8837f393f21b · outbound

This paper cites V.; Zhou, D.; et al.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions V.; Zhou, D.; et al

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.360868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.360868Z digest=sha256:aa0a5782cd8e7702f7e49c7e20661462edf53b21f1315aaaa5f083eac038237f

Observation 8a3068a2-2413-4b17-893b-59a5f87d8200 · outbound

This paper cites Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.506693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.506693Z digest=sha256:908992ec3a5ef8338beb18b9e82d6ce9647c8ec3a51cf584b5a16073ac8b2563

Observation c37b3efa-6603-4150-a9d7-3b8038644537 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.685804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.685804Z digest=sha256:9d362a82cdda96cb39012b1bee949143bc31b0beee4b64ba0975d67def05fb6f

Observation 937d58e2-100c-4b34-8746-37cb3ae1f068 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:05.824403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:59:03.815934Z digest=sha256:72e45af78d61f109921bdfa0e7820c36365c94fa47f0a2c5cef76df435212f5e

Observation 607ad875-3aa4-405e-8f2a-2638e625de80 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.989579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.989579Z digest=sha256:8f569eca38d68f9a74712c8345cf3f99439e5c5b52d1251dade6132a9e46aa06

Observation 10fcbea8-bb85-41dc-8eb6-083a24f9dbee · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.110925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.110925Z digest=sha256:f5046d4f31f6e104ada1bb3d13acc195962181c2f4584fe7387e526a6be0ea75

Observation fc824915-7486-47f8-a96f-94bace06dedf · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:05.549634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:59:04.244235Z digest=sha256:8ee2f6b6bd5a4778ce783d43979cafa4beadd5d9cbfdec41e0ed90a0d4de1a66

Observation b6b0a23a-ae10-4209-9a43-49c754fc910d · outbound

This paper cites J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; et al.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; et al

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:59:05.235334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:59:04.376600Z digest=sha256:357cdd8373357d37b496f6f264f4c3e194251670f7952c2a19f6bea504103aa7

Observation a04cce51-79b9-43fb-ada7-aafa8657335a · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.551521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.551521Z digest=sha256:8cd32340949601da89d0add12ff54cd43575ae065dcac88eae3e463664c2c8da

Observation e7a23773-4d3a-433b-bee1-e3af492139df · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.663968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.663968Z digest=sha256:35fcf42a846813940758b16ca5c6c8eb6bdc97b6aa0a6025a6479315056a481b

Pith citing papers

Observation cfaa48a2-6d08-4c1e-8e5f-e7a9f4ae9992 · inbound

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer cites this paper.

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T17:38:20.785896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:38:20.785896Z digest=sha256:897a9fff75be185f80f91658b9a3cd01715ae6da7ca9ef793b7a19df7f55eeea

Observation 6f36ca5f-1c77-4b7a-acf9-74fb39086d8e · inbound

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models cites this paper.

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:57.258483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:36:19.995405Z digest=sha256:771a3445b6e238405eb271c72458b8411bb164645e3f7d397d9a4f8eb7f19510

Observation 88d240b7-7eaf-43ec-a10e-27dd97ffe09f · inbound

TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation cites this paper.

TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:09.179305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T14:54:48.058400Z digest=sha256:ace6e69ca81325d220135f9dfbf464cf0976433ca778e7e330ab0a1d27f06cf1

Observation 7b5158b3-e675-40d2-8a40-aa4907ff9396 · inbound

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction cites this paper.

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T04:52:17.025802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T04:50:07.792757Z digest=sha256:f0a0240ec941813d7dc572efb10c053a4174ebd1243eec691c55851be804eb7b

Observation 5f93b8ea-0f8e-4842-9f73-437ff3206344 · inbound

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data cites this paper.

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:57:33.311170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-14T17:54:50.325820Z digest=sha256:082995cde77338b2f8a56312b4d9ceb11f7fb0f606f20ee910386f72b9a2dd99

Observation 963a8b05-0fcc-418c-907c-038d6af2d935 · inbound

DSSP: Diffusion State Space Policy with Full-History Encoding cites this paper.

DSSP: Diffusion State Space Policy with Full-History Encoding GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:21:24.247338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T10:17:44.804745Z digest=sha256:80fdc0fd2b941019df6901a14df625fbfe054fb694105463b4b3ab435a73fa28

Observation 8af57b5b-6e93-4cc6-9456-acc96e48fe0f · inbound

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models cites this paper.

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:01:20.749485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T21:57:19.195413Z digest=sha256:5c1d5ac89f2c880ed4d02866f69fa0f0a8034ee697b503535666b4459c3f6ddb

Observation 4a2bf18c-6494-44e4-b6a3-ad0de77e6d30 · inbound

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation cites this paper.

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T06:36:27.368310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:36:27.368310Z digest=sha256:aed3c3a710df3c3a8f38504ff7eebfcf1e5686251940fc50541911df695cfb36