Pith. sign in

Paper Citation Record · LEDGER

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

As of 9 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 12 inbound Pith citation observations for arXiv:2506.16211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16211 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:51:05.564651Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:08:36.570832Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:30.514537Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53acf680-4cd6-42a3-9738-8f482421420f · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:04.845336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:04.845336Z digest=sha256:bff0a73ee9d2d9c74d87349985bb087f8dadc57e2430b2c9d7c699f645a06eee

Observation 24b6b7d9-3775-4067-854b-24b403a0edea · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.413572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:04.971311Z digest=sha256:42d89b2cd92c33f77104bd519cf27c3df2326e48e56ee330058ada58d2e70b97

Observation 3743e857-a31c-4f6b-9dbd-c97790271b1d · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.020892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.020892Z digest=sha256:273842a66439fc05ce5e92bc78da3d45dd535aaedaab05718880e363881f67e2

Observation 10ec7206-8122-40ec-aca1-ab548f5291da · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.398694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.148613Z digest=sha256:b35a63ab17ef8ef5f4233e1df66d0e5993f2afa294693160625716fd5eb462ae

Observation 6709822b-3d59-42e6-a973-a5af90811532 · outbound

This paper cites G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.241297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.241297Z digest=sha256:475bcbf2564a40d08912d6bdfcc4349178cd5ae242c61bd1add9c70928a0ce2a

Observation aa23ab53-82a4-434d-8ab1-a3de0f8373d6 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.384171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.305317Z digest=sha256:2621b7bd14849e41dba05654d0b2db9f14d3c5ff29cae1b0d82dab3738bb051c

Observation f102da75-3367-41a3-bebf-4a1a35638799 · outbound

This paper cites SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.310332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.310332Z digest=sha256:2d522fe6abc49f62fddbaecd72d9bdc2dd4065b501c04f999bda112c185a15a8

Observation b9edca07-ac70-4075-a75c-2eb00d38d7b1 · outbound

This paper cites Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.315137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.315137Z digest=sha256:75e5b9a5d2a7a8729939056f56db0021f36c69f0371b9d5cbdf7a0962a49658f

Observation 12c2b4c4-c24b-4818-bc06-6557efe69113 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.370196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.320125Z digest=sha256:b45ea22039164989d925bab92ae30e11e4db27c28cdf9cc3cd84f12511cbcc43

Observation fb141498-28bc-4ab8-8223-bd5e1e935b2a · outbound

This paper cites DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.324527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.324527Z digest=sha256:4856559e99fad85645abd43dcc7c6ffbad69541eeca48152ea017b36c6493b15

Observation b67b89f6-dba4-4c1d-859a-a2eb8a8d0e91 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.329116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.329116Z digest=sha256:84bf18072b82b46ac9d609cdc057e912f781441c78ae92b705c4ee15e999589b

Observation 5434ca92-d28a-4d43-a5bf-a5ffb9011982 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.355296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.333582Z digest=sha256:e68f3ab45a5acb154f0656d8b88048b940111b1b660ddf1d1d043c89eaf7e79c

Observation 2920bfd2-698c-417d-b44f-3398e71fae21 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.340945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.337888Z digest=sha256:a82ce7899b1f426f55d596332cc40395d6db658e5908af21109d8d35b7920e79

Observation 0a42fde1-05fa-4965-9ab9-194ad301421d · outbound

This paper cites Huang, S.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Huang, S

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.326178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.341678Z digest=sha256:780cbe447383fc7fbc3e2c1a93acd65b6371a29233bb549b4002ff05326203f5

Observation 44ca8117-ffa3-4ba3-8b2f-4f61a5c76260 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.311543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.345758Z digest=sha256:7bbd7a97c09af7e93d3944848af42eb9c0989825c19ebf956d924f7d7d00df28

Observation e36abc5f-663f-4985-9208-3b9a1bc7c1bd · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.297310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.349416Z digest=sha256:52b10f2b6c0623f9a4580a5e08bd23b8c130145c5dd916e6e0ff924c09a94848

Observation 6a6d9936-f27f-4121-9e17-6715b23cb50f · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.353073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.353073Z digest=sha256:cb36dcf9f31583929c8b83758b39f1824997a6a08f8d44525c242cc3ebbd5907

Observation 998ae332-334f-448f-b49c-6788db8aa696 · outbound

This paper cites MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.357203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.357203Z digest=sha256:390ba1018e0b636a823ba451e694a2339be68732a60604ae1daea357013b9717

Observation ffc1e156-24ae-42ec-920f-4664b4d2668e · outbound

This paper cites RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version).

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.361677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.361677Z digest=sha256:bd33d625d4d8a4e12acdeb67f76a7bc44f0dda1977c2f6879441ab87574bdfd7

Observation ae9e3d4e-b8a2-4919-a1c5-79bbf547d64a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.365914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.365914Z digest=sha256:5548f7bfcccc080bcd500522db56093563870ac1722af1fda427f8150ba9e0b8

Observation 71d98caf-432d-4063-afd0-bcbc680928d3 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.370129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.370129Z digest=sha256:3b22b6d1f687c13401bec7b5d1510e76d540952f33368f94d112b2ed06836ec3

Observation 781c4e61-c402-41b0-8f99-a5c44169bc12 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.373806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.373806Z digest=sha256:d759b1850aabf83c8d4a5b19bfe2516d5cc833bbb71c9b53348049cfa061ec4a

Observation c6baf59e-1106-4231-aad2-8e8aeb2c6d45 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.377722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.377722Z digest=sha256:f96a45a6ec87b68f084ceab1e7e78feb5baab72bf23897fa8c7eb298f839a3c1

Observation 6f9c26fd-89b7-4d39-851d-3702d09d5286 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.382153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.382153Z digest=sha256:45adbfe18d7c710c201eea56db23455c5191ac833ff297a5f9cd00d33e093bd2

Observation 2089d7f5-93ce-4dec-9f61-c58a2b0a3e7e · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.386011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.386011Z digest=sha256:5923936f0fdf012a446d9fd2aa4eaad012c48577e541b0373e5a8d2ef8ada42b

Observation e4106095-b4d8-467f-8a49-324fed201f2e · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.390144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.390144Z digest=sha256:e4a588873fbcfe4e9fbc3c0b1e749f5e33f5f4996fe68720fa6a82ba2b2e934e

Observation a7eefe94-574b-4443-8831-d84e432eb660 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.266276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.394404Z digest=sha256:b5a0af23b2f79f577e299533a6689047983eed0ed39b4706a8c8b30d067116a0

Observation 2776f220-336c-4073-b42a-25e4523a694b · outbound

This paper cites Zhang, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zhang, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.398711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.398711Z digest=sha256:ada785fb45913459a5fdfd0a40351ff854b50bc1eac65bcef6aecc061cdfc2d1

Observation 31347ab5-89e4-4054-9245-a85a0314e0e4 · outbound

This paper cites DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.402554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.402554Z digest=sha256:00c0f4a6359824aa6fbc3642e39da0b8c988414ae7baaa6efd1c610a71469699

Observation 57f1110b-73d2-4ee7-99aa-ea5662ce3d31 · outbound

This paper cites Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.406504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.406504Z digest=sha256:278111c7974d2335e14b64ebb23c89ef4856516edf2b20dcf03616475ef51cfd

Observation ab0881a8-66fa-4b01-8888-c298d62784f4 · outbound

This paper cites Tyree, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Tyree, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.244679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.410625Z digest=sha256:38dbc4da94527918beea144ef26a6b1ace051845f4c765ac800d3fde1f5b2f82

Observation 6027aae2-becb-4268-bb2f-2747a4bc7ac8 · outbound

This paper cites Migimatsu and J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Migimatsu and J

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.414591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.414591Z digest=sha256:b77db42401692c42c98f27b869b173ecfcd13522ab97a3a8adfbf76439750a2b

Observation 69e1497f-f748-45fe-a3b9-700c6d65cb3c · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.222540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.418384Z digest=sha256:d501dad81e2a0e8c8f901dc37153df483af4b4c8a135d0bc1912d18e5b2da652

Observation 37e04de4-db34-4e8a-a077-eb8902a63ecb · outbound

This paper cites Devin, P.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Devin, P

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.208714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.423236Z digest=sha256:8acb85fbe0369067c4eb5f7bb384fa2209026fadce6157255e3098f3df60927c

Observation 6644594f-06b5-4192-9696-e405ca079be9 · outbound

This paper cites Locatello, D.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Locatello, D

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.194944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.427977Z digest=sha256:9732b8e7efcc5956a011431cbbef6ec3e59fa8ff856a8d4c26a2e1b834d6a4de

Observation 11d0a78b-0401-4773-bed6-2bbd59e90b06 · outbound

This paper cites MONet: Unsupervised Scene Decomposition and Representation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MONet: Unsupervised Scene Decomposition and Representation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.432378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.432378Z digest=sha256:38000a3e8cc1249895b6332e0d5b4acab5f98cb27e8a68c42010d2949891236a

Observation c012871e-e0af-44e4-895a-ac63d00c216a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.180647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.437120Z digest=sha256:a424ff5a7dddcf5e0b7318ed705b63dd6d7f0b47c343a57516632c5554d34098

Observation 7b2b6983-52c7-419e-bb41-46b3f5252e36 · outbound

This paper cites Heravi, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Heravi, A

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.166949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.441912Z digest=sha256:71859e9e7b023ac51014d2e737582eb0282b672967b42146e5d645efb234a009

Observation 168b2c4f-2d25-4599-86bd-93423394dbd9 · outbound

This paper cites Zero-Shot Object-Centric Representation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zero-Shot Object-Centric Representation Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.446892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.446892Z digest=sha256:4ae1e0765cda0564533edd30643d4b38234667b9714a3d83c3f41a5936dda25d

Observation 45466b54-eaf7-4a98-8d1b-44323e663ace · outbound

This paper cites An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.452350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.452350Z digest=sha256:1f08abbb87fc1a0c77cee4e10abf7649a61593673ab2fde599cad33b9cdbe24f

Observation 5ae50293-7817-4635-aab7-310d03f9b75f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.152241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.457334Z digest=sha256:58bdcae316a4bc6c22f95002d6afa400de31a56c6234f95ffef2adfea3aea121

Observation 1c96920f-812b-42e0-bed5-ede0ebce4e50 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.135773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.462541Z digest=sha256:2f07fcbdf4063ee1aee9b29616d94ecf495ff5215281fdb73ff77baa955d3741

Observation 9c1fbef4-b8ec-4ff8-882d-0f8b321c677c · outbound

This paper cites Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.122138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.467010Z digest=sha256:c85f9134521348fff60ed245a17fa42934078be23834aade0e3db1bacc69365c

Observation 9d5454d4-9a24-4458-9b93-c7625e4d3ee2 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.107327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.471119Z digest=sha256:f68618294a5e1e7db8ca166804372cb9dcd9698fb93e7364498b5b154fd3d28d

Observation c1a03e2e-f18d-46e2-be08-e8a6d00adf5b · outbound

This paper cites ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.475594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.475594Z digest=sha256:9fd0e587012d4c02a56c59421d937838b66e2fe4032d165bf3baa72d167ca698

Observation 6802722f-9f79-47f5-bd9f-89530512801f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.092961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.480005Z digest=sha256:7efc291d8d50d262620d71863db1a2be633c3aa401d03b91e18f6fdc9c86bea6

Observation 08c6105a-ee10-4731-8c6c-afa6600bf20b · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.077473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.484539Z digest=sha256:362c740a523421cbb9956bd4e7678fd68786323d012060cb097a78e22520e40e

Observation 73232784-6260-411b-a695-60a0ac9a5d2d · outbound

This paper cites Bar-Tal, H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Bar-Tal, H

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.488962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.488962Z digest=sha256:ec009f9cf26e9c8342c3210796b0df32e6d502c691e22d6578970d8c4f903f7c

Observation 8814db4c-fa02-4477-bad9-47d83a39da8b · outbound

This paper cites Dai, L.-H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Dai, L.-H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.053967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.493365Z digest=sha256:bddbaa9df416f9bdf452c679bc1bef2ae8ec22d8cbaae0e96f1f0b816fa2ca69

Observation 1fef668c-b1ef-4f1a-87a7-07fbe2a7171d · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.039106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.498683Z digest=sha256:9709f2b2bc1556bf007530c261a0bfa9d94fe50203df69430ba28d7a1c4f8212

Observation 094b105c-41c1-4408-b040-87c109eb108f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.503078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.503078Z digest=sha256:133925ab56bc4ea360739700ce216c3f58d3e7168b7d56c78c2b62a50f1897ce

Observation 0eda8c59-061c-45c6-a4ad-bc6afc8f021f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.508411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.508411Z digest=sha256:b74d86946cdf1955fd4797d7dc1b82505c6c8e439d6c4dbd8a4a1b6e1b84dd8e

Observation 257b6e7f-9da2-416b-8d1d-dd1e5d1e21de · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SAM 2: Segment Anything in Images and Videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.512735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.512735Z digest=sha256:4ad072494c71f1ff62fc85ebf8bce0cf665249aebdc6f72122f3b7045cf8c14c

Observation d5638925-aebe-47b4-9048-97992808df66 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.516993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.516993Z digest=sha256:9e4bbfc3e60b35c0be792b03d09a86af43d54a83ea42049cdf7a8c5669ba4ac5

Observation cb770695-661f-42d0-b598-c3c1fbeca1ab · outbound

This paper cites Krizhevsky, I.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Krizhevsky, I

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.522115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.522115Z digest=sha256:62181bd8c5a0f1ce04c502c4727d5da1ccb2a28662692d791fe315e4ea9bf0a0

Observation 9f592a5a-79e5-4b5f-9741-1bc67c350e3d · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.527593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.527593Z digest=sha256:d811d8fa86ab54be85cf987cdd395887853835269af64d5bd4b1ffa17544b790

Observation a31165a0-14fc-47c0-90b5-3d49cf2f0175 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:05.990154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.533167Z digest=sha256:ed6af62d0e12eadba7a87dbd711cd4d5d98c300ff50e7314144e39cb8b2a90f6

Observation cb1f02cd-b236-4b34-b417-427e5d224f48 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.538786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.538786Z digest=sha256:57c1329306ae6262fd94abaf00f13a1ea6cac6211412cde1343a1f5494e0a938

Observation 53e4b447-c84d-4971-a6ea-afd045c95f2a · outbound

This paper cites Khazatsky, K.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Khazatsky, K

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.543711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.543711Z digest=sha256:f178da20c23ad57e5ca03d0af5aab65ca3b21974d2c630507d226e27f0be963c

Observation a145d811-0bbc-4c71-9008-28c2e2228016 · outbound

This paper cites Radford, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Radford, J

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.549380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.549380Z digest=sha256:11e2bf4a8ff6f295bcc87c846734a111123de1a41b57833ba669f17d20ecb7c7

Observation 02556ad7-f4c2-421b-99d0-8e270d441f1f · outbound

This paper cites Campos, R.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Campos, R

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.950992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.554756Z digest=sha256:3ac3abcd19d5988b165274e2b101f4d3ea7d3c17cd3e7e491dfff9a844a23798

Observation 6177047f-aa54-44f6-aea6-ef3cb1efdf54 · outbound

This paper cites Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.936241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.559399Z digest=sha256:1bb6a28f36c9651ed617aebb64722fd302f3aeb9c9606214fdb824ae8c108020

Observation 5fdcae0a-3e61-4d20-a971-52375e32b597 · outbound

This paper cites Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.920767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.564651Z digest=sha256:0059a4b126015ee2c963458adfdedb90c5ea050c4f6cfa086b03414baaa1f0ad

Pith citing papers

Observation d33b1b39-c1fc-4205-be93-e74a269e4339 · inbound

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation cites this paper.

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:05:28.900178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:05:28.900178Z digest=sha256:0c01d6004ecd296bfc3b9c04cb3130c221f841024604c09f0d721a74fcc12646

Observation fff67d24-b737-4cff-8fe2-ff062d4198d9 · inbound

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation cites this paper.

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T10:43:40.991161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:43:40.991161Z digest=sha256:36330acb74a2e7c4c903e20bb0c4347496dd2a76a94796b431f5c31f5852305f

Observation 028fdbbe-87a4-4252-9064-713e605b5d95 · inbound

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation cites this paper.

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:25:32.979738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:22:41.611893Z digest=sha256:1d728f46b87b77adfbbcab94f563de4f64b932d5b3158d1e5cf87d3f9c824126

Observation 8fc455d9-e6c7-44c4-8e69-2d599f1370b6 · inbound

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment cites this paper.

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:52.618248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:27:12.286456Z digest=sha256:3de03bc1b546efe170567d12000858925105b63327dbe80361a58b2af067c531

Observation b58a1dcb-a01d-4e5f-a51f-6b8a8fc1da92 · inbound

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation cites this paper.

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:19.021480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:50:33.134927Z digest=sha256:a8fcd7727ca21c859cb0bf091b9d828be4fcabbdc097fd004549244a441ff1ff

Observation c4febba7-8387-4ccb-bdd3-5d0ebb6ffe1f · inbound

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training cites this paper.

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:11.595737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T08:07:47.793895Z digest=sha256:6e426ee676760ddb15dae7c0239ac39a7883f2d37196cef4667a0e604cd435f3

Observation 5813eb5c-32b6-470d-a900-9fd4888a5281 · inbound

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning cites this paper.

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:22:14.583427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T04:22:07.953637Z digest=sha256:85b0798f2f39fa976cac9f766838717d7bb23827ecba703ffe76e1dccfafb047

Observation ed470334-f0ba-422c-922a-939f8f57033f · inbound

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation cites this paper.

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.455458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T06:57:41.245418Z digest=sha256:c713a256996560ce7e616f29d4ae5fb00bfef14ee19c92c06d0a9f8282e8f971

Observation 27e3a5be-f2fb-4e2c-aae8-907fc7f80c20 · inbound

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models cites this paper.

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.825517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:55:53.760595Z digest=sha256:b728c4de50f7f0eee39f23ca9bacaecbba2c6d3e74e852e4966a62cef4143b77

Observation 19db67d5-ffa5-443a-90d6-5beb6bb7b0f7 · inbound

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining cites this paper.

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:37.882421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:23:36.941801Z digest=sha256:39c3303601dae5ba5ff9c50a595f24a6ca0f928edcb5e4706981c66d24abb7b1

Observation b2fab49c-65e1-4a40-94a3-3accfa9d07fb · inbound

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation cites this paper.

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:30.517227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T18:01:08.616612Z digest=sha256:d574eeac15113bec068e2c8f8501835a33cc1e1cad9f3fad5d50026a5e27be7d

Observation 49371b32-4dc8-43cc-a7d5-2b9dd8e113be · inbound

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation cites this paper.

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T20:08:36.570832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:08:36.570832Z digest=sha256:1282d00d30d88c1111e86e9a1f4c85d2ab3c2fea0b70bca098879cbc286e1827