Pith. sign in

Paper Citation Record · LEDGER

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

As of 14 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 12 inbound Pith citation observations for arXiv:2506.16211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16211 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:51:05.564651Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:08:36.570832Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:30.514537Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53acf680-4cd6-42a3-9738-8f482421420f · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:04.845336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:04.845336Z digest=sha256:4136ffc9beec592f1866360e2095618b4ec6c503185081cd9041ea6bfa88ade8

Observation 24b6b7d9-3775-4067-854b-24b403a0edea · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.413572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:04.971311Z digest=sha256:21fe203d16ede1afc3884db64cf646f9363ee91fd566e98b557a1aef9ccc5f7c

Observation 3743e857-a31c-4f6b-9dbd-c97790271b1d · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.020892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.020892Z digest=sha256:d2d727c2b2e0757866d4d602ab3f4ad537f3cc9610e1a530ba0ce27cef2dd332

Observation 10ec7206-8122-40ec-aca1-ab548f5291da · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.398694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.148613Z digest=sha256:df8629c4e74ddf3fcb16a162af5924079a24a2a1ebc749a908778b19a6afb026

Observation 6709822b-3d59-42e6-a973-a5af90811532 · outbound

This paper cites G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.241297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.241297Z digest=sha256:d0a08d02d288836413ac9e8c6a1da39735dda7df6172e839697798cfde35af37

Observation aa23ab53-82a4-434d-8ab1-a3de0f8373d6 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.384171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.305317Z digest=sha256:3aec26e93d49a99584a4e8d0e7ac29701087f83fe6386b960d93e2c0c1c11359

Observation f102da75-3367-41a3-bebf-4a1a35638799 · outbound

This paper cites SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.310332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.310332Z digest=sha256:5414a5e6592d0e1c53f71b41127566793af5331a613274c9afb41b130832f6e1

Observation b9edca07-ac70-4075-a75c-2eb00d38d7b1 · outbound

This paper cites Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.315137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.315137Z digest=sha256:25d4516f989fa4f66b6c769ad201202447ee10f258b9fbcd6c5560e052f319d9

Observation 12c2b4c4-c24b-4818-bc06-6557efe69113 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.370196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.320125Z digest=sha256:68dfea89a7454d9c31616840a29b6cfc527ff6bf315ca1c3f4f316faa8a66638

Observation fb141498-28bc-4ab8-8223-bd5e1e935b2a · outbound

This paper cites DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.324527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.324527Z digest=sha256:50e75ea2bf61615c8ae8257bba6e76f9c6cfbc8102f5911a1cdc46bc0e28714a

Observation b67b89f6-dba4-4c1d-859a-a2eb8a8d0e91 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.329116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.329116Z digest=sha256:ab098128120e172dac774ed3277b0e7d3e1d2460171bcf696eabc014170f8936

Observation 5434ca92-d28a-4d43-a5bf-a5ffb9011982 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.355296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.333582Z digest=sha256:7ee03e1195a3707a2fa6f230c9aa8ac10c019944a815302b98bb52d4540979ab

Observation 2920bfd2-698c-417d-b44f-3398e71fae21 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.340945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.337888Z digest=sha256:8e68dc279398c23594ac7eb8bfbf92133195c3f7feb6fe65b93425921050ac38

Observation 0a42fde1-05fa-4965-9ab9-194ad301421d · outbound

This paper cites Huang, S.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Huang, S

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.326178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.341678Z digest=sha256:359c852f18603af123421a2bf8d3eab0456d921cea378daa1d892bd59f66ebd5

Observation 44ca8117-ffa3-4ba3-8b2f-4f61a5c76260 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.311543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.345758Z digest=sha256:59af927a535c66492aee7e626f1557d5ae380964bd92085d50e5a627b111a8b2

Observation e36abc5f-663f-4985-9208-3b9a1bc7c1bd · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.297310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.349416Z digest=sha256:f76b1b7ce92499c4b42a20e79743259cb30fa3a69d8a8e3d4ac3c7b9878c9409

Observation 6a6d9936-f27f-4121-9e17-6715b23cb50f · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.353073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.353073Z digest=sha256:9cbb4967a0296b914512900fcae6124357b6ecb794d53a95c59589b01394ff4d

Observation 998ae332-334f-448f-b49c-6788db8aa696 · outbound

This paper cites MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.357203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.357203Z digest=sha256:ca71b89fda36982839823e723dc2471b0fbe7f29f772a15c83d5e6b90d0fec81

Observation ffc1e156-24ae-42ec-920f-4664b4d2668e · outbound

This paper cites RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version).

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.361677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.361677Z digest=sha256:cdc580be450c50067066fa815e4c90d576aae61170b83aae0395466e19bf4de7

Observation ae9e3d4e-b8a2-4919-a1c5-79bbf547d64a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.365914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.365914Z digest=sha256:87fdde001fc06656ad522e67dd8783c961cd8670f2cec19f196f6ef93863a9e4

Observation 71d98caf-432d-4063-afd0-bcbc680928d3 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.370129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.370129Z digest=sha256:962de5be06628b16ad0fc1f9280651c164007ecbf318309056211f8887a7e10d

Observation 781c4e61-c402-41b0-8f99-a5c44169bc12 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.373806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.373806Z digest=sha256:9ab18e71ebda6660eb1c911e7e4441924174c142de7bd926107488b3c8c72641

Observation c6baf59e-1106-4231-aad2-8e8aeb2c6d45 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.377722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.377722Z digest=sha256:7b1075e8c988a765531eb62f6f9397286f4fbdd9b97d87b27436e200a62f588f

Observation 6f9c26fd-89b7-4d39-851d-3702d09d5286 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.382153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.382153Z digest=sha256:d15bc825f7d105be166e4b7777d6b03609f785fe034e9278c2bba278982673e5

Observation 2089d7f5-93ce-4dec-9f61-c58a2b0a3e7e · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.386011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.386011Z digest=sha256:479e22d9a2072c9b03ca62ff630f6c3f88f38cc64f72e442c854538cd908ebc3

Observation e4106095-b4d8-467f-8a49-324fed201f2e · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.390144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.390144Z digest=sha256:e3104931629dc910baa5a52b4f1bee731791deec8e1591df2affc82424695859

Observation a7eefe94-574b-4443-8831-d84e432eb660 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.266276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.394404Z digest=sha256:86733a09105e4fa403e1dbc54c53df14a4b49e9b7e545f8cc0b99b9509548807

Observation 2776f220-336c-4073-b42a-25e4523a694b · outbound

This paper cites Zhang, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zhang, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.398711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.398711Z digest=sha256:5a97edf7104cabff3445cb57feae6201c40b9d9a125e8f7ff92bc6f802486c7c

Observation 31347ab5-89e4-4054-9245-a85a0314e0e4 · outbound

This paper cites DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.402554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.402554Z digest=sha256:fefd13f1e5ace956f97af4edc2e46c7bcc5f5e69b35716fb542c9afd4f154d7e

Observation 57f1110b-73d2-4ee7-99aa-ea5662ce3d31 · outbound

This paper cites Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.406504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.406504Z digest=sha256:58a19a74502bc388c52bbe0aae358c0fe26102c794bc2ce0261c1893250a4dca

Observation ab0881a8-66fa-4b01-8888-c298d62784f4 · outbound

This paper cites Tyree, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Tyree, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.244679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.410625Z digest=sha256:83b333e6a0bf3bede1b9cd26c80ac10c71e5311bcc72bb51a308554937fbcf92

Observation 6027aae2-becb-4268-bb2f-2747a4bc7ac8 · outbound

This paper cites Migimatsu and J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Migimatsu and J

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.414591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.414591Z digest=sha256:0462d89b4d7a2895b22c6f25b1fe46e2b1ee86ac36969b6b97e239ee9676745a

Observation 69e1497f-f748-45fe-a3b9-700c6d65cb3c · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.222540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.418384Z digest=sha256:f67ee19fd9ade8fca744d7e03aabd19442b10a5dd7d9fe34e4595a21c390bb68

Observation 37e04de4-db34-4e8a-a077-eb8902a63ecb · outbound

This paper cites Devin, P.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Devin, P

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.208714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.423236Z digest=sha256:af422a48c4a6ddffd6f01b653ed234003f6cf15859d0c29f0d6e7d708314788e

Observation 6644594f-06b5-4192-9696-e405ca079be9 · outbound

This paper cites Locatello, D.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Locatello, D

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.194944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.427977Z digest=sha256:a9fe3b64bde203af653592ef5ec854cf6f8446a77057efb482930a5d58864639

Observation 11d0a78b-0401-4773-bed6-2bbd59e90b06 · outbound

This paper cites MONet: Unsupervised Scene Decomposition and Representation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MONet: Unsupervised Scene Decomposition and Representation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.432378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.432378Z digest=sha256:34e055b154954b218a7b22587135455bcc6b7db8472eeae1f720c7e2498ca12f

Observation c012871e-e0af-44e4-895a-ac63d00c216a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.180647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.437120Z digest=sha256:a44b6425a34ebffe9bba6fc211c8052a524c2d2c1ed2a3cfa56bbb6c80367525

Observation 7b2b6983-52c7-419e-bb41-46b3f5252e36 · outbound

This paper cites Heravi, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Heravi, A

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.166949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.441912Z digest=sha256:e33c0a15623d9042c4e938cfd0ce9cca3ad045d5e10ffb52e4898429f4f6ebd2

Observation 168b2c4f-2d25-4599-86bd-93423394dbd9 · outbound

This paper cites Zero-Shot Object-Centric Representation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zero-Shot Object-Centric Representation Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.446892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.446892Z digest=sha256:f6bb4d3c322043ec557b52ac5ef6d40d07679d91e6e4a5a946b3cb7c6ff13896

Observation 45466b54-eaf7-4a98-8d1b-44323e663ace · outbound

This paper cites An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.452350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.452350Z digest=sha256:98e9b81582e3fe1dd2a067e58a13e77721203498315c6ea489c15dbf3287fbd9

Observation 5ae50293-7817-4635-aab7-310d03f9b75f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.152241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.457334Z digest=sha256:f0e428956bf34ade0ea410a34d2cb07d5d94a3080302cc753acf9c031b855436

Observation 1c96920f-812b-42e0-bed5-ede0ebce4e50 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.135773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.462541Z digest=sha256:897ee7e9e254c8e4c84642f6dfc38128c9cb1cce8005b20dd36db18ae8555b20

Observation 9c1fbef4-b8ec-4ff8-882d-0f8b321c677c · outbound

This paper cites Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.122138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.467010Z digest=sha256:88a6103412a73119a2a1221fbed6fe5c0604b04362ae8885720af3bbd071f88e

Observation 9d5454d4-9a24-4458-9b93-c7625e4d3ee2 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.107327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.471119Z digest=sha256:6f0863cd23565327aec7dfd57f3b8e474c147d72558b594ab6e607ba8aeb8f1f

Observation c1a03e2e-f18d-46e2-be08-e8a6d00adf5b · outbound

This paper cites ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.475594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.475594Z digest=sha256:fe9d93e592431b7b505ce089e66db564cbd238ce3a169e1ad4b727072e94543a

Observation 6802722f-9f79-47f5-bd9f-89530512801f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.092961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.480005Z digest=sha256:e7ee3a2d0129d408febb7dcd6e9a132311f18a1865785b5aa30484c80be178cd

Observation 08c6105a-ee10-4731-8c6c-afa6600bf20b · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.077473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.484539Z digest=sha256:9f3d7a30832633a3e1c5cd769d5dfeabd3bd62df727d9270da50efffeb33cae7

Observation 73232784-6260-411b-a695-60a0ac9a5d2d · outbound

This paper cites Bar-Tal, H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Bar-Tal, H

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.488962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.488962Z digest=sha256:724e7a5935e503856d680a85a9e43b15588690db2d2725512b4903a3d638818e

Observation 8814db4c-fa02-4477-bad9-47d83a39da8b · outbound

This paper cites Dai, L.-H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Dai, L.-H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.053967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.493365Z digest=sha256:2cf8938465fbf77cf1140001e81ae29a92f3a52bff4d5c2967f249dc6f7ba11a

Observation 1fef668c-b1ef-4f1a-87a7-07fbe2a7171d · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.039106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.498683Z digest=sha256:4f82502702e23daffd9e9083dcff097c98fcaa7836bdc8a450779c437dd7e92a

Observation 094b105c-41c1-4408-b040-87c109eb108f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.503078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.503078Z digest=sha256:d955a4547ec573473e350fa3d7ebb08b53ce8863b43485dec85189b5535cb7fb

Observation 0eda8c59-061c-45c6-a4ad-bc6afc8f021f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.508411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.508411Z digest=sha256:e0a3b5940e65833b5e0bd645e4bc04c8f0c3eff2d2f7107438032ae642cf71dc

Observation 257b6e7f-9da2-416b-8d1d-dd1e5d1e21de · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SAM 2: Segment Anything in Images and Videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.512735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.512735Z digest=sha256:35ecc6b9fab5dfde7429ceb4416a4cd8ec38ae14fbaf49b8777741a04c8cf9c3

Observation d5638925-aebe-47b4-9048-97992808df66 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.516993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.516993Z digest=sha256:76f44911e1adb2bef1c5d6b0e4e7ebd9cbaa39b20e6865dc51d76d930fb66c05

Observation cb770695-661f-42d0-b598-c3c1fbeca1ab · outbound

This paper cites Krizhevsky, I.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Krizhevsky, I

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.522115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.522115Z digest=sha256:cd4695f797c989868c07f8f6f78ecbf0e3e789aa3eb6b00e40774dcad12157ed

Observation 9f592a5a-79e5-4b5f-9741-1bc67c350e3d · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.527593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.527593Z digest=sha256:80a0ec21c2a97726758ad9201689fd9a9bb5f1ee44e53fe090e690f0fd556b93

Observation a31165a0-14fc-47c0-90b5-3d49cf2f0175 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:05.990154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.533167Z digest=sha256:ff835ffb6387a01c1dd5be720d8198bd04474fa2e5a209aec6cbae92847faf68

Observation cb1f02cd-b236-4b34-b417-427e5d224f48 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.538786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.538786Z digest=sha256:9b4d1c4112bea0b687fb75a4174728f42d7b02d1211550bca4737527b82c77b2

Observation 53e4b447-c84d-4971-a6ea-afd045c95f2a · outbound

This paper cites Khazatsky, K.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Khazatsky, K

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.543711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.543711Z digest=sha256:ad57640debb4d025148170974a495286f24f6f11debffbcdfad4d431c650c818

Observation a145d811-0bbc-4c71-9008-28c2e2228016 · outbound

This paper cites Radford, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Radford, J

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.549380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.549380Z digest=sha256:5cf64d5dd9d20f2e4f6e89a94879e0849e895bceaa4074a9b25b0ffc3bbea824

Observation 02556ad7-f4c2-421b-99d0-8e270d441f1f · outbound

This paper cites Campos, R.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Campos, R

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.950992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.554756Z digest=sha256:b2376796f68c5f3de271c9f1f782b3185fad92613ad37b9ad395a0c160f19269

Observation 6177047f-aa54-44f6-aea6-ef3cb1efdf54 · outbound

This paper cites Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.936241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.559399Z digest=sha256:3670f7970242ec2894f1c1678e3a30be6d7038a89252dc1a6f0409c2abd9227f

Observation 5fdcae0a-3e61-4d20-a971-52375e32b597 · outbound

This paper cites Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.920767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T23:51:05.564651Z digest=sha256:7a029418cc074ac356770d400e65018ccb890891bea36d019068719579c3debf

Pith citing papers

Observation d33b1b39-c1fc-4205-be93-e74a269e4339 · inbound

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation cites this paper.

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:05:28.900178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:05:28.900178Z digest=sha256:0d0e327a5728a1c4b996e59352f87a8dadf651beaa1945f7951869b5fa50a032

Observation fff67d24-b737-4cff-8fe2-ff062d4198d9 · inbound

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation cites this paper.

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T10:43:40.991161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:43:40.991161Z digest=sha256:b054b4045d38529395a52d0226722016d242dbd54e9b20444a0e985a5dd390a5

Observation 028fdbbe-87a4-4252-9064-713e605b5d95 · inbound

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation cites this paper.

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:25:32.979738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T00:22:41.611893Z digest=sha256:ec9b6a30a3ebf96a63729b017067f0862d415fed54b2be8ba70677f42a9a2396

Observation 8fc455d9-e6c7-44c4-8e69-2d599f1370b6 · inbound

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment cites this paper.

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:52.618248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T19:27:12.286456Z digest=sha256:3aa14b4c14516781544e83f051da3aa840686595272756e7cfac3bdf11c90281

Observation b58a1dcb-a01d-4e5f-a51f-6b8a8fc1da92 · inbound

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation cites this paper.

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:19.021480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T04:50:33.134927Z digest=sha256:c4bd1128b00391766fe156e90a9bc9dd5e293c32e4f56a126f8ff6f2ca35fc0b

Observation c4febba7-8387-4ccb-bdd3-5d0ebb6ffe1f · inbound

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training cites this paper.

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:11.595737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T08:07:47.793895Z digest=sha256:d929602b65d8e7a8afce784d0623f22d914d400ac612f18dc4692bcbf15818a7

Observation 5813eb5c-32b6-470d-a900-9fd4888a5281 · inbound

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning cites this paper.

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:22:14.583427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-13T04:22:07.953637Z digest=sha256:370f7815299ddf82fd17dac59d3f2487bb77378bd591d4235d2f49ed3da7418b

Observation ed470334-f0ba-422c-922a-939f8f57033f · inbound

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation cites this paper.

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.455458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T06:57:41.245418Z digest=sha256:a14d642a5a167617ea1464f2068dd657341fb6f3d618bd1252a63f51a85efaa4

Observation 27e3a5be-f2fb-4e2c-aae8-907fc7f80c20 · inbound

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models cites this paper.

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.825517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T06:55:53.760595Z digest=sha256:b39334257421c9ca9aeb652513afcf6958c5c480bc6d00a40301181178fd16c3

Observation 19db67d5-ffa5-443a-90d6-5beb6bb7b0f7 · inbound

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining cites this paper.

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:37.882421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T11:23:36.941801Z digest=sha256:61a9edec6efa36ca2a5bc7fbf2f6e1fab5ed070eb4639f77abce1dbec3d64bbc

Observation b2fab49c-65e1-4a40-94a3-3accfa9d07fb · inbound

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation cites this paper.

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:30.517227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T18:01:08.616612Z digest=sha256:1a542912134c4aa84d16019f08711ba21c9a663e5c0da8c543ce24ffe82693af

Observation 49371b32-4dc8-43cc-a7d5-2b9dd8e113be · inbound

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation cites this paper.

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T20:08:36.570832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:08:36.570832Z digest=sha256:39ea1a0ca909780fd753af7900d4e3f3abea765a6b7352fac1e989ad0e5d35f5