Pith. sign in

Paper Citation Record · LEDGER

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 25 inbound Pith citation observations for arXiv:2508.10333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10333 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:47.949431Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:43:02.573394Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 016cf2e5-6ffc-416b-b2ae-b2216e04b144 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.157327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.157327Z digest=sha256:4e36772d591c881dd7d3110db66f21b5f022467a2073c60b4e5b456a2e8938e2

Observation 51f44ab3-2844-4713-9497-581cd0c11de9 · outbound

This paper cites write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.256938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.256938Z digest=sha256:5449c05c4e2b99158ad510355bfc2b16c8ac4da7b0e40a44f2a6f1a0a9655632

Observation aeeb2696-acee-489f-a136-bc444a61c5b6 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.335373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.335373Z digest=sha256:ca46bd1b6acdf8b4386c2476383bb7208ea3dc3464c445331b5fa75055176039

Observation 55b44d0f-e875-4922-be3e-02fe9f8fd682 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.445023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.445023Z digest=sha256:3e8eb0f6130bd73be83632234750b0c4cbbafe7ee7c5a2e2ece682f81ac604fe

Observation c00ad31d-fa76-43a3-90a4-ad61a16998c0 · outbound

This paper cites R.; Finn, C.; Kumar, A.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Finn, C.; Kumar, A.; and Levine, S

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:55.339761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.532847Z digest=sha256:a4d1e3ec5966b065f9d4be4040e4aa947154680989b6a7dad644f21a929635d6

Observation 06acae65-edc6-4903-8305-99374a8069ba · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:55.111512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.649791Z digest=sha256:67930b5e00fe9017e7ca58f0fe0d353a0ea10ca7e76b0962846f10c16930dd1c

Observation 198b9eed-5966-4a5b-925f-906ed4ef824f · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.733209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.733209Z digest=sha256:5e534e799b99f6eddf5fd161f58174c7ba752c72e1e4eeac61cd73584e60fbf0

Observation c547134b-6b44-49d3-8058-63b9bcfabe75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.891716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.870859Z digest=sha256:c71e22436e4f1f5732297c952b9d45c128caf283e2b824985496e177e55b0a0e

Observation a0da2fd6-ec44-4efb-b3ca-6e38b27f6f4b · outbound

This paper cites WorldVLA: Towards Autoregressive Action World Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver WorldVLA: Towards Autoregressive Action World Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.981102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.981102Z digest=sha256:4dd7fa19361bf8376593889a41a2c21ab8ddf5d8564b3478f4a6f5fb603097eb

Observation ab4bafc3-28af-4703-879e-7fe899692918 · outbound

This paper cites Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.065743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.065743Z digest=sha256:a56f82839ec3b83221d042d8234a103ec7ca01b6a12fa380965792ae51caf068

Observation 714ce289-afb3-4d49-997b-99ce921d69b2 · outbound

This paper cites 2016--2019.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver 2016--2019

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:54.667834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.146408Z digest=sha256:98825e3c5e445fb44aa82bcc9a5be7c01bca770340ff10f2924465d1115eefe4

Observation 69e98311-e672-4d82-a735-259fddfe5f32 · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.242879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.242879Z digest=sha256:d349d94169152e04caf1bb2b4f1bb75f18d4d169d757a88089d7856b5110b2b8

Observation 551c752c-1b81-4542-9afb-cc94e77c9086 · outbound

This paper cites GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.318557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.318557Z digest=sha256:9ab9ae882003f3d91b8ae8754ce5dcfe2bf00de272da9bda5f2bc16975c76f80

Observation 938d2317-9e60-46c1-a3f9-030549681f0e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.442791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.423867Z digest=sha256:9a5515e59e1bf8e038d2fea67b321862c3e4ef7f8e475581576b4dc79b2685f7

Observation d35117dd-cc1c-4143-9414-317818c7bc8b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.217460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.559992Z digest=sha256:4bd17e4838ab3c21344b728d1d7aff1edb3f978b746a0c668fd79f652e3eaf57

Observation 53a7fbd4-1b51-41ca-82a2-d25b7325feb7 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.003307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.659725Z digest=sha256:94b3a327f7a9c4b41357a28d14b98bd5bad18d31acbd4590df549a001ec8d1b7

Observation 1202615c-7b2a-4fb4-ae69-df1faa8315fa · outbound

This paper cites Prediction with Action: Visual Policy Learning via Joint Denoising Process.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Prediction with Action: Visual Policy Learning via Joint Denoising Process

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.779814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.779814Z digest=sha256:65f8142f8ef9ea96b051b1420708bfb371f48a84a319922006fdf920bbb94b17

Observation 5ea62d55-fde7-489e-a6e8-98527acce49e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.866349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.866349Z digest=sha256:a5cb7ae8ff8364cd28e48ac33b1cbae1c973b2b0657a5356cb31f48ecc5f92b0

Observation 1689ff9b-2280-4c5d-98aa-43d842ff1973 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Denoising Diffusion Probabilistic Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.969912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.969912Z digest=sha256:7713e76d8a478c8f44db08a240bb571aecec79689227b651ae1438cdca00d86b

Observation 6087975a-809b-457d-9aa9-8bbee619fd25 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.809157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.077675Z digest=sha256:4251f14a0ad575014a7f22ba5fbbdcc015253e3b9f8adb9610e9b311260d6fd1

Observation ab0ec46e-731b-4f0b-abf5-efcf72132ebd · outbound

This paper cites Elucidating the Design Space of Diffusion-Based Generative Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Elucidating the Design Space of Diffusion-Based Generative Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.168164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.168164Z digest=sha256:45b977b968ff09b447fefa7f2f9db5f7fae381b88c439c5aac3ce21d1dc796f2

Observation a9876cd6-eab2-4244-8bdd-b648e0c9944c · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver YOLOv11: An Overview of the Key Architectural Enhancements

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.246525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.246525Z digest=sha256:4f26b292cfb2df50ee52342ad99b3c400bcb9b2a8a50effdb14a302b61f6bce3

Observation dcc27518-8402-45ab-8d2d-e921d14404e3 · outbound

This paper cites J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:53.600727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.342040Z digest=sha256:0ce165e5fbb1b503506a1c3c984436e7edcd2ffab3d267a2614f7e6181e711df

Observation da5c8f71-2c0d-4347-a6e9-a711655e2601 · outbound

This paper cites Auto-Encoding Variational Bayes.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Auto-Encoding Variational Bayes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.447704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.447704Z digest=sha256:0accbf5522c1ebcfa3b7b70610b80525e1d20b92035c743bfa6d0b679d39f282

Observation 9fca09ee-667d-40b0-9ce7-6108f3bd8d1f · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LISA: Reasoning Segmentation via Large Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.554728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.554728Z digest=sha256:58cdab50bc2fcf8538e3a8a789ea800fd126a7f507189a7cea317440c628f4f6

Observation b70c07cb-7301-41f9-bb17-c5cbbd37e5a1 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.342812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.669023Z digest=sha256:0a7d5525d7804e66522ababf49315950f5c5859905985aa06a1fd7483626cc2c

Observation c6d3a858-2cef-4a18-868a-14a5e7a7fa19 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.096576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.765284Z digest=sha256:e4ee7d9a58b9797a343abb77105f5cdf0af9020c30c3e49e0e169f35d8b9e880

Observation c4ffcb0e-a213-4910-8ad6-d0482e7ffc4f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.824198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.866675Z digest=sha256:6500e67c3408b8db8e489ddd3fd9d6e755e22fe47bbca33c93043d829c3b1571

Observation 07a1d701-131b-4fb7-81de-748660bc9c78 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.552907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.956249Z digest=sha256:4caf97027cc8b89f954190049af277d8ed09e07782654bbbe74b9522f2100e68

Observation ae041e0c-9b5f-4608-8038-a44070d822f3 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.262406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.066281Z digest=sha256:23c420fdb544c004af3180dd9a0c380530cac24dea536b1582e7688bb80a7927

Observation cc708740-53df-4841-9066-d53e6a38ec97 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.212485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.212485Z digest=sha256:ad6f30093139816b27216260964378f6e6b999b5c4b5fe6deda03d2b4adb8938

Observation 19505c9c-a177-4790-8423-3693c94b90c0 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.940390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.297037Z digest=sha256:118bfcfba77661687dd0c176e05406ebfed33b08c210fbc15a79185e6cd024e0

Observation dc6d773c-9e4a-4678-94a5-6d73163ec34e · outbound

This paper cites LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.390488Z digest=sha256:7ba5622b447edc5fddca4c00a8226892674a0221cceb92872462827eda4f08a9

Observation 89925814-2c7a-48a3-96e5-87396f0cc01b · outbound

This paper cites Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:51.617250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.482182Z digest=sha256:ff39cabb4b5ded0a05a0686222249f7a4a85e62c3284ef13b71180991c426dce

Observation df7fb02b-7e2b-4311-a8b1-676ab029924f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.298656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.579047Z digest=sha256:0b1916e6271df8d5f9349493fa945794656eee2d71c9a1285d0b81644a94be04

Observation 8276046b-c484-4278-8264-49318ee9d348 · outbound

This paper cites Scalable Diffusion Models with Transformers.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Scalable Diffusion Models with Transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.709820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.709820Z digest=sha256:3e94aea6af53e0d2a6e2662e425f54b38a711a8a945f8c66774a247a9e226351

Observation 1cdedee4-544a-4af8-a9b4-0b3a8e676f7c · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver High-Resolution Image Synthesis with Latent Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.824230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.824230Z digest=sha256:0900da947132692f6de443bc85cf6eee13ba51147bfa2bd7896745c007452fd7

Observation be481c99-2867-4366-a198-bc684b3d9df4 · outbound

This paper cites CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.934582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.934582Z digest=sha256:366f92af4f9b5e0358eaf31d5e86bf07ec2c74fe3081eec50ed2eafaca8ce3ef

Observation dc72166d-d190-4471-abc2-81fee668cfae · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.054694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.054694Z digest=sha256:c433d52113861c18e43854bffc93bf0b56695deaa1a1f9564645da456dad39aa

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · outbound

This paper cites RationalVLA: A Rational Vision-Language-Action Model with Dual System.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:7606201e1555677c0d5f620f7a222aa1641686a63b30e519e164535be34375b7

Observation 962310b2-29a4-435c-8c94-a14bacd57d74 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.007699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.259758Z digest=sha256:24edc744133e8b2d5a4b8cddccad8e9b5e9003fb749958b4102c2c5d613483e4

Observation c8a4849e-a5f7-4daf-ab83-e3007ad3de02 · outbound

This paper cites Generative Modeling by Estimating Gradients of the Data Distribution.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Generative Modeling by Estimating Gradients of the Data Distribution

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.370745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.370745Z digest=sha256:89fb80ee6a4bd0a9b40c1fd2aca187d053d46036addaad61c2ade1865d7f152b

Observation 555eac16-8dc4-49c0-8c20-45642aa8b9e4 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:50.728268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.479452Z digest=sha256:c6fdc7fbafce151bc4a4e61a51b4b0fe71df736c161fc0c673cb98c83d6d453e

Observation d881a641-d32d-4186-baf1-a0d8bba7cae6 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.559044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.559044Z digest=sha256:a7d086b862953e693b2764c6e1a9df65d5c8591e793b0e03c32e82e5b139f49d

Observation 04fd7e57-eaf3-4ceb-b882-4eea9e122eba · outbound

This paper cites QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:32:48.264716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.642262Z digest=sha256:be24eefdcc9e6e0f391f3812e6ae4b06a066b5d7d029eedb96a0488cfe19c59c

Observation 78b3481b-d7df-4f6c-b1ad-bf84737163a5 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.725496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.725496Z digest=sha256:6c6657b1d003434ed61e0bd0221c7fa6a8d47df74894832c55f48e99f1eafdd5

Observation c1cd9d61-8884-436f-b469-261b8210cb6e · outbound

This paper cites BridgeData V2: A Dataset for Robot Learning at Scale.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver BridgeData V2: A Dataset for Robot Learning at Scale

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.783305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.783305Z digest=sha256:4b0b9db57a2fdb65639a34e04da2f1bfcc0064626eccb7b4fd13a8db6723a852

Observation 5e261572-3b55-4c88-b4be-eb504e83cbf3 · outbound

This paper cites R.; Black, K.; Zhao, T.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Black, K.; Zhao, T

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:50.320404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.873909Z digest=sha256:6db6decc77dc565570bd94d341d6fca13f083a4052b5b1c7cf9eccb2cb253475

Observation 961c6b19-704e-4cf5-8b3d-43130886e1a4 · outbound

This paper cites Reconstructive Visual Instruction Tuning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Reconstructive Visual Instruction Tuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.962180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.962180Z digest=sha256:9cf4a6becdf0ee24f279053ef8393842e41b3fabfc6ad7a03c7f7e805cf0a7cb

Observation 35480d38-3d8d-435c-a75a-6db8a4a4b601 · outbound

This paper cites Unified Vision-Language-Action Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unified Vision-Language-Action Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.040368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.040368Z digest=sha256:963889d6c26483abd5815a51f674c87e7aeb3f70ab4dad9435fc7ff1beb4897d

Observation ab9de928-841d-4d9c-b2d6-9bac4f94c513 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.945212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.125046Z digest=sha256:ce71559be1e44459699019e3d1970c013483b41e73b3b2c308b1c9b54673c503

Observation f7912514-9f6e-441a-a431-939058de3eed · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.736982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.234493Z digest=sha256:47c194c4d78f3c4fc83d69076e38de13f2b7572e8eb0bc3c90b2dcd2f4792a24

Observation e3675e46-39e6-40c1-8a83-8b3963868738 · outbound

This paper cites Qwen2 Technical Report.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Qwen2 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.323607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.323607Z digest=sha256:14a994b15336b5db40f8ac99037eb4f560be75ba5392039f995bfe047c36951a

Observation 9b121ff5-6903-468a-b8b1-5072aa742e75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.535472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.410676Z digest=sha256:ab28bc7d67406d096ec1f93b40b5051c8945f6d765281d20ccc4e90bb73a75cc

Observation c72d3213-3f3d-4590-99c1-76b9ef2cd74e · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.498001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.498001Z digest=sha256:45a73e8cd50707b63cb05a30d28dd60ccdb8177ed01c480c3ea7bf14080a5ec9

Observation e5c2536f-6e29-4fd2-9a13-780548522aeb · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Sigmoid Loss for Language Image Pre-Training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.563746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.563746Z digest=sha256:e65ac0a2ec5a1f451ec649c5ff8f2e4a36effae22c72401cbe117f6218180c02

Observation 4e7d3ff6-6072-4b93-ba59-62fe787ce832 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.340025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.644076Z digest=sha256:ddbfbd55a096f5ee3ae2703a32c0c5a1a97f25423666b16d665820c92186019d

Observation 2878e17c-26e5-4437-902a-464de73e7722 · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.702884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.702884Z digest=sha256:7abb3f7cbda7d9fee02bcaaf3fcbd94281853fa11e836fe50cf2248a5eea1f25

Observation 30b05cf0-1631-4f7f-802e-65e17540558d · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.158036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.766165Z digest=sha256:85a93badff290460eb30054713370d518b26a3e2ff357bb9702211f0b83be4c4

Observation c5a20b2a-d22f-4280-bcaa-1753625b240b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.985244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.881596Z digest=sha256:40032da32df4b61c3d4337dd9769882a1c6e23058d4d6da27998d3c84700ca82

Observation c1bca469-ffdb-400d-bdd3-e767492d0800 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.824489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.949431Z digest=sha256:1ce1e67213599dc961b47b2235b7cf75c86e8d0c8dec9cd08ef22685106b8d56

Pith citing papers

Observation 19af9271-5ab4-4013-9b05-dd7b6cf08d83 · inbound

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models cites this paper.

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T09:34:44.154129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:34:44.154129Z digest=sha256:04a6f43abecaff39037d84b50516b6aa1a89087e47817a544c3de1d0bca4a154

Observation 0fe86c0c-8d69-4c55-ab72-8922135f7214 · inbound

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention cites this paper.

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:29:09.956059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T06:28:22.652509Z digest=sha256:3c275aed21c46b636eae741bbc5c473b8b14e62477c6aeb7a1054ba04562f352

Observation fa147f79-109b-48ed-8d65-e62d9182f7f1 · inbound

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation cites this paper.

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:24:11.084402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T13:22:16.242427Z digest=sha256:dcbeea5643ac7fb03ba062157ff9c0ff6bacda2d7cea8a7cf78c469927dee98a

Observation 687336c7-bd92-4c6d-9459-ecc1d17ac7e6 · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.270525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T09:49:26.446868Z digest=sha256:905415281906640636ab9519bd2190109543a7ab6056a9ed4f917af6b9481e45

Observation ac8c78d4-7a12-45ed-a724-9bff414a72fd · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T00:13:51.580603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T00:13:51.580603Z digest=sha256:49676939a0b81254f2bbd6fe83f4219088ae84ea2cc70be3714a9668299da625

Observation 61734d00-9f8d-434e-97c7-60a46713027d · inbound

Grounded World Model for Semantically Generalizable Planning cites this paper.

Grounded World Model for Semantically Generalizable Planning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:05:32.233313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:05:29.465402Z digest=sha256:474238becc32b9b91d8a39217211c217647ee995c48c12c60f14c574082ca78e

Observation 37ad6c86-c8ef-4dee-9050-fe0b8a291e96 · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:05:29.539890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T14:01:50.429917Z digest=sha256:e02580d8d4e70293ab3ef3f3166e00e697aaf423b6fb2fae46eeba684c0e0372

Observation a89019a4-ca2b-49da-b1b5-e39e3a5f0e8f · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:52.410964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T01:51:46.849464Z digest=sha256:7da6f2504cefb9910bf04c76ac1b7ae9959feee91d7fb9e3927e55b66cb44dbb

Observation 8b519540-cf1a-4cb5-bf3d-dd72203cd8ee · inbound

Mask World Model: Predicting What Matters for Robust Robot Policy Learning cites this paper.

Mask World Model: Predicting What Matters for Robust Robot Policy Learning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:06.115031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T02:14:17.676675Z digest=sha256:78e45d605ae2e0aaa85b27ce5bb5026bef9f872c82fbf2ebc5f2f09392016c3f

Observation 8596f8aa-4994-4f4a-bdb5-6a505effefb7 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:25.118343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T22:07:24.208555Z digest=sha256:24cd1203d409897fb9462e2ac59e6692b46d3150d1f830359049f5ae953f5f51

Observation 0b5e01cd-11ba-4333-a186-a3d1e3e36d45 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T18:39:49.173828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:39:49.173828Z digest=sha256:3e1c4d217564cbec918a022ed35e5e061e9b3edde570c011d94920a8343c1849

Observation b834c9a3-a4f5-42f6-9030-286defb1a253 · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:16.014701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T03:13:36.437080Z digest=sha256:0338ab24d9730fd0b1888c8de9933dd779c1a78861bf9de980a2947d95786f8b

Observation 120946ea-c5c6-4db0-bf1b-2c4431ab940b · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T18:17:43.564400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:17:43.564400Z digest=sha256:f159cf2af805b818245c5ca541ecbdcdea63d1c2077285875073ac7bc92eac12

Observation ebe3b632-0602-4730-bbed-32727529d609 · inbound

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models cites this paper.

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:26.986030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:07:52.566360Z digest=sha256:07ae4f9e403391ca821fe043e5d1a9c189be980955bc1a1e9c906bbba7487f6f

Observation 633e7192-f355-4783-9050-9e01dab00802 · inbound

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete cites this paper.

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:57:17.450518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T04:55:23.145492Z digest=sha256:fdf2e4dd042396834ec2eb6387f3b3e96d378ef11b0c52f2d226838d60996b26

Observation 3f82cfcf-3285-4442-82fa-fb2cff8805a0 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:06:08.676554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T06:05:52.461348Z digest=sha256:9cd765deae9765aa4de3952861c234fa559366946fa1136d88dad21de6becd15

Observation b45b1ea2-afc0-4140-a249-687f7ee2fdc2 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.253115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:8b2855c095d014e4eed28a748e4384fa0224eb55a4e43d5510884a5a77a581c3

Observation 9fb6d44f-2a71-4c89-bed3-bc563272dbff · inbound

GEM: Generative Supervision Helps Embodied Intelligence cites this paper.

GEM: Generative Supervision Helps Embodied Intelligence ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.942277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T13:38:27.263726Z digest=sha256:935c3d275b47863509c0b1ee6fe8756527c4c4bc9a2a18ba09b65ffce435b2c3

Observation 97f062d7-ed85-48d7-9bdd-387dd4e394a0 · inbound

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving cites this paper.

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:02:45.777855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:57:01.742536Z digest=sha256:2943387a017449bf10fe1e31973883240373a41157f6b591a6877b07a9d946fc

Observation c7e9f867-668a-4f8d-abba-10e39c14a3f0 · inbound

OneVLA: A Unified Framework for Embodied Tasks cites this paper.

OneVLA: A Unified Framework for Embodied Tasks ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.110096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T17:05:08.124096Z digest=sha256:a9a172674ce075719919842505b9ca36582cf18c9725ba8bd678a6717694067e

Observation def8b06d-2585-490c-9a1c-43e1c96c8825 · inbound

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding cites this paper.

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:16:59.252395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T01:23:02.576098Z digest=sha256:0d3bd58eecf86e5eb291fbebc034ad49f9473cb4af5172b982aef21be4622e3e

Observation 679555cc-f140-41bf-872f-17c2adabb100 · inbound

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging cites this paper.

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T15:19:27.489381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:19:27.489381Z digest=sha256:2a2333e5cbf457aa415eb7b161e581170e764b9af32dea86f0adc0fee87ccaa7

Observation 432ac8cd-3b5a-4d82-bd7f-38db2c5679f7 · inbound

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment cites this paper.

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T05:16:41.306473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:16:41.306473Z digest=sha256:4153fdadb33257609ceb9f9e5b461c2eda9fb6e51d425b66ed762971a6418da6

Observation 7c3528a9-9a6d-477e-9437-b8c2a6af53d9 · inbound

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation cites this paper.

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:30:08.940833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:30:08.940833Z digest=sha256:aa1919b3f93ffae19b92cb3625f9a14f73b0be3b38fce955216630e3ec307f41

Observation a229022d-8beb-4610-9978-6d1de4977e9e · inbound

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight cites this paper.

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:43:02.573394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:43:02.573394Z digest=sha256:0c77c35c9f9ac08c51ebc183d1c611244b3ac49a0ec98846176bc334d518d6a5