Pith. sign in

Paper Citation Record · LEDGER

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

As of 20 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 28 inbound Pith citation observations for arXiv:2508.10333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10333 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:47.949431Z

measured 89 of 89 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:27:01.402929Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 016cf2e5-6ffc-416b-b2ae-b2216e04b144 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.157327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.157327Z digest=sha256:4e36772d591c881dd7d3110db66f21b5f022467a2073c60b4e5b456a2e8938e2

Observation 51f44ab3-2844-4713-9497-581cd0c11de9 · outbound

This paper cites write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.256938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.256938Z digest=sha256:5449c05c4e2b99158ad510355bfc2b16c8ac4da7b0e40a44f2a6f1a0a9655632

Observation aeeb2696-acee-489f-a136-bc444a61c5b6 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.335373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.335373Z digest=sha256:ca46bd1b6acdf8b4386c2476383bb7208ea3dc3464c445331b5fa75055176039

Observation 55b44d0f-e875-4922-be3e-02fe9f8fd682 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.445023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.445023Z digest=sha256:6eec33a0e0eab45c7351152a3262c06c375cd7ca92829e1f22c3b142555cdbfe

Observation c00ad31d-fa76-43a3-90a4-ad61a16998c0 · outbound

This paper cites R.; Finn, C.; Kumar, A.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Finn, C.; Kumar, A.; and Levine, S

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:55.339761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.532847Z digest=sha256:a6211f07d8a60a01fdf096fbdaf56444bda3129022a8881a890630fe0999d828

Observation 06acae65-edc6-4903-8305-99374a8069ba · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:55.111512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.649791Z digest=sha256:dbcb0a0444f4a754d12ea62aacea0dfccc81f08b11f27d3ecdce2be829a6a38f

Observation 198b9eed-5966-4a5b-925f-906ed4ef824f · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.733209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.733209Z digest=sha256:c7d0229a8df643d78b3ecf1aa8820d876e6466545d9cd45a1e1236f9905213e1

Observation c547134b-6b44-49d3-8058-63b9bcfabe75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.891716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.870859Z digest=sha256:94aba5f9f0c3ed46865d9147243e229beda06558f6d154ffa9e89f36c5cfc93a

Observation a0da2fd6-ec44-4efb-b3ca-6e38b27f6f4b · outbound

This paper cites WorldVLA: Towards Autoregressive Action World Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver WorldVLA: Towards Autoregressive Action World Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.981102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.981102Z digest=sha256:3e218bec21e8d118d49bd14e7c19b24cac89c9c64efb268e8b32e919461af008

Observation ab4bafc3-28af-4703-879e-7fe899692918 · outbound

This paper cites Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.065743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.065743Z digest=sha256:970529cf4d2329ef267a444a1b895520c7a49e86e9ec61be5172926bc37a412c

Observation 714ce289-afb3-4d49-997b-99ce921d69b2 · outbound

This paper cites 2016--2019.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver 2016--2019

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:54.667834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.146408Z digest=sha256:7ab24fab0bb1ea7c8634bbc139490f8ddf9bbf47c590d2ddce69123867bd2322

Observation 69e98311-e672-4d82-a735-259fddfe5f32 · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.242879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.242879Z digest=sha256:952b4001c86abf35bf069788e4a83c77979ba0129e32f77b9d63e3365459d50b

Observation 551c752c-1b81-4542-9afb-cc94e77c9086 · outbound

This paper cites GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.318557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.318557Z digest=sha256:0bfb4b3915fdf37c432bf86bef6f7bb5d1b2e877dca029704839ce233ed96095

Observation 938d2317-9e60-46c1-a3f9-030549681f0e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.442791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.423867Z digest=sha256:651f1d66d9993d04f522baf60e4e39a717639c302165adcff8609dcfd6e53a6e

Observation d35117dd-cc1c-4143-9414-317818c7bc8b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.217460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.559992Z digest=sha256:49f4505138ba112bbdd3e50272fd48570fe2af22a35fde928b2b60cc71a00b9c

Observation 53a7fbd4-1b51-41ca-82a2-d25b7325feb7 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.003307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.659725Z digest=sha256:eeb20d309d010019402a2d6b6f44167000b543b3922865b5ddab596ad04fde6d

Observation 1202615c-7b2a-4fb4-ae69-df1faa8315fa · outbound

This paper cites Prediction with Action: Visual Policy Learning via Joint Denoising Process.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Prediction with Action: Visual Policy Learning via Joint Denoising Process

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.779814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.779814Z digest=sha256:225051b373d95201c4b30584f082d2117fca4335ca873a3ba0033b68b9a97298

Observation 5ea62d55-fde7-489e-a6e8-98527acce49e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.866349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.866349Z digest=sha256:a5cb7ae8ff8364cd28e48ac33b1cbae1c973b2b0657a5356cb31f48ecc5f92b0

Observation 1689ff9b-2280-4c5d-98aa-43d842ff1973 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Denoising Diffusion Probabilistic Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.969912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.969912Z digest=sha256:36a6dee59299b3f183d200081083ff4662a7f3239fe8e84c23c52b6b1b8847b2

Observation 6087975a-809b-457d-9aa9-8bbee619fd25 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.809157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.077675Z digest=sha256:e0d38a2c311c160061c317e47c8aad8df9ea3aed1e45eec337f33a80deeaab44

Observation ab0ec46e-731b-4f0b-abf5-efcf72132ebd · outbound

This paper cites Elucidating the Design Space of Diffusion-Based Generative Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Elucidating the Design Space of Diffusion-Based Generative Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.168164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.168164Z digest=sha256:45b977b968ff09b447fefa7f2f9db5f7fae381b88c439c5aac3ce21d1dc796f2

Observation a9876cd6-eab2-4244-8bdd-b648e0c9944c · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver YOLOv11: An Overview of the Key Architectural Enhancements

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.246525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.246525Z digest=sha256:d8eeeaaf6f662b1c86716abb569851296d0d0bcff0bc884c0e9115b182d9a817

Observation dcc27518-8402-45ab-8d2d-e921d14404e3 · outbound

This paper cites J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:53.600727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.342040Z digest=sha256:9b2b26753f8f7c271aea96ae9ab48b9feef65960ceda87dc2aa6a42522dbcc17

Observation da5c8f71-2c0d-4347-a6e9-a711655e2601 · outbound

This paper cites Auto-Encoding Variational Bayes.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Auto-Encoding Variational Bayes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.447704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.447704Z digest=sha256:2a0f54700801baf5d4df563f2f13acfb5a333d6ed538d9b10d066e7a5dc6b44c

Observation 9fca09ee-667d-40b0-9ce7-6108f3bd8d1f · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LISA: Reasoning Segmentation via Large Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.554728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.554728Z digest=sha256:32b720abe454b459f86c2b070582a243aceb887a2c78ce51663a4354e14cff45

Observation b70c07cb-7301-41f9-bb17-c5cbbd37e5a1 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.342812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.669023Z digest=sha256:1abebbcad85c8f4bc8f95bade7dd2f270b2a57ceff8ed86c50730b84e547e5e3

Observation c6d3a858-2cef-4a18-868a-14a5e7a7fa19 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.096576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.765284Z digest=sha256:a28c6399a1b294e6d83518ce0ffb1e311b94fc5e948013e3f236ab1f1b401617

Observation c4ffcb0e-a213-4910-8ad6-d0482e7ffc4f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.824198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.866675Z digest=sha256:c6e05e4c666975f74c20474f771f9fffa98b3d5c591f025cae3c74a652b566ae

Observation 07a1d701-131b-4fb7-81de-748660bc9c78 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.552907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.956249Z digest=sha256:23428ac2845bb893682d6c26ae0f571182fae66a71c94b6279b887fe1c0198ee

Observation ae041e0c-9b5f-4608-8038-a44070d822f3 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.262406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.066281Z digest=sha256:73609583af2e1849d7e90706b1073148bd892eef0d48f2977edac9f3894a5a4e

Observation cc708740-53df-4841-9066-d53e6a38ec97 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.212485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.212485Z digest=sha256:85bfc173ae47443f76b19f978d0e79d0b6b4991909d6d2b3c9f8a062f62ecc4e

Observation 19505c9c-a177-4790-8423-3693c94b90c0 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.940390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.297037Z digest=sha256:946464d9f4c6afca0ade868f83a7e5694dc3a923081a0584803e1592deaa4a24

Observation dc6d773c-9e4a-4678-94a5-6d73163ec34e · outbound

This paper cites LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.390488Z digest=sha256:012853c76bfff95161847ca20d03b047ebfc2b14f5ba6236a348f65be1d70048

Observation 89925814-2c7a-48a3-96e5-87396f0cc01b · outbound

This paper cites Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:51.617250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.482182Z digest=sha256:03a55625cabb879408eaf3fb3294caf00dc2f044c74a9df254a2874646ce946c

Observation df7fb02b-7e2b-4311-a8b1-676ab029924f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.298656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.579047Z digest=sha256:3967260e9289964b3477cb9dc4daaa288c9d300086d344db0348b2c8fc148ecd

Observation 8276046b-c484-4278-8264-49318ee9d348 · outbound

This paper cites Scalable Diffusion Models with Transformers.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Scalable Diffusion Models with Transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.709820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.709820Z digest=sha256:b14265d8fafa139b728524b7f3c7413fc23e019a2a356caa3ddf75e6565fb601

Observation 1cdedee4-544a-4af8-a9b4-0b3a8e676f7c · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver High-Resolution Image Synthesis with Latent Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.824230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.824230Z digest=sha256:0900da947132692f6de443bc85cf6eee13ba51147bfa2bd7896745c007452fd7

Observation be481c99-2867-4366-a198-bc684b3d9df4 · outbound

This paper cites CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.934582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.934582Z digest=sha256:73fa71070c66c44e473c9906977f8cd232b6e3f3e715ab6e79858a9f86a6460d

Observation dc72166d-d190-4471-abc2-81fee668cfae · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.054694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.054694Z digest=sha256:c433d52113861c18e43854bffc93bf0b56695deaa1a1f9564645da456dad39aa

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · outbound

This paper cites RationalVLA: A Rational Vision-Language-Action Model with Dual System.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:8c5dffef2a020337bdaee5010db5e8d8bb9775e7f8a63753b5995a0cced1351e

Observation 962310b2-29a4-435c-8c94-a14bacd57d74 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.007699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.259758Z digest=sha256:7561260917351e3988b998af57f42a82fcac47c7f9f5ba378ecd7e3bcaab9e60

Observation c8a4849e-a5f7-4daf-ab83-e3007ad3de02 · outbound

This paper cites Generative Modeling by Estimating Gradients of the Data Distribution.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Generative Modeling by Estimating Gradients of the Data Distribution

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.370745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.370745Z digest=sha256:89fb80ee6a4bd0a9b40c1fd2aca187d053d46036addaad61c2ade1865d7f152b

Observation 555eac16-8dc4-49c0-8c20-45642aa8b9e4 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:50.728268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.479452Z digest=sha256:3517eae6ddd8113aa43bf5e0bc64b593735b5d89561ab6b7608790b31ba5d0aa

Observation d881a641-d32d-4186-baf1-a0d8bba7cae6 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.559044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.559044Z digest=sha256:b1b314230bd63f1ae85af3cc40ab9ee17535946848b06508c9a907b611da839f

Observation 04fd7e57-eaf3-4ceb-b882-4eea9e122eba · outbound

This paper cites QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:32:48.264716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.642262Z digest=sha256:dcf26ac40980156690889acfe10e71eafab4a861911460b852f13c065a3b4d84

Observation 78b3481b-d7df-4f6c-b1ad-bf84737163a5 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.725496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.725496Z digest=sha256:6c6657b1d003434ed61e0bd0221c7fa6a8d47df74894832c55f48e99f1eafdd5

Observation c1cd9d61-8884-436f-b469-261b8210cb6e · outbound

This paper cites BridgeData V2: A Dataset for Robot Learning at Scale.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver BridgeData V2: A Dataset for Robot Learning at Scale

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.783305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.783305Z digest=sha256:a68a21590a9f77c84a44195666455b8351f0daae2ef2c19ea016b759c96e0cc3

Observation 5e261572-3b55-4c88-b4be-eb504e83cbf3 · outbound

This paper cites R.; Black, K.; Zhao, T.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Black, K.; Zhao, T

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:50.320404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.873909Z digest=sha256:ee13e16e4df3afa1868c391dc6e5b997611add86264acc718185f801cce05fc0

Observation 961c6b19-704e-4cf5-8b3d-43130886e1a4 · outbound

This paper cites Reconstructive Visual Instruction Tuning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Reconstructive Visual Instruction Tuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.962180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.962180Z digest=sha256:326d268a8d5364001096cc7fb2828ce6918fe853ee60064aa4dd02211cc4c3dd

Observation 35480d38-3d8d-435c-a75a-6db8a4a4b601 · outbound

This paper cites Unified Vision-Language-Action Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unified Vision-Language-Action Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.040368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.040368Z digest=sha256:cffedbab83024112281e36c250c2f0c79d6f6a5bb487ded0c1a744f0a96b43d0

Observation ab9de928-841d-4d9c-b2d6-9bac4f94c513 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.945212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.125046Z digest=sha256:17a25056c32b4d42ff336f9e16ef4e92decd4ab199fc819dbe1c6b32f0eaceaf

Observation f7912514-9f6e-441a-a431-939058de3eed · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.736982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.234493Z digest=sha256:4e202afa55048bd509460d726f45557f54f22907c810017cc823d0a7bedbd818

Observation e3675e46-39e6-40c1-8a83-8b3963868738 · outbound

This paper cites Qwen2 Technical Report.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Qwen2 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.323607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.323607Z digest=sha256:a28e941921887d1199e6583bbad4aa3a02f93fce16656fd1be16fd561b79b904

Observation 9b121ff5-6903-468a-b8b1-5072aa742e75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.535472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.410676Z digest=sha256:a4ed885003381350bc08f7b531f009678568074549474667bc7e37664c5bf44c

Observation c72d3213-3f3d-4590-99c1-76b9ef2cd74e · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.498001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.498001Z digest=sha256:45a73e8cd50707b63cb05a30d28dd60ccdb8177ed01c480c3ea7bf14080a5ec9

Observation e5c2536f-6e29-4fd2-9a13-780548522aeb · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Sigmoid Loss for Language Image Pre-Training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.563746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.563746Z digest=sha256:e65ac0a2ec5a1f451ec649c5ff8f2e4a36effae22c72401cbe117f6218180c02

Observation 4e7d3ff6-6072-4b93-ba59-62fe787ce832 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.340025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.644076Z digest=sha256:684d9a90edfdea217379478845c93158cae008225d859d62edcde8c2df871ee6

Observation 2878e17c-26e5-4437-902a-464de73e7722 · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.702884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.702884Z digest=sha256:647ef6ce80c52008b7aed756950ddc9e23bc3a9f7dd22c36ea97d95983687869

Observation 30b05cf0-1631-4f7f-802e-65e17540558d · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.158036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.766165Z digest=sha256:1e6eecd9a601d27781f95534c425b6bb41eb9c9d7321552e17016bf907fce1a9

Observation c5a20b2a-d22f-4280-bcaa-1753625b240b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.985244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.881596Z digest=sha256:a4aaf6c1f6e49008fe7e4576d859e8194fa55f6e14d0fd0ce13e760f83bcb738

Observation c1bca469-ffdb-400d-bdd3-e767492d0800 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.824489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.949431Z digest=sha256:d29d208178aadc509216f3a75a0a591037438ab983c13c34ff93774b64699b29

Pith citing papers

Observation 19af9271-5ab4-4013-9b05-dd7b6cf08d83 · inbound

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models cites this paper.

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T09:34:44.154129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:34:44.154129Z digest=sha256:f16da175896d411c9bda13399d36f87376180c807cd8537c0f0f00e46af6f3b4

Observation 0fe86c0c-8d69-4c55-ab72-8922135f7214 · inbound

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention cites this paper.

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:29:09.956059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T06:28:22.652509Z digest=sha256:99a444db8d9ab6665eb19c484ffbbb0ed35f4e51dfbc5fb6550e47b6996d5f75

Observation fa147f79-109b-48ed-8d65-e62d9182f7f1 · inbound

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation cites this paper.

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:24:11.084402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T13:22:16.242427Z digest=sha256:cc5853e7482c651a5a00f6580920d391e25cbf8071406c408f27b5c03991c4f3

Observation 687336c7-bd92-4c6d-9459-ecc1d17ac7e6 · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.270525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T09:49:26.446868Z digest=sha256:4e2414a0930b075bd82b42c22c56ae61b06bf5328d134f4e9b5fc8580a97fb65

Observation ac8c78d4-7a12-45ed-a724-9bff414a72fd · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T00:13:51.580603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T00:13:51.580603Z digest=sha256:96f7227cdaa1506d637fad8eb2b14ec4b065a26fc648b32b0776dad4c090366d

Observation 61734d00-9f8d-434e-97c7-60a46713027d · inbound

Grounded World Model for Semantically Generalizable Planning cites this paper.

Grounded World Model for Semantically Generalizable Planning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:05:32.233313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T15:05:29.465402Z digest=sha256:4f95edaa8410f29eba100330e31d8dc8a1542252c18bfd4d84518979925f5800

Observation 37ad6c86-c8ef-4dee-9050-fe0b8a291e96 · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:05:29.539890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T14:01:50.429917Z digest=sha256:a201d258dd27e53ccabb47af76a21d33519dfc9332513b1264bb3118b49521fc

Observation a89019a4-ca2b-49da-b1b5-e39e3a5f0e8f · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:52.410964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T01:51:46.849464Z digest=sha256:52c6269350b881b963b8d2c3f3021cbb23e512a01db9dbf780fd70b129d39ef6

Observation 8b519540-cf1a-4cb5-bf3d-dd72203cd8ee · inbound

Mask World Model: Predicting What Matters for Robust Robot Policy Learning cites this paper.

Mask World Model: Predicting What Matters for Robust Robot Policy Learning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:06.115031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T02:14:17.676675Z digest=sha256:e6c7f3aac05d291479a23dd8cdd6657581b8eaa3300171508975a5156f19c215

Observation 8596f8aa-4994-4f4a-bdb5-6a505effefb7 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:25.118343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-09T22:07:24.208555Z digest=sha256:cc06fbca421dd22f3818d3a4c2d0de00870dc02ff3e66f1f4be8d185fdc8aaf3

Observation 0b5e01cd-11ba-4333-a186-a3d1e3e36d45 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T18:39:49.173828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:39:49.173828Z digest=sha256:3e1c4d217564cbec918a022ed35e5e061e9b3edde570c011d94920a8343c1849

Observation b834c9a3-a4f5-42f6-9030-286defb1a253 · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:16.014701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T03:13:36.437080Z digest=sha256:e1e81c4a333ed25df1e83b967e5d4025a2de0ff6303b6005c4a063f86bd21e30

Observation 120946ea-c5c6-4db0-bf1b-2c4431ab940b · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T18:17:43.564400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:17:43.564400Z digest=sha256:068e5e76d85b18d849981b1ea2c426eb260252416ad05481f9de1d6184b6451d

Observation ebe3b632-0602-4730-bbed-32727529d609 · inbound

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models cites this paper.

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:26.986030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:07:52.566360Z digest=sha256:adf6ffaf099d7181ae70d97f94e50916f85176c3266dbff5f89c0519f5b17b77

Observation 633e7192-f355-4783-9050-9e01dab00802 · inbound

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete cites this paper.

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:57:17.450518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T04:55:23.145492Z digest=sha256:79db77b2692a59702eeebca5360f5b8a037a6c4d49a95fa1f6c57a66f5255029

Observation 3f82cfcf-3285-4442-82fa-fb2cff8805a0 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:06:08.676554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T06:05:52.461348Z digest=sha256:a4ec378307023cac94fdff0ad85e4159749fd95958ea30d5348a2535aaf45893

Observation b45b1ea2-afc0-4140-a249-687f7ee2fdc2 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.253115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:faca94d2fe3b3841c1418267259abab0fd4211aa168487e6294ef28f161b0577

Observation 9fb6d44f-2a71-4c89-bed3-bc563272dbff · inbound

GEM: Generative Supervision Helps Embodied Intelligence cites this paper.

GEM: Generative Supervision Helps Embodied Intelligence ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.942277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T13:38:27.263726Z digest=sha256:47e54cb9f1012b964320cc8c1fa033788db20120bb7eb36d085a2ea357ff4d40

Observation 97f062d7-ed85-48d7-9bdd-387dd4e394a0 · inbound

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving cites this paper.

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:02:45.777855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T22:57:01.742536Z digest=sha256:d07f3e621c3b910b0986b29fdb967d700444a507b1db77a266ef4658fd253932

Observation c7e9f867-668a-4f8d-abba-10e39c14a3f0 · inbound

OneVLA: A Unified Framework for Embodied Tasks cites this paper.

OneVLA: A Unified Framework for Embodied Tasks ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.110096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T17:05:08.124096Z digest=sha256:cb3f26949723b9ebf18f197cc4e38d57b145de5dabfbcc11832cc610924b4b45

Observation def8b06d-2585-490c-9a1c-43e1c96c8825 · inbound

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding cites this paper.

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:16:59.252395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T01:23:02.576098Z digest=sha256:a1cb213210fe28e97fc9affc716d768071c77d7823094aff88ca3254bbdd3d92

Observation 679555cc-f140-41bf-872f-17c2adabb100 · inbound

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging cites this paper.

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T15:19:27.489381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:19:27.489381Z digest=sha256:0c38f8189c53142b6167e272ffc71f7d7e7e9f862cee83db2cf695cbef216438

Observation 432ac8cd-3b5a-4d82-bd7f-38db2c5679f7 · inbound

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment cites this paper.

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T05:16:41.306473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:16:41.306473Z digest=sha256:e31d0aa583b3e2c19e8704b776fa345f7b4a48520c809ee5d9ab9373a4d63f07

Observation 7c3528a9-9a6d-477e-9437-b8c2a6af53d9 · inbound

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation cites this paper.

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:30:08.940833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:30:08.940833Z digest=sha256:ec3c66e6433377ac7ccd77b6afd9cee44bde8f577c844c6fd5c5832fd5bd04dc

Observation a229022d-8beb-4610-9978-6d1de4977e9e · inbound

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight cites this paper.

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:43:02.573394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:43:02.573394Z digest=sha256:eb073428515c73fe6b53515663afc3d496c295d7739de4d7ad52d214960b0435

Observation f2086ba7-6ef6-43ed-bffe-b0771fb280dd · inbound

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight cites this paper.

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T18:04:05.199011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:04:05.199011Z digest=sha256:ed52dac0f47aa509a2d3106f937cfec2ab8dac8821beddd87c37e0029864950e

Observation 1041b97c-be14-4232-a8e7-8e454e373539 · inbound

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models cites this paper.

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T04:29:48.065577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T04:29:48.065577Z digest=sha256:aa778f0dd08d59828d793493996a486ed9a1d4c369d84b2626693fae9006ce72

Observation 6a50ebe8-ed3b-405d-baff-8aefda09b651 · inbound

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models cites this paper.

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-14T04:27:01.402929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:27:01.402929Z digest=sha256:712a423dc69fc14c1c3afcce36622a77f1543af3b05843358cb0cb492c2526cb