Pith. sign in

Paper Citation Record · LEDGER

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2605.31286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.31286 v2

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:58:48.611343Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T20:46:04.923901Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact31
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d31bef59-942d-4b01-b1d9-40da007ba36e · outbound

This paper cites Qwen3-VL Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Qwen3-VL Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.394231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:1007ddf8f7e974a84910edea16a72f77e992c16a704040735e87c839631d3cf7

Observation dc6a7625-44a6-42e1-9a5b-c858f355c4d5 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.428612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:af5030f8cbd53cd165a09ad0b6fa95470ce6401e9b425da8eeb6dd4f1e436f10

Observation d86c246f-5270-4827-81ed-4d82d80fa67e · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.340717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:70852c9aefa93fa1741b8297299ecb19599bdd4d3e054a5465e2a3d91bd0e765

Observation 86a0d7f4-aa26-40ff-a8cb-1ff68843b1dd · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T19:56:10.423826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0a8c04690dec562e0e1bf8e14c30afd944ca4e9cd8c9b9fd5f7895a94ac0b2f0

Observation b8b71efe-b976-4e16-8f61-badf360c53e5 · outbound

This paper cites Training-time action conditioning for efficient real-time chunking.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Training-time action conditioning for efficient real-time chunking

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.412825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:453c76ee768cc9ea70c1c63bac771d49814c39ae6ac578ead46e17bf52bd6c91

Observation 3baeb24a-5909-4a9a-9bb7-c2ffca01c367 · outbound

This paper cites Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:8b021dac3fb8537c9d5edfd834e415fb7a4d0c6e3d2bb90ef3f18319896106b5

Observation 0bc8d3e1-d01f-4537-93fe-2d7e6d1492b5 · outbound

This paper cites arXiv preprint arXiv:2602.12684 (2026).

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation arXiv preprint arXiv:2602.12684 (2026)

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.421785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3a136e877ddaa59ace7d5929cf9b419108380a3fa3042df34d679c09dcd3fc43

Observation 3b44f106-5c57-48d6-bd86-8313d80eb5d9 · outbound

This paper cites Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:df4645b68b142fbaea65dada4adad957d11daac73aa0ac9180a8bb7d2922e069

Observation 649ee648-8e63-44a7-8195-f24c1d4bb078 · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.427156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:d9c21b66ab9ffcdb43797077a5893880dc0972bdb3863b2e88c3905aade474b7

Observation 3546b866-9125-46e2-872e-6903d860429e · outbound

This paper cites GR-3 Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR-3 Technical Report

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.424772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:27feb1405ea76833b60585beeba73f5503c918177e3f42d0ef666516ec141221

Observation baae2a12-7ed8-416f-a29f-65fc208ddc12 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.429643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:f743e26a6723ce47938cd4dce734660068f22fbe28fb743a7cb0f0c77d529b7e

Observation 69b1b96c-2985-4494-996b-88bd41ff2eaf · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.379180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:ca2d3af027e77ee4b46ca3b836c91df64332f6d6404f82ff95f462a55d350050

Observation 00efa83e-7c76-44d8-ae8b-4b4977cf3de8 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.381185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:cec38b09953dbf3289e49b156d06f17fc8f3dcaf0447af9fba67dcca0de5a0d8

Observation 5b63e49f-108f-4948-98ab-13e3850f0b70 · outbound

This paper cites ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.409160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:a8295ec436f9e8c88bcbec633bf7dc25740429f05c5bae11030b0c9e1fa3c794

Observation f3531602-0d1b-40cc-8a46-a9a56b74f387 · outbound

This paper cites RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.410282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:14d970a8174a1eb0b752a5a3fb009756087dacbad6da2871de99c6860bd0dc7e

Observation f25ab638-847b-491d-a10f-6304913606e0 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.403716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:45164c08ac03d390f4b40183405e2ab01c51bb670818cf389aae6b6496e27f6d

Observation 6502b855-75d6-437b-b569-94d3fa4dfb64 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.398578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:b623fc6df073aa6229fa11169901ed12f85e9bcd3da79c1b3b6b2eae4b9c59ee

Observation 5f1ceea0-ab82-46d1-80f0-3a8b7b3fb9e9 · outbound

This paper cites Galaxea Open-World Dataset and G0 Dual-System VLA Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Galaxea Open-World Dataset and G0 Dual-System VLA Model

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.401207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:7407b3fe2646cba1063168279aa1aad0af1c5f97dceeba4a99632e9c35fd3991

Observation 9971aaf8-5eb7-4986-a1e1-84f9dcfa6e4b · outbound

This paper cites Hg-dagger: Interactive imitation learning with human experts.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Hg-dagger: Interactive imitation learning with human experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:dbb0fe9a15d195d734e9e071520e64977edb6747abf85468f643dccaf9376677

Observation 661eb9c4-c784-4ecd-9c89-3e375c28f63c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.426409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:43625c12b43bcd3660187cd9b25903267e1ed36540af8215c9282daa9430e4ae

Observation c6209d33-3571-4bce-b111-00e70158b444 · outbound

This paper cites Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:85b6b77158c0b585fea02200a2465933d8a14d39079949172f38cdceba5e563e

Observation b2c245e3-e9de-43e1-a95d-67ed23389d15 · outbound

This paper cites Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.417816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:83479c91bb821365bf6135c182fdc8d9925d8dacabaa20b671c29eae7dec0176

Observation 303ed987-7efa-4c1f-9f94-64bdbeb09045 · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gr-rl: Going dexterous and precise for long-horizon robotic manipulation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.396125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:96281087b4ef09c48fc5530f64b7464b80dae2022083a2e8d0b8a987ef3ce1d7

Observation 7efd9857-db2a-4235-adf5-185c32fc30f5 · outbound

This paper cites HoloBrain-0 technical report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation HoloBrain-0 technical report

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.374801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5d622aaedcc9561d0886e568058a3acb945f16d4a0ff5146a60836ecdb218b94

Observation 0bfdb06d-9720-4581-ae5d-4454bc18913c · outbound

This paper cites Flow Matching for Generative Modeling.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Flow Matching for Generative Modeling

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.361856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:87a20f47e4f00b2238509ec4c3bdff14e41a5a5c8c1e8a9905dc4c50f71bfcfc

Observation e0dcd32a-d2b5-437d-931b-7a0379f56b7d · outbound

This paper cites Being-h0.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Being-h0

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.387419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:1e47f33dee1fceb2c91010802439d9d6277d7b3120a546cc67fe48b1565893fd

Observation 02bc5ddc-ea70-40d3-9c9f-bd502fcc6c9d · outbound

This paper cites Human-in-the-Loop Imitation Learning using Remote Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Human-in-the-Loop Imitation Learning using Remote Teleoperation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.370614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3b44348969df32d4e5a184b00e8687e0bc009fd55758be84eae1e5aa12740e55

Observation 6b0e00ea-2806-42c1-b3f1-4da682061fc8 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.383901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0d948d0ddd7053415be0e17112c22b76d6d32dd33dbf24a010eb9ee59f337bf2

Observation 282d1af0-5a82-4121-8dda-60230e9a3c7a · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:e22104c10a81d1e0cb4c29d3d852c5bb818d5261d9c3dc0f83986a7433e2ac32

Observation 30b04f0f-7b75-4878-8525-60865447498e · outbound

This paper cites Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.390427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:ca374f4cb1e1d23cb3b4b0c7f44d4fa9d0f79c7ddf5b04246561fcd0182e0c92

Observation 144a15e2-87f3-4934-bac9-a89fd6d816c5 · outbound

This paper cites RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.411949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:9f94df9a6c2d7cee6dea770052cfc27f0b3373cff14fe5052cc4683eaecc3500

Observation a827281b-00c4-4db8-a95e-0a4cb8330a04 · outbound

This paper cites A Pragmatic VLA Foundation Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A Pragmatic VLA Foundation Model

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.415298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:208466e184c0827978c50cefde233c555ca10e099b9afa6cc4235027d886c51d

Observation 3226f6d6-6d58-4bd1-882d-ebfc51706b56 · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.393321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:8833485fc57e1fbe9882f789f54e0dacd5f67dccc931b3d2df5af6b1336771ef

Observation 1e35cda1-f2a5-42f0-b844-59321df3fb55 · outbound

This paper cites χ0: Resource-aware robust manipulation via taming distributional inconsistencies.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation χ0: Resource-aware robust manipulation via taming distributional inconsistencies

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:cc2ae1a32da59bbc9fa3c5910e293a408784471b9b16568279c1fba540afb3e5

Observation 5d4ce01d-0765-4e20-a9ba-43dea1aa92a6 · outbound

This paper cites Igniting vlms toward the embodied space.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Igniting vlms toward the embodied space

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.376197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0b2f2243f744bfdd41a4acf923d27e118becaad9aad4de1fb5a266dea0cc7352

Observation de67d6c5-8687-4df8-acd0-8354226005d1 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.353861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:fc9d4a5261d62963853757cb7020c0fe79d2bb84cad993e3e28f887220087cda

Observation af956799-8209-4fc5-81da-52d926799d12 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.363522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:970942d38b0ab133e8688f27e13f3976981147caac12dda20c608ab355036c3e

Observation b3ef4f82-d55c-4413-bd89-f10d377a319c · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Rt-2: Vision-language-action models transfer web knowledge to robotic control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3947ba8095666e133e2f1239c36e5fb611e0a19a8ea784bea734adaec9fb443c

Pith citing papers

Observation a8523749-3fbd-4682-9e55-639488fd01b0 · inbound

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects cites this paper.

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T20:46:04.923901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:46:04.923901Z digest=sha256:364f18f2570afc00c38b42d35175994fc8f78f9afab0570ad950ca8c0c39e00b

Observation f8f84011-ab4c-4c8d-a5ac-eb0591ac4d79 · inbound

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation cites this paper.

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T11:21:18.588143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T11:21:18.588143Z digest=sha256:c32ea1e6966414f4c9ac96a119f526f45849eb32a199f884487007eda6398caf

Observation 1ccb8264-a599-42c5-bf2e-e74afb398eb1 · inbound

Learning 4D Geometric Priors for Inference-Efficient World Action Models cites this paper.

Learning 4D Geometric Priors for Inference-Efficient World Action Models DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T14:23:57.266710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T14:23:57.266710Z digest=sha256:b7072a4a34778841df94ebfba391b3f1d35683691ff36abb7ab3e85772e029b4