Pith. sign in

Paper Citation Record · LEDGER

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

As of 20 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2605.31286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.31286 v2

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:58:48.611343Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T20:46:04.923901Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact31
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d31bef59-942d-4b01-b1d9-40da007ba36e · outbound

This paper cites Qwen3-VL Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Qwen3-VL Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.394231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:90708fedcf7853ff42062084e4f9b00b89a1bba7443077ccba43cf0d3234037b

Observation dc6a7625-44a6-42e1-9a5b-c858f355c4d5 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.428612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:8b8907fb5a4da04352a90e7403847583eec1fde3f914089ad99d33153886c985

Observation d86c246f-5270-4827-81ed-4d82d80fa67e · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.340717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0581fe3809ee3be278303519eb2d6147ad526e2649a383672c561d9f14300f92

Observation 86a0d7f4-aa26-40ff-a8cb-1ff68843b1dd · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T19:56:10.423826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:7c1da4af44cccb78251bf82a3b9ac0bb4ef439c7d4965e29f335859e8190e4fb

Observation b8b71efe-b976-4e16-8f61-badf360c53e5 · outbound

This paper cites Training-time action conditioning for efficient real-time chunking.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Training-time action conditioning for efficient real-time chunking

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.412825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:1d6b0c7621dad93e0779e2bec68b73f524f854314888826b5b7b65d5c19a9b47

Observation 3baeb24a-5909-4a9a-9bb7-c2ffca01c367 · outbound

This paper cites Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:8b021dac3fb8537c9d5edfd834e415fb7a4d0c6e3d2bb90ef3f18319896106b5

Observation 0bc8d3e1-d01f-4537-93fe-2d7e6d1492b5 · outbound

This paper cites arXiv preprint arXiv:2602.12684 (2026).

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation arXiv preprint arXiv:2602.12684 (2026)

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.421785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:6ef11098b04872e870cd61ec1bc807fab4abdadaa18883d4000d54dd656cce3f

Observation 3b44f106-5c57-48d6-bd86-8313d80eb5d9 · outbound

This paper cites Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:df4645b68b142fbaea65dada4adad957d11daac73aa0ac9180a8bb7d2922e069

Observation 649ee648-8e63-44a7-8195-f24c1d4bb078 · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.427156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:9dc48c156de0c1eb8ab9d8d1d3afc145cd08f9e5eae73d2de41a000037146239

Observation 3546b866-9125-46e2-872e-6903d860429e · outbound

This paper cites GR-3 Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR-3 Technical Report

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.424772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:1b0fc763cca05be84ce3fd4ffcb49b5a42c24d2a612e54a44457970221bfbaaf

Observation baae2a12-7ed8-416f-a29f-65fc208ddc12 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.429643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2040263a0021531dc593fb70e78a8a29cdf7b3557699b2fefcd331134a91afd6

Observation 69b1b96c-2985-4494-996b-88bd41ff2eaf · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.379180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3d4be9d27814abc043befe35bdeb2ed68dad34d4c9474ee71011082366d7a999

Observation 00efa83e-7c76-44d8-ae8b-4b4977cf3de8 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.381185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3241f742f095116ee251d9e4ac15d3dea3ac2ebbbdd4bd42147444892acd0e1f

Observation 5b63e49f-108f-4948-98ab-13e3850f0b70 · outbound

This paper cites ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.409160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:4e42cf29a7f0e7bd43fd7b1fd9eb200ead37e576217448306e7315ac5aff0217

Observation f3531602-0d1b-40cc-8a46-a9a56b74f387 · outbound

This paper cites RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.410282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:b95e6545b4393834934e299c6166fd1045d65870e7c001291dcf7ca0dc8c5cef

Observation f25ab638-847b-491d-a10f-6304913606e0 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.403716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0d7b4677bb0bce59b4f9edc832a106b11991df9e7292b1fa38be26fa746b47e0

Observation 6502b855-75d6-437b-b569-94d3fa4dfb64 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.398578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:c45d18e59df30a6cbcb7444c90ca931d82edfb5c99c8b9b263bcab392669d46d

Observation 5f1ceea0-ab82-46d1-80f0-3a8b7b3fb9e9 · outbound

This paper cites Galaxea Open-World Dataset and G0 Dual-System VLA Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Galaxea Open-World Dataset and G0 Dual-System VLA Model

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.401207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:e1bf1b79751f4cfa7ec5b6e03179af72ba19f8a948a5d8963bbe549950e563e1

Observation 9971aaf8-5eb7-4986-a1e1-84f9dcfa6e4b · outbound

This paper cites Hg-dagger: Interactive imitation learning with human experts.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Hg-dagger: Interactive imitation learning with human experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:dbb0fe9a15d195d734e9e071520e64977edb6747abf85468f643dccaf9376677

Observation 661eb9c4-c784-4ecd-9c89-3e375c28f63c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.426409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5cbb5d48e2e05030fcd6ea8385b83cd27af87ee87a7593c0522968da44ef53b8

Observation c6209d33-3571-4bce-b111-00e70158b444 · outbound

This paper cites Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:85b6b77158c0b585fea02200a2465933d8a14d39079949172f38cdceba5e563e

Observation b2c245e3-e9de-43e1-a95d-67ed23389d15 · outbound

This paper cites Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.417816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:e38924a7d7b9dde768408f09558909ab1a9bc085d570af21926855d951da11c3

Observation 303ed987-7efa-4c1f-9f94-64bdbeb09045 · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gr-rl: Going dexterous and precise for long-horizon robotic manipulation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.396125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2d08f0f5e067388289e3160f62a33027940a611b6c434dd307cbd0cf1941d1ee

Observation 7efd9857-db2a-4235-adf5-185c32fc30f5 · outbound

This paper cites HoloBrain-0 technical report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation HoloBrain-0 technical report

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.374801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:6b7d4254aae28e4f7738e569ce5f37b643a48565ed1ef8d89fd19bea3efa5a17

Observation 0bfdb06d-9720-4581-ae5d-4454bc18913c · outbound

This paper cites Flow Matching for Generative Modeling.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Flow Matching for Generative Modeling

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.361856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:c5beb7e30067a2e5a9c990d05820b0a907feab4f0d815e34cc6908664266a1ed

Observation e0dcd32a-d2b5-437d-931b-7a0379f56b7d · outbound

This paper cites Being-h0.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Being-h0

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.387419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:054ac3754059d2a964aa3613cb3983d3be30a9d60a1022af096db1ce13529fcf

Observation 02bc5ddc-ea70-40d3-9c9f-bd502fcc6c9d · outbound

This paper cites Human-in-the-Loop Imitation Learning using Remote Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Human-in-the-Loop Imitation Learning using Remote Teleoperation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.370614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:ea4d1bc04f915f2104bc5cad4ff339d193fae8d05c08aaa5e3397d090644145d

Observation 6b0e00ea-2806-42c1-b3f1-4da682061fc8 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.383901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:7afd95fbb219d5186cb4f270a7c85337b71034c96c4fdf76af0fd2207257fb43

Observation 282d1af0-5a82-4121-8dda-60230e9a3c7a · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:e22104c10a81d1e0cb4c29d3d852c5bb818d5261d9c3dc0f83986a7433e2ac32

Observation 30b04f0f-7b75-4878-8525-60865447498e · outbound

This paper cites Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.390427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5d36faeac1b337b8711394bf5ca92b976a208ff416ede57b2e5d468c6b3a9bae

Observation 144a15e2-87f3-4934-bac9-a89fd6d816c5 · outbound

This paper cites RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.411949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2f9ba2f78f15dd53078a5f0bf3586015d5c94c1b7f2aa7cdc789b751f7a685ee

Observation a827281b-00c4-4db8-a95e-0a4cb8330a04 · outbound

This paper cites A Pragmatic VLA Foundation Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A Pragmatic VLA Foundation Model

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.415298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:ae148e667a819e02275087550eb37673034abcb8e45ecacfa3625289f588a8b4

Observation 3226f6d6-6d58-4bd1-882d-ebfc51706b56 · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.393321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:760c42958858545b714ba2b75a48458c90139daff3119e0f90638d99b24c7c2c

Observation 1e35cda1-f2a5-42f0-b844-59321df3fb55 · outbound

This paper cites χ0: Resource-aware robust manipulation via taming distributional inconsistencies.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation χ0: Resource-aware robust manipulation via taming distributional inconsistencies

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:cbcfff362bb8ed7710b9ba62625ca8ada79eeac1cdf87fe9e7f5d42be09f554c

Observation 5d4ce01d-0765-4e20-a9ba-43dea1aa92a6 · outbound

This paper cites Igniting vlms toward the embodied space.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Igniting vlms toward the embodied space

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.376197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2979e94b7a36500347ab77d00321604a0b5e3e6f7fdaae61db837b8082a6db68

Observation de67d6c5-8687-4df8-acd0-8354226005d1 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.353861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0235fe78a315cc07b014c980db1d6e565c4b2cc3e6527f0fa9578160edd26404

Observation af956799-8209-4fc5-81da-52d926799d12 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.363522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5d4121fc4f7dad363b6485784b5d57436f9dc8534010d7216e94f9b8d14bb288

Observation b3ef4f82-d55c-4413-bd89-f10d377a319c · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Rt-2: Vision-language-action models transfer web knowledge to robotic control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3947ba8095666e133e2f1239c36e5fb611e0a19a8ea784bea734adaec9fb443c

Pith citing papers

Observation a8523749-3fbd-4682-9e55-639488fd01b0 · inbound

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects cites this paper.

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T20:46:04.923901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:46:04.923901Z digest=sha256:364f18f2570afc00c38b42d35175994fc8f78f9afab0570ad950ca8c0c39e00b

Observation f8f84011-ab4c-4c8d-a5ac-eb0591ac4d79 · inbound

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation cites this paper.

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T11:21:18.588143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T11:21:18.588143Z digest=sha256:c32ea1e6966414f4c9ac96a119f526f45849eb32a199f884487007eda6398caf

Observation 1ccb8264-a599-42c5-bf2e-e74afb398eb1 · inbound

Learning 4D Geometric Priors for Inference-Efficient World Action Models cites this paper.

Learning 4D Geometric Priors for Inference-Efficient World Action Models DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T14:23:57.266710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T14:23:57.266710Z digest=sha256:b7072a4a34778841df94ebfba391b3f1d35683691ff36abb7ab3e85772e029b4