Pith. sign in

Paper Citation Record · LEDGER

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos

As of 13 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2607.11397.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.11397 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T05:52:34.589171Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e13f6412-e76a-491d-a514-5974249018dd · outbound

This paper cites DINOv3.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos DINOv3

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:2485b3eb85a83207c2064c82d863c95935e0552b72fc4ffe9bde28cdc12489b2

Observation 409e3c19-06cf-4ffa-891b-bd889ee5365b · outbound

This paper cites Depth Anything V2.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Depth Anything V2

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:f85632091894bca2adafe3a60a47c17240fa96a1c5903232eb2c9653fe83b70e

Observation a15d9ece-09f5-4fa5-b969-7da18dbe3a8b · outbound

This paper cites EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:020460f821d81944882ddd16bebe48115d928bb5890dc28f3cc3179135552073

Observation 693c8148-6448-4a28-852a-30879f48ee49 · outbound

This paper cites RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:eb5f4ad8de18cbb6a1b7b8b94d46584ff09728a3a5b719feeecb93a710a12b30

Observation a04ded9e-f9bd-4821-9c43-87d8e9c28221 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:7d6cbe64763371ed7aa89f5644e19947003f36307ea6acef2b703126d629ceed

Observation 6aae0788-def2-4fb9-ab50-20ba9b9ee5ea · outbound

This paper cites RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:819f826373e7155c70ecc2210392891cec363b74292a36632e6d6c832094ab3d

Observation 4527748e-65b8-4904-ad8e-6775d6d91f0c · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RT-1: Robotics Transformer for Real-World Control at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:44e343024ade07c9bb78ab005f0aff0451de8901060ebd0c354cb9c009f10259

Observation 0130bd01-cadf-4479-8120-f560a3cb7e40 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:a4fb834213dce45a5b453e417db4b7902892edf20360d1684d1a92b5c8e8488c

Observation 38929368-2b45-493c-bc76-698097edb913 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos OpenVLA: An Open-Source Vision-Language-Action Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:c93454a7e65132c5cc49f6c1e52a03afbf5dac77e6cdcd3260fd1dfc2c7c290f

Observation 19b90573-6c63-4a67-b551-b9df87433143 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Octo: An Open-Source Generalist Robot Policy

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:6204d5c244e0e3a27e261e0a3b3a34908df5cfce6ddb37edc5420c8ec8559c13

Observation 0224757f-32b9-4aeb-b1b6-13e7010c3f64 · outbound

This paper cites π0: A vision-language-action flow model for general robot control,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos π0: A vision-language-action flow model for general robot control,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:0837c38df43808c85a28d5567df8e6e2b127b240f16a95b6c0ede301150dfaba

Observation cdd1f35a-a47d-42fa-81fb-954f2e188e66 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:27313c66c7a717634416bf1801d2889e9fc698a9335e4301260705788fa1efae

Observation 6c9b1331-017e-47f0-96cb-78b8c9b13883 · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:49d112820f2c2e837ca2d4e5c74e0933b989a12585620b444ab17a4a6e3e3fbf

Observation 94a0c9c8-3670-45be-9e14-c3a5147e7001 · outbound

This paper cites Causal World Modeling for Robot Control.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Causal World Modeling for Robot Control

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:ff149070a944090ea67fae0962725f11f07b5d6194345d73e087fd1c997dbd54

Observation 9625656e-7c4b-40b8-acb9-2e7fda05eb94 · outbound

This paper cites LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:b0f89e8cb52634f3cc25b83bc5eb758fb34da0eb051de44130755d557086b7d4

Observation a1a21988-ddb1-4cae-a221-8c1dd96609ff · outbound

This paper cites World Action Models are Zero-shot Policies.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos World Action Models are Zero-shot Policies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:9e31fd128d470241287085d94d64af19290b2590a8747cedf86b4f78b060f4a3

Observation a8ddf504-e257-45e0-a1ad-54027e2d70ab · outbound

This paper cites Motus: A unified latent action world model,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Motus: A unified latent action world model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:d30c75282297a51aaf7a73c982d7be63ea96f1c96e724ca690be7018831cfde7

Observation 59bb24ef-efbe-433b-a9b7-d23c44443329 · outbound

This paper cites Latent Action Pretraining from Videos.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Latent Action Pretraining from Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:a1462bca63db3035dfead73d49b20a79a8515e1a29dabfbb193e44591223f1cf

Observation 4103b645-e698-4860-9693-bd739ef6c59d · outbound

This paper cites Moto: Latent motion token as the bridging language for learning robot manipulation from videos,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Moto: Latent motion token as the bridging language for learning robot manipulation from videos,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:2c768f03258b0aa81318856d2e66d978e7b5116eb4b6daa36ec755452d6f98cd

Observation c6393ffb-9f0a-4715-970f-f519d05edc6e · outbound

This paper cites Univla: Learning to act anywhere with task-centric latent actions,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Univla: Learning to act anywhere with task-centric latent actions,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:be6932a56c040f86d2223f2d339f91836edb29dbcd65f40ce1ab8c60bff54040

Observation 340c69ea-265c-4627-b9ac-346ef3594dfe · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:577cfc58451a73e4aa3cfb4f5dc6a423b3d1319e37ad0a31f2fc44d59450010c

Observation 0882fa01-a819-42df-8335-27a82fc14cfd · outbound

This paper cites UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:b961a71a3e05d86f48895f5de4ab8d72c4b26b871ad6221feb9cf584dd96fb92

Observation 1b21bf43-ba97-4099-a825-80c480399f5c · outbound

This paper cites villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:17e1ec6e61f579cf1c442a048bec579809d18c960727a59049c2d9aa60f19277

Observation 9850e7a2-7e73-4f09-ae20-2f45c696c936 · outbound

This paper cites Qwen3-VL Technical Report.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Qwen3-VL Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:92592e7a508d91b44140749be12890b6d8a48a3f0118a8466f78614ba990204d

Observation ebdff35c-a361-4e3a-a531-07472ad81ea8 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:f27e3a8a2747116bc808c2f9ba1c6b56b6fa216e266bf7faefc5c45af7f31158

Observation 4590de90-8d21-40b2-9f49-59f7c3510fa1 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:cd2553db82be32f3145d19c41f73c215ef9e60e3c1b32b28de6a2593948d401d

Observation 8c1c41b8-d55a-49b9-9e6e-399c3d54419e · outbound

This paper cites StarVLA-$\alpha$: Reducing Complexity in Vision-Language-Action Systems.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos StarVLA-$\alpha$: Reducing Complexity in Vision-Language-Action Systems

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:583ca0660f9c7ca749f54527d9a8b1eb7706b898b72f5f833d9e75603e61bff6

Observation 6d76f414-18a5-4d28-ae28-3d2d9491b22e · outbound

This paper cites Internvla-a1: Unifying understanding, generation and action for robotic manipulation,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Internvla-a1: Unifying understanding, generation and action for robotic manipulation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:915c06d031e79bdf6b5e411aa72ea96deaf83604da5694944b91b5489acba226

Observation 52a0440b-bfff-464e-b8ca-d995cfa3364d · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:b9771b5c024fb6ce208af2a6c0c6f6a40a915debae0d9232889c3ec6e955c9cb

Observation d231e741-0f75-4edc-9e22-1e7fac3bf2aa · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:6a13d08cd75dc16e3aaf515a927b5565ce843dec63550348df3ab6fe056df808

Observation 14bbfdd0-4c0d-4537-b4d1-e77f57c3cfec · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:47be457e8518e1b9d77f31643a9cfcfb949f9ee4d5979985e0361247f10145d4

Observation 2670a3db-1c2e-4455-a9d0-c1bffb91d88b · outbound

This paper cites RLDX-1 Technical Report.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos RLDX-1 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:e997ab6f75d74207125deefc96e4d38ff641eb23f2fca3c4e8adc08583a9091a

Observation 284b221b-72d1-4f40-958e-52b4fd8ad0a6 · outbound

This paper cites FrameSkip: Learning from Fewer but More Informative Frames in VLA Training.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos FrameSkip: Learning from Fewer but More Informative Frames in VLA Training

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:776e9b1e3cfd6526f8d7dbe5cd9e5bae0f33a76c04ad5c672a5496ba30a4b9c6

Observation 63faa1a1-8b7c-45c4-bc74-99818282425f · outbound

This paper cites Dit4dit: Jointly modeling video dynamics and actions for generalizable robot control,.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos Dit4dit: Jointly modeling video dynamics and actions for generalizable robot control,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:078b095566913068c70b48c6d2942236f5c8735d80db7cf250513795911eb2cc

Observation f7bb466b-d7c2-499c-8d49-404f97cb8a5f · outbound

This paper cites DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA.

WALA Learning Executable Latent Actions from Action-Labeled Demonstrations and Action-Free Videos DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T05:52:34.589171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:52:34.589171Z digest=sha256:b83a02327b894857c8c9a5e50745e34151d3e04758528709cb2b4f7d5b390a33

Pith citing papers

No inbound Pith citation observations are available.