Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:15:34.751856Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 4 inbound Pith citation observations for arXiv:2508.13104.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:15:34.751856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-11T17:36:01.254174Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T12:49:53.218336Z
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 60256d56-1918-4bb7-b7ce-f244880d90be · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1daa2b1b-c9ed-47ff-9302-d7b07c4255a8 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Cosmos World Foundation Model Platform for Physical AI
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a4cb098-bed9-4131-b721-3a24e55fcf06 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts InterDyn: Controllable Interactive Dynamics with Video Diffusion Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cfdc77f-8959-4e24-b4a0-2b4519aab69d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Diffusion for world modeling: Visual details matter in atari
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5075b4ea-386a-4733-be38-7dd996117d1d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70878d32-88ff-450a-9db4-193333146606 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Introducing hot3d: An egocentric dataset for 3d hand and object tracking, 2024
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a755468e-3c01-4254-b336-fec4b6e284bb · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Lumiere: A space-time diffusion model for video generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d3787ea-af80-409d-af2e-01179bc12842 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Automatic rigging and anima- tion of 3d characters
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78af8e13-775a-43df-af28-0c2d8a3144e8 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4738a33-af55-4a7b-b4a4-8dc1d9898485 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c495f89-e5a7-439f-8d27-ec80194f2e00 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts RT-1: Robotics Transformer for Real-World Control at Scale
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c203145b-f52a-4401-9386-5b8a5b1a257a · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Genie: Generative interactive environments
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d5a3834-d60e-4088-bd2b-457991f717b8 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts 1C filter: a simple speed-based low-pass filter for noisy input in interac- tive systems
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b456ac54-02b6-492e-a7f8-ee9f91467e62 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts GameGen-X: Interactive Open-world Game Video Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfc4a088-9b51-4689-8f1e-d636e08ec80b · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Scaling egocentric vision: The epic-kitchens dataset
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2da638c3-b55e-4aac-ade5-58b95b99c301 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Oasis: A universe in a transformer,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3df8d26-b67a-4672-97e2-625ed288367b · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Genie 2: A large-scale foundation world model,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0596b06a-5d90-485a-89fc-5d833fc7f272 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Motion capture from internet videos
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 745f3d2c-3b99-4f3f-b138-c0e89ca35bb0 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Tenenbaum, Dale Schuurmans, and Pieter Abbeel
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c33008-0a3d-4f81-9981-1bf6483dec1d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Arctic: A dataset for dexterous bimanual hand- object manipulation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9187ef7e-f0f3-4c0b-bb64-a00df3bfad4d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts The Matrix: Infinite-Horizon World Generation with Real-Time Moving Control
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bc7d8dc-607f-4fc9-a95c-df317c42ff33 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Gigahands: A massive annotated dataset of bimanual hand activities, 2024
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d6592d7-74cf-42d5-9acb-ce5335339722 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Vista: A generalizable driving world model with high fidelity and versatile controllability
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67a4cd9-5fa8-426e-9919-e6c2f4ce8062 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Motion Prompting: Controlling Video Generation with Motion Trajectories
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ea2ced9-fdde-4c0d-88cc-59a759ccb189 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts The ”something something” video database for learning and evaluating visual common sense, 2017
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2243be9b-abab-4a12-abb2-846fb1109606 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Ego4d: Around the world in 3,000 hours of egocentric video
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b01cae2c-f5d8-4898-a66b-b83f3949bae1 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Ego-exo4d: Understanding skilled human activity from first- and third-person perspectives
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d956fe0-0779-45b0-a5fc-97ce2918c6dd · outbound
Precise Action-to-Video Generation Through Visual Action Prompts MatchAnything: Universal Cross-Modality Image Matching with Large-Scale Pre-Training
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c1a7c24-223f-41cc-aef4-50d0f1328e73 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Hand-eye calibration
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b25704f9-b7e5-4ec6-9c08-b1ef38a1f23c · outbound
Precise Action-to-Video Generation Through Visual Action Prompts LoRA: Low-rank adaptation of large language models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ff48e37-f7db-43c7-8cc9-9007c51c5501 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Animate anyone: Consistent and controllable image- to-video synthesis for character animation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aa7adb4-b970-4a1b-b514-7fbeeba0fa7d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813318ee-f29c-464b-922a-6c6e56ba610f · outbound
Precise Action-to-Video Generation Through Visual Action Prompts CoTracker3: Simpler and Better Point Tracking by Pseudo-Labelling Real Videos
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9132b549-d6a0-4de2-88d8-b5841df3d597 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts The kinetics human action video dataset, 2017
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d50df9b7-850b-4f6a-bcc9-30dccab68493 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2725ab3-fc2f-4225-aedf-0470386e666b · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Sapiens: Foundation for human vision mod- els
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1065572-00dc-42ae-aabf-adf95f136f96 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Learning to Act from Actionless Videos through Dense Correspondences
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0685dd8e-9c77-4bfd-9ec1-7256dccc8706 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Pose space deformation: a unified approach to shape interpolation and skeleton-driven deformation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3e93b1b-f504-47ef-9f7e-fb46ff0a4d09 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Wonderland: Navigating 3D Scenes from a Single Image
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0742e22e-6796-4822-8fdd-5499e550eeab · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Taco: Benchmarking general- izable bimanual tool-action-object understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce376c2-119d-4b1b-bb34-8dbc05b709fb · outbound
Precise Action-to-Video Generation Through Visual Action Prompts The BabyView dataset: High-resolution egocentric videos of infants' and young children's everyday experiences
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d10c0c7f-277f-4683-866b-e7f5739a9199 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts MediaPipe: A Framework for Building Perception Pipelines
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c9e602f-c4d6-4f7f-befa-3be7ee9e27df · outbound
Precise Action-to-Video Generation Through Visual Action Prompts MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ca9129-a756-4b67-9b64-e36f1bd4e604 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Do generative video models understand physical principles?
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81e2e1f7-a9e4-4e20-a133-5215c25c9563 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts A survey on deep learning for skeleton-based human animation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d7e9f7-ad75-4f5d-9473-5e51d7c31b0c · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Openai sora, 2023
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17bc754a-3a3e-4632-99da-c975150821bb · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5880a6a6-32bb-4b1c-9441-42314804c31a · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Computer animation: algorithms and techniques
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66b4011d-db36-4198-9a9a-cb93a57612ff · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Scalable diffusion models with transformers
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ad72c32-8f59-4344-b9e1-ce4d8a19e61a · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Sampson, Shikai Li, Simone Parmeggiani, Steve Fine, Tara Fowler, Vladan Petrovic, and Yuming Du
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 333eddb6-26d6-48b0-97e6-d63793da210e · outbound
Precise Action-to-Video Generation Through Visual Action Prompts WiLoR: End-to-end 3D Hand Localization and Reconstruction in-the-wild
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7867f6f-f8ec-4e64-a853-463835f4df2c · outbound
Precise Action-to-Video Generation Through Visual Action Prompts SAM 2: Segment Anything in Images and Videos
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1b1493d-301d-40c6-b942-fbfa0484ab2b · outbound
Precise Action-to-Video Generation Through Visual Action Prompts World-grounded human motion recovery via gravity-view co- ordinates
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56c0324b-d41f-4071-9e1f-4162b3135e57 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Motion-i2v: Consistent and controllable image-to-video gen- eration with explicit motion modeling, 2024
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a760d21-6dbc-4a72-b1c9-c4d977d0d174 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Genhowto: Learning to generate actions and state transformations from instructional videos
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7ddd85e-195c-4d2d-9b7a-6795245cb9b7 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Optimal hand-eye cali- bration
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17874761-d284-4e65-ba06-b254102ef742 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Controlling the world by sleight of hand
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ca3260-b9bd-4bf6-8d98-38558533cbfd · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Raft: Recurrent all-pairs field transforms for optical flow
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a59459fc-d4f8-468d-a52a-4ba3aedb1085 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts A new technique for fully autonomous and efficient 3 d robotics hand/eye calibration
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caf6a23d-d957-4466-a80c-918dc681d3d6 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Fvd: A new metric for video generation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69492a8f-750b-4b60-bbb9-21a373efdea6 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Diffusion Models Are Real-Time Game Engines
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e07eb07-a377-4dd4-b077-448780dd9982 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Boximator: Generating Rich and Controllable Motions for Video Synthesis
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d70202-8869-4e07-a7eb-a46f1a171958 · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Motion Inversion for Video Customization
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d4ee067-2a01-470c-a0e8-d8cbe31222ec · outbound
Precise Action-to-Video Generation Through Visual Action Prompts EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9e55f6f-982f-4e2a-91d3-5804930451cb · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Image quality assessment: from error visibility to structural similarity
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b236e7-d8f2-49a4-9292-c167d04d171d · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Mo- tionctrl: A unified and flexible motion controller for video generation
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8435d1d5-5001-42f7-a240-8e5b2d7f2baa · outbound
Precise Action-to-Video Generation Through Visual Action Prompts ivideogpt: Interactive videogpts are scalable world models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3efa7fda-a3a9-4916-9a96-170102a4f36a · outbound
Precise Action-to-Video Generation Through Visual Action Prompts Samurai: Adapt- ing segment anything model for zero-shot visual tracking with motion-aware memory, 2024
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5146c375-6962-4624-8511-7a7893229b71 · inbound
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints Precise Action-to-Video Generation Through Visual Action Prompts
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92200f97-2258-4608-8c8a-e82fe275d20b · inbound
PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Precise Action-to-Video Generation Through Visual Action Prompts
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 63f2e8db-c8e4-4c18-b917-fd3e1bf5c50d · inbound
PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Precise Action-to-Video Generation Through Visual Action Prompts
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 58c0452d-2ab4-4841-a0c3-ca847f61293a · inbound
Mask2Real-WM: Segmentation Masks as a Sim-to-Real Bridge for Controllable Dexterous World Models Precise Action-to-Video Generation Through Visual Action Prompts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.