Pith. sign in

Paper Citation Record · LEDGER

Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2206.11795.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.11795 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:46:06.424312Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

50
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 986c05f3-628c-4994-b984-5400d767e7ae · inbound

A Generalist Agent cites this paper.

A Generalist Agent Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:24:49.880515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:24:49.833638Z digest=sha256:c83f85914d66b26641e7bf87334859dcbc446ed76c873a83165cd58b504a2b7a

Observation 277ff47c-34e9-4d15-8d37-cc93fd2ce4df · inbound

Mastering Diverse Domains through World Models cites this paper.

Mastering Diverse Domains through World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:08:22.082585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T09:08:21.677362Z digest=sha256:5235bc4518812c8e405a9887828eb88a86a7e2a06124a0c259c0c6b046d02938

Observation ef1029a3-d927-436a-bacf-a94fcdbd3cdf · inbound

Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents cites this paper.

Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T03:27:40.743136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T03:27:40.524895Z digest=sha256:d9e1d0f9c216cbecfdaa825610586e4baf9c7e738da363a374f604efb754e701

Observation f8c883c3-b099-4fcc-95d4-879a812a740b · inbound

Voyager: An Open-Ended Embodied Agent with Large Language Models cites this paper.

Voyager: An Open-Ended Embodied Agent with Large Language Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:11:41.059888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:11:40.995345Z digest=sha256:8a338f26299118935a408745275ed8ebe3f04126c799fbda460822315b80fa1a

Observation fce01e15-727f-481a-a06c-901bd4b06853 · inbound

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution cites this paper.

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:12:31.274505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T08:12:30.984870Z digest=sha256:e9ae93ee829ac521b87b69da143ebcdb9fcbf6f8fcf54a8cbb9b9c09ffb9a352

Observation 617edcce-b5d0-4431-ab19-6ed9ae3abe8b · inbound

Learning Interactive Real-World Simulators cites this paper.

Learning Interactive Real-World Simulators Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 219

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:15:18.445102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T02:15:18.265190Z digest=sha256:6ce623b05c902ebec32e9ff61df62cc8e44e16ccd8cafc8ffe4b3b616bc3fbac

Observation 676efef6-edf1-49f2-ac33-e253b0d5d32a · inbound

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies cites this paper.

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:22:44.385372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T08:20:05.898025Z digest=sha256:d8bf61f9c5d27c83172e08273fd2a17bef8326b2fcea2f3dc7fb586d5c4836c0

Observation 45edeb9c-7083-458f-8200-889add293bd6 · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:50:45.650952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:bf67cfdeb472cecade85b512949ed49ea9f1ab5f35fbd48f7cf312264708438c

Observation dcf482dd-1dd7-426a-ae45-b605024bf623 · inbound

Scalable Multi-Task Reinforcement Learning for Generalizable Spatial Intelligence in Visuomotor Agents cites this paper.

Scalable Multi-Task Reinforcement Learning for Generalizable Spatial Intelligence in Visuomotor Agents Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:06.424312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:46:06.424312Z digest=sha256:a47a0abcbb189d69d318a3340881e060ad41afa67cb80af6a0a04c214e98f24f

Observation 6737d5a0-3899-4a16-84f7-71eba7eb3c2a · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:45.747217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:45.747217Z digest=sha256:d87ee85bf0f5d901aeac24416fe49972d725dd37e08cc1744b691e94191b99e1

Observation 49b175a3-22db-4fad-a340-e10665f6aa08 · inbound

PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments cites this paper.

PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:58:44.340988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:58:44.340988Z digest=sha256:07ac29e41254aed2f46412938f2e85f577b014a969cb87af52e2671085e3c500

Observation a9059ad9-2301-4408-90f9-24b24924b0ef · inbound

CA2: Code-Aware Agent for Automated Game Testing cites this paper.

CA2: Code-Aware Agent for Automated Game Testing Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:04.989946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:53:26.558042Z digest=sha256:be07a943bf802e2d6acea584a6d12aff7286361ad711057b6d0de26b54168f54

Observation 9ce58f77-7b49-406b-897e-8d4da5292e23 · inbound

ASH: Agents that Self-Hone via Embodied Learning cites this paper.

ASH: Agents that Self-Hone via Embodied Learning Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.482966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:51:53.328977Z digest=sha256:5e110822a238c1ebbae9218e90e29c736fda64cfe3e3f0f68123b3c737ee70e6

Observation cdb9c999-d720-4bd8-8ec8-da56f66bc0d6 · inbound

ASH: Agents that Self-Hone via Embodied Learning cites this paper.

ASH: Agents that Self-Hone via Embodied Learning Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.116021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T21:12:54.453886Z digest=sha256:5b183a3be9d2f756486df5155d7ed3cfc931bf9599b8ac4e65828e16c2dd9eae

Observation bc03e82d-e8dc-42d2-879a-8df64c52aee6 · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:33:04.182122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T05:28:50.354662Z digest=sha256:c05025d8adab1d87980392fc790f4877b6421f3959e022ce89d917f414e9addc

Observation 302b840e-9cec-4864-b8a9-e6955826f61d · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:39:49.272258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T07:36:12.214949Z digest=sha256:fa1a02c42c5e404faeb9c0760d2bf53bf0512311a4322c592bbc8c8350296dee

Observation d357f8c1-c2f0-4f40-8fb9-2991c3e585e8 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.589878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T19:12:22.513577Z digest=sha256:8d51afa8b50f89a25cf08073d77b5e0c90e4be0bba35d1b70d1d699c77bda533

Observation dda212b4-94fa-47a6-8991-c5c1ed9fd39d · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.595338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:11:07.089829Z digest=sha256:94a6ae9a449f1b6cb95575b69b285766206bab780ffc09ac3c10c593bbf6147b

Observation de3919ae-247d-4613-8294-24c54513fe49 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T12:05:57.682386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:05:57.682386Z digest=sha256:3d0cefb2e533bc293aea9c5bd687b6b566a5384660744d4912838ba0b275e4c8

Observation 1b6a113c-79dd-44de-884c-bd4feec47771 · inbound

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models cites this paper.

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:53.233336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:46:21.198781Z digest=sha256:7e60fa197ffcf512a03ee2c76430f8557c508afd33be849b07f3d8390fa57aeb

Observation 1801af48-6d70-4ac6-95db-e513da3ce0c3 · inbound

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models cites this paper.

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:04:39.292257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T10:19:06.268547Z digest=sha256:89c34c6a59e9c0760c8ecf5e1674f454d73b27b4bc85e9930b72117326c7548e

Observation f7f8dd81-8c56-470d-a611-dedaaa8bd49e · inbound

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models cites this paper.

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T04:50:59.096166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:50:59.096166Z digest=sha256:b194096c21ed0c3d04be4f8a47764d085fee8414a8cbd892041c45bb6a803078

Observation c22d914d-ad88-437d-8db1-25ec4149b7b7 · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 122

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:08.507221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:08.507221Z digest=sha256:3607ad84761fa17f98f862f79169349733b46436ec9d0afef552cbea428f5d23