Pith. sign in

Paper Citation Record · LEDGER

Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2206.11795.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.11795 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:46:06.424312Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

50
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 986c05f3-628c-4994-b984-5400d767e7ae · inbound

A Generalist Agent cites this paper.

A Generalist Agent Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:24:49.880515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T06:24:49.833638Z digest=sha256:03684fe09a04f0945f5a62d7deb6518544c17e9fbcdd6ff952b8d1f3084e6399

Observation 277ff47c-34e9-4d15-8d37-cc93fd2ce4df · inbound

Mastering Diverse Domains through World Models cites this paper.

Mastering Diverse Domains through World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:08:22.082585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T09:08:21.677362Z digest=sha256:bcf315aa1c6dfe291fc62d480e1accfde341c7f0867f4e9d75e8355fe98cc049

Observation ef1029a3-d927-436a-bacf-a94fcdbd3cdf · inbound

Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents cites this paper.

Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T03:27:40.743136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T03:27:40.524895Z digest=sha256:d890a8f7665adcfc6cee9f910cf2f6ae23c0b780729ac0621fd45a2599c4d0c9

Observation f8c883c3-b099-4fcc-95d4-879a812a740b · inbound

Voyager: An Open-Ended Embodied Agent with Large Language Models cites this paper.

Voyager: An Open-Ended Embodied Agent with Large Language Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:11:41.059888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:11:40.995345Z digest=sha256:d6d331174c6b948ee012da13fd07168b5f9083120f2614a9e1cf920a7f0a7bea

Observation fce01e15-727f-481a-a06c-901bd4b06853 · inbound

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution cites this paper.

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:12:31.274505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T08:12:30.984870Z digest=sha256:917b9342382e7b08ddf920846b5201f94e5e578c2cfae6223d4a51c95896e00e

Observation 617edcce-b5d0-4431-ab19-6ed9ae3abe8b · inbound

Learning Interactive Real-World Simulators cites this paper.

Learning Interactive Real-World Simulators Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 219

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:15:18.445102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T02:15:18.265190Z digest=sha256:45b7ef732f2012384f8c0d4f708cbfb0faff30290f529a99319ee90695b09af6

Observation 676efef6-edf1-49f2-ac33-e253b0d5d32a · inbound

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies cites this paper.

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:22:44.385372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T08:20:05.898025Z digest=sha256:b270704e0a1b42c4f3f25337fabf04819ed081c7c13c886308d927b7d9dc10cf

Observation 45edeb9c-7083-458f-8200-889add293bd6 · inbound

DreamGen: Unlocking Generalization in Robot Learning through Video World Models cites this paper.

DreamGen: Unlocking Generalization in Robot Learning through Video World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:50:45.650952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T23:50:45.332466Z digest=sha256:3e5d96e1bb16b30aa382f9cd89c9296d07d7e6c33b6233db8e3f0c0ac2c194d3

Observation dcf482dd-1dd7-426a-ae45-b605024bf623 · inbound

Scalable Multi-Task Reinforcement Learning for Generalizable Spatial Intelligence in Visuomotor Agents cites this paper.

Scalable Multi-Task Reinforcement Learning for Generalizable Spatial Intelligence in Visuomotor Agents Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:06.424312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:46:06.424312Z digest=sha256:10531003cd44cf9893bb9c611a866c8fad8eac35bd1637f697042f9c2e5a3d4e

Observation 6737d5a0-3899-4a16-84f7-71eba7eb3c2a · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:45.747217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:45.747217Z digest=sha256:cf7127ad3092e359d0d2160de3256744e8947f256844e983bcd45ff4c2d04398

Observation 49b175a3-22db-4fad-a340-e10665f6aa08 · inbound

PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments cites this paper.

PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:58:44.340988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:58:44.340988Z digest=sha256:d5873d99575f4d8ee09d36934a237f02959faf20d30a2f89bb8c5396b5656ca1

Observation a9059ad9-2301-4408-90f9-24b24924b0ef · inbound

CA2: Code-Aware Agent for Automated Game Testing cites this paper.

CA2: Code-Aware Agent for Automated Game Testing Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:04.989946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T05:53:26.558042Z digest=sha256:bbeee40ea8e20b562921bce9783ccd8c3bf6632ee6409868359865e05d613146

Observation 9ce58f77-7b49-406b-897e-8d4da5292e23 · inbound

ASH: Agents that Self-Hone via Embodied Learning cites this paper.

ASH: Agents that Self-Hone via Embodied Learning Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.482966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:51:53.328977Z digest=sha256:7c4b4b917e3d44a8b733abf0abdf6c751601e17f8ce89ebd575c4690e1e5a13e

Observation cdb9c999-d720-4bd8-8ec8-da56f66bc0d6 · inbound

ASH: Agents that Self-Hone via Embodied Learning cites this paper.

ASH: Agents that Self-Hone via Embodied Learning Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.116021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:12:54.453886Z digest=sha256:9622b9e7e0bc078ac09fc1c710eabdc0f3ac9d4aa67537bf1d385df4ed9acaa0

Observation bc03e82d-e8dc-42d2-879a-8df64c52aee6 · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:33:04.182122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T05:28:50.354662Z digest=sha256:683d6bae0bf6f21f55be57c3d936bf0aeff2477152c37b4960702326825ac6a1

Observation 302b840e-9cec-4864-b8a9-e6955826f61d · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:39:49.272258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T07:36:12.214949Z digest=sha256:2cf444396ff2f5993e8c1a6c9390c655217c7807cd62e832e265eb1499b754cf

Observation d357f8c1-c2f0-4f40-8fb9-2991c3e585e8 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.589878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-25T19:12:22.513577Z digest=sha256:0c473dba3b88ffdb01eed2c644cb16c78ac8b7bba4984dc1a5ae838ec106a1ca

Observation dda212b4-94fa-47a6-8991-c5c1ed9fd39d · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.595338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:11:07.089829Z digest=sha256:774cd08e3d78379a4f7c884966711b43f0ad2bdfd54b177dfe22bb67a1ab46aa

Observation de3919ae-247d-4613-8294-24c54513fe49 · inbound

In-Context World Modeling for Robotic Control cites this paper.

In-Context World Modeling for Robotic Control Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T12:05:57.682386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:05:57.682386Z digest=sha256:3d0cefb2e533bc293aea9c5bd687b6b566a5384660744d4912838ba0b275e4c8

Observation 1b6a113c-79dd-44de-884c-bd4feec47771 · inbound

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models cites this paper.

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:53.233336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:46:21.198781Z digest=sha256:db5b837363c34b2a0773ac52e9a149b489c86835615e9661d6cf1cd78ad4c7cd

Observation 1801af48-6d70-4ac6-95db-e513da3ce0c3 · inbound

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models cites this paper.

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:04:39.292257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T10:19:06.268547Z digest=sha256:41a48db99a21d9103b2ffb8dc2ca7a16fdd7cd267af29713c1cba7eed8b9eb73

Observation f7f8dd81-8c56-470d-a611-dedaaa8bd49e · inbound

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models cites this paper.

Causally Debiased Latent Action Model for Embodied Action Conditioned World Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T04:50:59.096166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:50:59.096166Z digest=sha256:b194096c21ed0c3d04be4f8a47764d085fee8414a8cbd892041c45bb6a803078

Observation c22d914d-ad88-437d-8db1-25ec4149b7b7 · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Reference 122

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:08.507221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:08.507221Z digest=sha256:3607ad84761fa17f98f862f79169349733b46436ec9d0afef552cbea428f5d23