Pith. sign in

Paper Citation Record · LEDGER

EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2411.08380.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08380 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:56:40.227298Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:09:35.541046Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eba3de72-db0f-477b-ac34-c9d316dc5fd1 · inbound

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning cites this paper.

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:40.227298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:56:40.227298Z digest=sha256:cfc2d32cef9888b1620c63f7dd41634f9291671c7e03598a0606528616d33740

Observation 259f6f08-efe0-4d4b-ac2c-913566405df4 · inbound

Seeing in the Dark: Benchmarking Egocentric 3D Vision with the Oxford Day-and-Night Dataset cites this paper.

Seeing in the Dark: Benchmarking Egocentric 3D Vision with the Oxford Day-and-Night Dataset EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T10:50:51.168830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:50:51.168830Z digest=sha256:a75e2ce6a30183aa85c33aca79bed405927ec5b7b1faaeee2ac51da9e1df0bcf

Observation d96b73dd-335a-4a97-8689-3d79f91c0368 · inbound

EgoM2P: Egocentric Multimodal Multitask Pretraining cites this paper.

EgoM2P: Egocentric Multimodal Multitask Pretraining EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-07T05:31:42.617889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:31:42.617889Z digest=sha256:eef9dd597895366434cde7f45b66001a49a108ae0ebbf012939f7024de52d46d

Observation e2e62075-90b9-442d-8a0f-54f627e3b7a1 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:17.840081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:17.840081Z digest=sha256:c7bbaa81cfd822f51a4db7cac5dc554cc4dacc06a2c3b793319a515da46151a0

Observation ec597a67-6a35-4e64-a58c-56dfcbe03a3f · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:16.373183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:dfbf1656ff6551e8c1684f8dff814c04d60e35549f2209e937bd53537bf7362f

Observation 8d4ee067-2a01-470c-a0e8-d8cbe31222ec · inbound

Precise Action-to-Video Generation Through Visual Action Prompts cites this paper.

Precise Action-to-Video Generation Through Visual Action Prompts EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T19:15:34.292314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:15:34.292314Z digest=sha256:247c284503f2e472b2e5b2ed232c3382194423ea70f5bfac6c2c6642921f73de

Observation be883004-0e67-4bd2-a90c-745b6dfc6a8e · inbound

SFHand: Learning Embodied Manipulation by Streaming Egocentric 3D Hand Forecasting cites this paper.

SFHand: Learning Embodied Manipulation by Streaming Egocentric 3D Hand Forecasting EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:54:18.270974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T17:53:19.002603Z digest=sha256:7e5d78f24ed94833b0fb4bd1ea8eb4727bea40d9fccba2318b5514eb1ca71575

Observation 67625f16-802b-41f8-acc7-3a7e45e13f3c · inbound

EgoSim: Egocentric World Simulator for Embodied Interaction Generation cites this paper.

EgoSim: Egocentric World Simulator for Embodied Interaction Generation EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T14:40:24.818842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:40:24.818842Z digest=sha256:ec3be178f313845dffb3e24d7264ac5e9401e88642c8bfdf75cbb6de2f0b0820

Observation ff58e91a-20bb-474b-8274-95c618a83ef0 · inbound

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks cites this paper.

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:02.978634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:10:36.209152Z digest=sha256:938294471a7d514f04195dab21ab419f7c6409905859688b298be7ff46ace8e9

Observation 57fcf62c-8edd-4aee-8b9b-a8a04ea19581 · inbound

Robot Learning from Human Videos: A Survey cites this paper.

Robot Learning from Human Videos: A Survey EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:41:29.720260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T04:55:44.273643Z digest=sha256:b26a50d048c67408bf7f8755f8eb5bc9ee41cb0bf5c544ccde4fc82ec56e24b7

Observation f066cf9c-ec7e-4459-8a22-426517529e66 · inbound

VISTA: A Controllable Platform for Generating and Auditing Egocentric Assistance Scenarios cites this paper.

VISTA: A Controllable Platform for Generating and Auditing Egocentric Assistance Scenarios EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T14:24:57.896385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:24:57.896385Z digest=sha256:daec8e3037fbd58c75b1785ace53223d9ae6496c5b023b60932e9fc385ce3756

Observation 9a614153-e28a-419a-b3dc-dbd1de4d3215 · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 183

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:07:18.085461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:331f93a36537a971790fb68901a5fa6e4bac373da2c1edd1a22660ad95b9bcc1

Observation 03223971-dea8-4844-936a-34fb0f43d6ea · inbound

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control cites this paper.

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:34:02.491593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:24:38.389294Z digest=sha256:f70b6d34b8b56b43f7e3818902f94f18b00155ddae0b99468af4fe0417c1d505

Observation f7401e87-89f8-4008-8fab-a9f396fbf08a · inbound

World Action Models: A Survey cites this paper.

World Action Models: A Survey EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 163

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:09:35.543085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:11:12.686936Z digest=sha256:b1b48882027fcd789a0a312f9c982b6c2a8bf15d1f4153e834d9eb58ab0e6696

Observation 2d0fbbc3-546e-4783-98f4-cc931057e2bc · inbound

HandsOnWorld: Unconstrained Egocentric Video Generation with Camera-Disentangled Hand Control cites this paper.

HandsOnWorld: Unconstrained Egocentric Video Generation with Camera-Disentangled Hand Control EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:58:37.380383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T15:52:18.371702Z digest=sha256:322d5c37571592b0bc88d52f578e54602ef3183f0899090ee2d569859c10cb86

Observation 58d28c96-4fad-4bb4-b6a7-adabeebe6cf0 · inbound

iFLYTEK-Embodied-Omni Technical Report cites this paper.

iFLYTEK-Embodied-Omni Technical Report EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T12:21:16.963000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:21:16.963000Z digest=sha256:e4664db2686c473be878c96a91276e3e95deb1c919269b9a8c35c9be31a94eb2

Observation 3f261905-45cb-4fb0-b2d3-10ff1f52fe30 · inbound

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding cites this paper.

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T14:08:29.408596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:08:29.408596Z digest=sha256:6f1f1a41023a85da2378248ac223a3875822a45c4b17d7e596fde5ab7ce550b6