Pith. sign in

Paper Citation Record · LEDGER

Training a Large Video Model on a Single Machine in a Day

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2309.16669.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.16669 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:22:25.180917Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:26:05.896375Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2f028e33-ae02-4d0b-8d47-ea889ce8eda2 · inbound

Extending Video Masked Autoencoders to 128 frames cites this paper.

Extending Video Masked Autoencoders to 128 frames Training a Large Video Model on a Single Machine in a Day

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T16:22:25.180917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:22:25.180917Z digest=sha256:7319eb5e7b82053fd2dab378c29e480c34fee95ef6329b23ca7a4235e6b06596

Observation a7309c89-6de1-4361-8b8b-09e82da33cdd · inbound

Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model cites this paper.

Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model Training a Large Video Model on a Single Machine in a Day

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:14.736293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:14.736293Z digest=sha256:306ede612b9c29ac1eb310ba0164dc0277ddc11f958b8c55f07879bd449ca683

Observation a7b81eb3-a405-45a5-afa6-8dfe0d3a0b67 · inbound

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models cites this paper.

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models Training a Large Video Model on a Single Machine in a Day

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:35.247796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:35.247796Z digest=sha256:fe9e5b2b2c756a4bcee1d1117d32896c7dc15cfccc0978081c9166aac4b424a0

Observation 416ae5dc-9eea-4ae5-8c37-38a17c72361d · inbound

EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization cites this paper.

EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization Training a Large Video Model on a Single Machine in a Day

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:26:05.951232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:25:59.818943Z digest=sha256:0a3bce1c28272fa2e8c4417814feeb7f1e1ec83652c4f2945017cba6c8bf95b5