Pith. sign in

Paper Citation Record · LEDGER

Training a Large Video Model on a Single Machine in a Day

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2309.16669.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.16669 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:22:25.180917Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:26:05.896375Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2f028e33-ae02-4d0b-8d47-ea889ce8eda2 · inbound

Extending Video Masked Autoencoders to 128 frames cites this paper.

Extending Video Masked Autoencoders to 128 frames Training a Large Video Model on a Single Machine in a Day

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T16:22:25.180917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:22:25.180917Z digest=sha256:abbcc3e6ad83d4d88e97c465cc324a3f61b577384edfcf5cc98d41d12c935310

Observation a7309c89-6de1-4361-8b8b-09e82da33cdd · inbound

Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model cites this paper.

Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model Training a Large Video Model on a Single Machine in a Day

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:14.736293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:14.736293Z digest=sha256:9da821dac659994747a210e8213b2839b8aa6778b9394c3d05d75aed39cf193f

Observation a7b81eb3-a405-45a5-afa6-8dfe0d3a0b67 · inbound

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models cites this paper.

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models Training a Large Video Model on a Single Machine in a Day

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:35.247796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:35.247796Z digest=sha256:bf9a86b783b34fb0c4cf58530a70c5a3249219357cab0b695783621340d8f3fa

Observation 416ae5dc-9eea-4ae5-8c37-38a17c72361d · inbound

EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization cites this paper.

EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization Training a Large Video Model on a Single Machine in a Day

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:26:05.951232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:25:59.818943Z digest=sha256:eeb73ae9b9504530b1f390428d9b3a48d22c7cb245c40472e7903d84cdf74e72