Pith. sign in

Paper Citation Record · LEDGER

Lost in Time: A New Temporal Benchmark for VideoLLMs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2410.07752.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.07752 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:26:20.299457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T19:07:17.473192Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0f9ee443-d1d2-492a-bd74-25d790104880 · inbound

HD-EPIC: A Highly-Detailed Egocentric Video Dataset cites this paper.

HD-EPIC: A Highly-Detailed Egocentric Video Dataset Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T23:26:20.299457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:26:20.299457Z digest=sha256:0c43b71d213b8138d7bc475fb822a8f84d3906729d52005367fa3fae3b48ce03

Observation 769e071d-a57c-4a42-a863-a23d22964c86 · inbound

Seed1.5-VL Technical Report cites this paper.

Seed1.5-VL Technical Report Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:26:06.098537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T05:26:04.960844Z digest=sha256:d4354505e13fcb68626808eb1b8c94079c0b053e99c240c8e1d6e16ca5975f3d

Observation ec8e0ee3-de3c-4682-8d6c-becc925384ba · inbound

Fostering Video Reasoning via Next-Event Prediction cites this paper.

Fostering Video Reasoning via Next-Event Prediction Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:40.921513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:40.921513Z digest=sha256:10a4e176e7227e3b8c826080506e271004655df5bd4cbfdb1817ce9b8d93fc25

Observation 0773ce32-d25b-4c9a-a11e-7d9cc1e2eb03 · inbound

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models cites this paper.

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:34.028134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:34.028134Z digest=sha256:69c6a45c34f6cc44b13b87a9a307f8a966f0fbbeb7dd6e6bf6e7438cb9f2ac59

Observation 03cff194-f5a0-4074-90c0-a2701db95f6b · inbound

How Important are Videos for Training Video LLMs? cites this paper.

How Important are Videos for Training Video LLMs? Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:07.790019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:07.790019Z digest=sha256:4ba97b0c0395bc7982dd64aa197169915a56b5436fc2df18762324f6eea0ad64

Observation 2ddcf280-f912-4b04-a12e-5e84b7c485e2 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:33:50.755434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:9f2ae03be8bbc3906072ccb6a438c5a311d62c90a1864014d0deffcb8bc6ace4

Observation 3f7009cb-616d-44a6-bd44-2b6e59c6e59e · inbound

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos cites this paper.

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:22:55.823070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:22:55.823070Z digest=sha256:cbe83aa0d387beba36ed404fdb40cc5e482717fee3ff08535098d8cd6d0fc558

Observation 0b62f425-e764-4385-a3c0-7c7e35137384 · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:56.477678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:56.477678Z digest=sha256:f54172c157ff3e871608a5def82e4c29d24565b32ca1314bdda570d526e7fca1

Observation 8bda65d5-9efd-44e8-89ea-24f8062f3b04 · inbound

EgoExo-Con: Exploring View-Invariant Video Temporal Understanding cites this paper.

EgoExo-Con: Exploring View-Invariant Video Temporal Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T07:23:03.592902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:23:03.592902Z digest=sha256:4c16c5d47e78fa9cdce82174992032d8d6a9e8aefa79a40b55bc8d4ce9c250a0

Observation 916a81a1-e40b-4e6d-a595-405f1fcbb1e5 · inbound

Adapting MLLMs for Nuanced Video Retrieval cites this paper.

Adapting MLLMs for Nuanced Video Retrieval Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:21:18.917253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:20:09.051957Z digest=sha256:1dd034bf61725ac0ad69612bde95704b9991ca7009becc1fbe951b77f5d7bc89

Observation 8e55d311-7132-4f32-8d5c-1edf6eb53b06 · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.381093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:b98fdf77d7f0a3fb6c3159a27950ba1106d936d6b9efe6e1c6f012ce04135078

Observation 97dcf00a-7928-4d4c-98d7-667d7095f848 · inbound

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging cites this paper.

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:06.215848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:42:43.948462Z digest=sha256:1a981865381c8e8eb84669c3e400d8ed29716b1b5f87c6db58269ad78b76799f

Observation 2c98fe22-2bb7-43dc-9a7b-2927aa9238db · inbound

PushupBench: Your VLM is not good at counting pushups cites this paper.

PushupBench: Your VLM is not good at counting pushups Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:08.857453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T08:32:47.354181Z digest=sha256:72684afb5959688671444462337cba54a8f93d2be4615c5781facb2ef8802165

Observation 8fa4fcd6-dd2b-4686-abb4-51a589436198 · inbound

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs cites this paper.

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:53.171872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T03:04:27.841522Z digest=sha256:18c7fb7ebdd5ed0077c70c7cf9385a30d859d6bea7f4de2eadf292394a0479fe

Observation eb6ab733-4622-4b3b-85ec-36cb1c12f2cc · inbound

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding cites this paper.

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T16:03:08.026421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T16:02:53.887605Z digest=sha256:bbe01365c5068e424655264c972a59c028afbeb34ed306b56e4df3721fa9de57

Observation e8bd32cd-f16d-464d-ac4b-2b0408354d5c · inbound

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning cites this paper.

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:46:14.915462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T07:45:56.473188Z digest=sha256:b79ff5931d467a4df2620c042fbcbf12a186d71f26b4bd8e5c081f87d5fa7190

Observation b9223b89-0eee-487b-a66d-6691a801b119 · inbound

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis cites this paper.

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:14:42.891160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T07:12:02.612292Z digest=sha256:0a750b2ae411ea7d4b923779a487123afa444c9215a74096396e5ebf3e793de5

Observation 54d8863a-6e9a-4d9b-94d0-2e2186023ae7 · inbound

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs cites this paper.

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:44:38.756827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T05:41:39.396469Z digest=sha256:9d6ae4510f3cdb54e9dfacc230ec6b939922330f7edefc10742918fcb335cbba

Observation 41b3e6fe-42de-40a4-820c-17a67258d8c3 · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.474609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:dc711488f73b9eef35541a88286de06d4beaf6407d1be199632c77c5acb24d1c

Observation 00b81ec2-478a-4f2e-81e9-cce4e09940af · inbound

Do Video-LLMs Actually Watch? Diagnosing Character-Tracking Failures in Long-Form Video cites this paper.

Do Video-LLMs Actually Watch? Diagnosing Character-Tracking Failures in Long-Form Video Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T07:12:52.433947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:12:52.433947Z digest=sha256:bdff96e17fe77b5b6b728737059d85b71c7f8ab697f9ad4b33de77ecf4f6bf6b

Observation 9bddb493-8b2a-4d5b-9e5e-2f5ace7445cc · inbound

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks cites this paper.

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T05:39:58.511387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:39:58.511387Z digest=sha256:30c3fe4fe6c9b90740559b79467c1bd61d510886b8c845b9431fa9aaaeec0a34

Observation 8e3ba013-2d83-496e-9ee5-82869176d604 · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:46.152973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:46.152973Z digest=sha256:fe235462a27c937c345cce86792ad98cda185a6c17dc8b854633b09d5cd7c0b4

Observation 24a8a509-1e7d-4172-b845-151ab4876e45 · inbound

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning cites this paper.

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-04T17:24:16.126597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:24:16.126597Z digest=sha256:735f5256efaa69e3f8632ee61dc090a9a3b36a29b0344930b8e5514305f333c7