Pith. sign in

Paper Citation Record · LEDGER

VTimeLLM: Empower LLM to Grasp Video Moments

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2311.18445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.18445 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:06:35.130897Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:45:34.766676Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c34e4c35-76cd-442a-8933-209119e7c252 · inbound

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs cites this paper.

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs VTimeLLM: Empower LLM to Grasp Video Moments

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:44:53.448438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:44:53.284345Z digest=sha256:68d4703fb249914c15d047e537f7f9c4998858826207c570024c2071a51b3da2

Observation e2794c98-6dde-432a-abc5-45cdba567235 · inbound

PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance cites this paper.

PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance VTimeLLM: Empower LLM to Grasp Video Moments

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:33:15.634988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T17:31:59.030963Z digest=sha256:0b363dbf95e4a574b8be1ae0cdade180cec11dde3c9e9c65095db7d40e2b7ec8

Observation 76ce522a-51f9-4e33-8d59-0eba8a33ac6e · inbound

VideoRoPE: What Makes for Good Video Rotary Position Embedding? cites this paper.

VideoRoPE: What Makes for Good Video Rotary Position Embedding? VTimeLLM: Empower LLM to Grasp Video Moments

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T20:06:35.130897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:06:35.130897Z digest=sha256:8e5100e5afb9867fc214a7e54b34a4a83e638f7c36d0ad087d21dcd7b8b2e5f5

Observation 33e18d71-395a-4d2a-886f-d54016a41e1b · inbound

PDB-Eval: An Evaluation of Large Multimodal Models for Description and Explanation of Personalized Driving Behavior cites this paper.

PDB-Eval: An Evaluation of Large Multimodal Models for Description and Explanation of Personalized Driving Behavior VTimeLLM: Empower LLM to Grasp Video Moments

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:36:55.781535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:36:55.781535Z digest=sha256:a285cc4717362546b03d89b595bfaf6b758b2a80d1379b4416dd825e3e0c07bc

Observation 08bdc58b-990a-4b98-8c63-4042fc607a58 · inbound

A Paradigm Shift: Fully End-to-End Training for Temporal Sentence Grounding in Videos cites this paper.

A Paradigm Shift: Fully End-to-End Training for Temporal Sentence Grounding in Videos VTimeLLM: Empower LLM to Grasp Video Moments

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:48:11.527741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T19:43:57.298338Z digest=sha256:1c7c85c19efa0307a99fa88cfb636acf585858c2f81fc631a4d143f13851095f

Observation 755d4af9-53eb-438f-8fe2-696c6546b606 · inbound

MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding cites this paper.

MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding VTimeLLM: Empower LLM to Grasp Video Moments

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:46:13.688106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T14:10:27.416341Z digest=sha256:d379aef290045d80191d902b8ab1d78e63f698f3e9200ca4bfd030870e3b3449

Observation b8b842b8-d96b-4015-8c38-0b713131a6bb · inbound

MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding cites this paper.

MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding VTimeLLM: Empower LLM to Grasp Video Moments

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:45:34.768754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T08:41:31.700367Z digest=sha256:a12d6b0e4e16a6bae5fac18de601318b3ad5c68c2116cb9f38b62074102bdc21

Observation 28687dfa-c2a3-48e0-af69-1c6b34771ef1 · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding VTimeLLM: Empower LLM to Grasp Video Moments

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:50.295963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:50.295963Z digest=sha256:1113da8917727955189998eae7d7ae413e09fedee5725788f172b149467dba64

Observation b4672433-9574-44ed-8088-14daf013796f · inbound

TimePLE: Rethinking Temporal Representation for Video Temporal Grounding cites this paper.

TimePLE: Rethinking Temporal Representation for Video Temporal Grounding VTimeLLM: Empower LLM to Grasp Video Moments

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T23:32:51.395681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:32:51.395681Z digest=sha256:0d3391c35aadde1b64631949749c7271db6e7e458087272b1b4a8692fe27f623