Pith. sign in

Paper Citation Record · LEDGER

QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2107.09609.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2107.09609 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:29:22.165769Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T14:10:15.171643Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 66f1ab7a-3e27-4924-a21c-034eb1d2399d · inbound

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering cites this paper.

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:29:22.165769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:29:22.165769Z digest=sha256:ccfdc407a86d9e4be03595d35a86c8330a717b1babc7b43324eb881b086de521

Observation 9edbfe3d-8028-429e-842b-6423ac9b5069 · inbound

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents cites this paper.

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:10:15.175406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T14:10:14.929207Z digest=sha256:52f848ed3de4e791b426d639750997e00a74e10faae199e771f097c5960f3969

Observation 277290e8-e626-4bc7-96ec-cd21110b8e3a · inbound

Tempo-R0: A Video-MLLM for Temporal Video Grounding through Efficient Temporal Sensing Reinforcement Learning cites this paper.

Tempo-R0: A Video-MLLM for Temporal Video Grounding through Efficient Temporal Sensing Reinforcement Learning QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:47:52.690022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:47:52.690022Z digest=sha256:4feeed3a1ceafb8d981784a848566da7f53a8ff6d569cda05a23c10fa5a4e6da

Observation 7771e64a-d31b-4f33-9b37-3a799931a030 · inbound

The Repeated-Stimulus Confound in Electroencephalography cites this paper.

The Repeated-Stimulus Confound in Electroencephalography QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:08:00.723791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:08:00.723791Z digest=sha256:85ea185e7a8cb837e90ba8ee11e91b1aa742f1eccd42fb6b6105250c48ef1e73

Observation 0ad83622-e530-41cc-a48f-e15a1d1481ef · inbound

SoccerHigh: A Benchmark Dataset for Automatic Soccer Video Summarization cites this paper.

SoccerHigh: A Benchmark Dataset for Automatic Soccer Video Summarization QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T12:34:38.941509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:34:38.941509Z digest=sha256:0493eeb33069f3e2e87176c4553f5f64247e3f83be0b274bd1d5f8ed2f5b9d2b

Observation 6f8f3397-a73b-4a74-8bc3-4435e3ae69e5 · inbound

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models cites this paper.

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:01:11.429339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T07:28:09.725591Z digest=sha256:82881577d350d34e11192b894f113026a50c9c6c81912bb2a963589fdbaf9b5c

Observation a9735ae1-a460-4008-b20c-7a75cc8a287e · inbound

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models cites this paper.

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T15:38:00.249144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:38:00.249144Z digest=sha256:c4f5bfb8c82c659f2f3e3b465f13b18d21aba33d9d275f43b7953e7bc32c37ef