Pith. sign in

Paper Citation Record · LEDGER

A Simple LLM Framework for Long-Range Video Question-Answering

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2312.17235.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.17235 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:53.167691Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:03.011665Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation df1358d4-cea1-4665-8cf1-34ee1e4e3d41 · inbound

PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning cites this paper.

PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:21:58.044985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:21:57.873354Z digest=sha256:6d8eeb63dac9bc0666a81ec0ec4fa37d2b36ec281cfd2980b8b051447ce35b3a

Observation 5fbec168-da9c-4c0f-8ff8-5d4de93b152a · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output A Simple LLM Framework for Long-Range Video Question-Answering

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.857614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:ee6d1d7445c395ca45899c61e2d2f73facc398d6bd4a18597f7ef0725d4475be

Observation d88c3ee7-3d28-4528-8fc2-6a4f70996126 · inbound

Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles cites this paper.

Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles A Simple LLM Framework for Long-Range Video Question-Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:53.167691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:53.167691Z digest=sha256:81f85f28788138d863057cdc8f7503ac552c45c0cd130f6ccf52de8b1923ff26

Observation 89b6721f-3457-49b3-88e1-24ab7fba52b6 · inbound

HCQA-1.5 @ Ego4D EgoSchema Challenge 2025 cites this paper.

HCQA-1.5 @ Ego4D EgoSchema Challenge 2025 A Simple LLM Framework for Long-Range Video Question-Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:47.594214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:47.594214Z digest=sha256:00ca591838d1ee374fe9ac991ed21c9192dcc1ed0c54add99af5c9553c0c9049

Observation 0cd89e7c-11c0-45cd-8bc6-4f53a320ddf1 · inbound

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering cites this paper.

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:29:26.732647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:29:26.732647Z digest=sha256:c440fb48551e104826fd16c8371262eb7cb3332ee458b321d0caa18585c5840f

Observation d6452b5a-89f2-4f23-9ca3-fe3a857482c8 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding A Simple LLM Framework for Long-Range Video Question-Answering

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:56.959356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:56.959356Z digest=sha256:9fa04c52646e1426a5f4c0b28b6a7e8722c32e9268481df738d63a4cb24aba20

Observation cd904f4d-844d-430f-a3a9-23449a5f5372 · inbound

LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering cites this paper.

LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:36.153446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:36.153446Z digest=sha256:6ef95eab8a0f8b1a2cf81229398f052e857f4f9a5cc62c7ff82dec77ccea1449

Observation 34e11361-29a3-4833-90e7-58ddcf13c67e · inbound

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering cites this paper.

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T04:49:42.312571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:49:42.312571Z digest=sha256:a729d60cd9e39ca1cd36e42328456448134b84f6630116f600d63194ae68e768

Observation 4c6df0f6-e239-4eee-b50c-4bf0e9b1d78f · inbound

CAViAR: Critic-Augmented Video Agentic Reasoning cites this paper.

CAViAR: Critic-Augmented Video Agentic Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T21:27:33.767446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:27:33.767446Z digest=sha256:6cbbedcf2196e006a4587e2edd40a89747a6958ed7661f0e8f9107b7a1487a85

Observation cea25f6f-a7db-49e6-bb95-189a99b560d1 · inbound

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents cites this paper.

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents A Simple LLM Framework for Long-Range Video Question-Answering

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.621623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T12:40:19.544260Z digest=sha256:6157c90893265ba64d6d5962450391dc6e2656970ec8e7add3fcb4fa4309d945

Observation b693d7da-96f3-4119-aeba-636ba2d90386 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.682531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:46d50ae1e4a98a0c9a859cafe645f29ee24528713a0aa9d74f4f13a2ae6b3b7b

Observation 39f82adb-aefa-44aa-b2f5-9b44fd964678 · inbound

Towards Sparse Video Understanding and Reasoning cites this paper.

Towards Sparse Video Understanding and Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T23:33:16.311655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:33:16.311655Z digest=sha256:0a1b45f20c3389ea0ff431d86de4fe46e6ada70ff1fef5d9becd1c3c347a6486

Observation 290ab08c-93d4-4de8-85c2-b2ab3dc664c7 · inbound

Progressive Video Condensation with MLLM Agent for Long-form Video Understanding cites this paper.

Progressive Video Condensation with MLLM Agent for Long-form Video Understanding A Simple LLM Framework for Long-Range Video Question-Answering

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:43:14.889680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:40:41.380829Z digest=sha256:c8a361e9ecb4cc223f928938bc287e2c9c05b409b077cdae370314376088c668

Observation 488288cd-eb57-4fc0-a317-92bbea7afcb2 · inbound

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks cites this paper.

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks A Simple LLM Framework for Long-Range Video Question-Answering

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:13.224993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T08:41:42.061219Z digest=sha256:056cc2813adfae0a5770e04c64e668bb65f3eabdb99faa0f146c080f7f5d30dc

Observation 058c212c-0e18-4f04-9a59-49c33874108f · inbound

UNIVID: Unified Vision-Language Model for Video Moderation cites this paper.

UNIVID: Unified Vision-Language Model for Video Moderation A Simple LLM Framework for Long-Range Video Question-Answering

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:07:09.283746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:56:26.674841Z digest=sha256:2b991c4d8635a6ff16a574021a45d91fa746d18a4cabb660cd0dda48b077590d

Observation 4760a439-14da-41c5-b5d7-935380a77d67 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 170

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.012988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:b3254b289d8208fc7e51915792a59a0b14ca681878ea877a3be5f4deaf43aad1

Observation 5edf0bcf-564c-4d65-9551-ccf929071349 · inbound

Agent-Computer Observation Interfaces Enable Dynamic Computer Use cites this paper.

Agent-Computer Observation Interfaces Enable Dynamic Computer Use A Simple LLM Framework for Long-Range Video Question-Answering

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.233670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T06:59:20.295818Z digest=sha256:a5ef6fb5d07ac96c503974c729826783aa290827daa3aa096bfece733d8954e4