Pith. sign in

Paper Citation Record · LEDGER

MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.17176.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.17176 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:56.970276Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:38:56.117133Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c968ea4c-2577-487c-92db-2822ddbffd9b · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.683009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:03e0ec7749419c13713cad69fceb2014192c5d7edf3a2b103145155f05c421de

Observation 0e915a04-0f0e-4302-bbfd-d9e68fd29c5d · inbound

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks cites this paper.

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:05:29.177851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T07:05:08.716223Z digest=sha256:a5e7f5f984c1e1290583823c8f7303e2bfc8d616596bebadda72ab5fde33d237

Observation a4bee17c-1f8f-44a3-8887-3919a4db3b45 · inbound

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding cites this paper.

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:20:00.203242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:19:59.603343Z digest=sha256:c4af1673562206048dae46afc117d1ee4e852578f9bf0182621e11af5af619e4

Observation 0b115c45-2019-43e6-ad60-3a1f2add900f · inbound

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval cites this paper.

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:31:40.772346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T14:26:59.015559Z digest=sha256:497c3155d8c7b05b0dceea868bff36847917de2708920bf0ca0cfcf5b40452e7

Observation a72b4ca9-7698-4895-9ef0-3b6547706f43 · inbound

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding cites this paper.

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:56.970276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:56.970276Z digest=sha256:3cf4547aeea490ac312ea93a2221fac6789b0d62e6424989b21a6b31ec3cae48

Observation 89b8f273-1564-497c-bb65-37c560ab0e08 · inbound

MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding cites this paper.

MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:00:07.170081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:00:07.170081Z digest=sha256:a685fd933cca243295a76129336c8ff4e3dbe204268153978de7d6a96d6c317c

Observation c4ebeb15-4443-4c06-9dcd-0a56d5531c28 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:54.353632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:54.353632Z digest=sha256:8b61936bdedf59eed998ce52c13825371db5d4ee914b71915666a9655fc709a0

Observation bcea2d96-461a-428d-bce5-d7f7a744bc27 · inbound

InterAct-Video: Reasoning-Rich Video QA for Urban Traffic cites this paper.

InterAct-Video: Reasoning-Rich Video QA for Urban Traffic MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:52.956161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:53:52.956161Z digest=sha256:905ae868640baa15c265015d7a4ee974ef39ea49e28f90bc14a64970f16152d1

Observation 46d85ce4-206a-4de0-b05c-25ed74047074 · inbound

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding cites this paper.

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:46:46.573054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:46:46.573054Z digest=sha256:f18fb08ebb91ee17dfdf8c3e9f174a38acea346128da47f0c170f8de8aeba33c

Observation 5f6696c8-d6b7-40a4-a677-b56b1c03cd55 · inbound

AdsQA: Towards Advertisement Video Understanding cites this paper.

AdsQA: Towards Advertisement Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:20:36.871295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:20:36.871295Z digest=sha256:731a319f4b917fc9ec41f2e60d21e50a098e1ccc8b34c8f79ca9849c390de33d

Observation f5be6b5e-1180-4d3a-8fdf-bff484473869 · inbound

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding cites this paper.

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.889758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T22:11:01.690237Z digest=sha256:c3064eb7e505d2804a0a0c74da190954b75c44ac826c29e769d6c2370f390897

Observation b17d68cf-e15c-4649-9cce-ea65ece727a3 · inbound

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams cites this paper.

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:56.118616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T01:12:46.295455Z digest=sha256:eb0f5cc6a875b25b09c2c3079d58acbfa45abe4db2efb61933693ac45384e5f8