Pith. sign in

Paper Citation Record · LEDGER

Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2501.13468.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.13468 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:13.898584Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:19:29.878882Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1812f63e-756e-42b0-9447-5e044e886567 · inbound

Seed1.5-VL Technical Report cites this paper.

Seed1.5-VL Technical Report Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 154

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:26:05.736163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T05:26:04.960844Z digest=sha256:db67489bd23cc5f46ef072c32af2c071b8d0f91e68640db96afdbb651ae25016

Observation e57d3da4-cce2-4184-93a0-e8ea7e0e0757 · inbound

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought cites this paper.

Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:13.898584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:13.898584Z digest=sha256:967cfa7e8a81404190652965b96851ac3f0808c16a6179d87b22890f36a5d135

Observation bec48dd9-f5e8-467c-bf47-6c9ff591f81e · inbound

Diffractive electroproduction of light vector particles: leading Fock-state contribution in the presence of significant higher Fock-state effects cites this paper.

Diffractive electroproduction of light vector particles: leading Fock-state contribution in the presence of significant higher Fock-state effects Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T05:25:57.560559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:25:57.560559Z digest=sha256:22004e1563abd8d7cd99b7598ff50a5e626dda81c587bc0ab62969f0ff05e62e

Observation ea305080-f6c1-45e1-b84d-4a875a2b6e31 · inbound

AdsQA: Towards Advertisement Video Understanding cites this paper.

AdsQA: Towards Advertisement Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T20:20:36.908527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:20:36.908527Z digest=sha256:bd87634ea02da00ef200db65a224123061f8e30e0faf2ba913e0fdc886e53dc1

Observation 96708282-2c6f-4961-b694-275db5656e7e · inbound

StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos cites this paper.

StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:51:28.098757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T02:49:12.987772Z digest=sha256:22bf7db302cbc74051d383683cded0647c5bb1504b5ad9dcc0c928a9d0172026

Observation e95cc4b0-ed20-433f-a143-c9f4b56b3e17 · inbound

Streaming Video Instruction Tuning cites this paper.

Streaming Video Instruction Tuning Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:48:21.837952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T19:44:11.032898Z digest=sha256:1a635a4cb941e85afff8b977f5951cce3261a801f5a72fd64b2f37e1491a37ff

Observation 10b406f3-de90-4880-aaa4-951d4684bfcd · inbound

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding cites this paper.

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:57:53.814670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T12:55:04.564442Z digest=sha256:d4d605a12a832d0e74a8ff664809cd7a2d9559adc6a44bc5c6e581baaa6355fe

Observation 4640387a-ed4b-40e0-98ef-a9f34eedddce · inbound

Position: Modular Memory is the Key to Continual Learning Agents cites this paper.

Position: Modular Memory is the Key to Continual Learning Agents Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T19:33:04.902520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:33:04.902520Z digest=sha256:cd97f769231945d84a23517f0d2963c82703d1706a729899e138e6c18f3bf18c

Observation c2babc68-f2f8-44f2-af5c-bb9e26c79346 · inbound

StreamMeCo: Long-Term Agent Memory Compression for Efficient Streaming Video Understanding cites this paper.

StreamMeCo: Long-Term Agent Memory Compression for Efficient Streaming Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:45:50.440534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:19:41.543165Z digest=sha256:1e713990756b7346a94a13cbbbf1f9b48b2343810c0aed90cebce4065235bbb0

Observation 9ee79b02-aca4-4767-8d12-6b0124b5ed74 · inbound

Online Reasoning Video Object Segmentation cites this paper.

Online Reasoning Video Object Segmentation Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:26:02.856020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:29:03.843441Z digest=sha256:1b3f9649850a40ca1c4474a1e1805b8a1891768c33664f17c50371bea64f261a

Observation c1dcbb80-6dad-4874-a696-0d5d79bdd3e5 · inbound

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning cites this paper.

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:47.801677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:51:52.861981Z digest=sha256:dd8bb47e6bdba989771141d4517abe44b66358c8294e9a691a762aa157490082

Observation 7430da5b-7941-4ebb-849a-1d994f51f855 · inbound

Don't Pause! Every prediction matters in a streaming video cites this paper.

Don't Pause! Every prediction matters in a streaming video Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.108827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T04:32:01.379605Z digest=sha256:d604af79871ae588682ada72626b29270c9c6d286a9102f634bccec533c4017f

Observation 83bcec37-707f-44f6-8587-a9687b7594e2 · inbound

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding cites this paper.

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:20:56.376815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T02:30:55.939351Z digest=sha256:8247f08bd6973dbed17c54ce671f2af9beaea762a8b563bdb7cdbc05e8e8edb1

Observation e68a8438-83f1-4109-b5cd-36c9b9750cb4 · inbound

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding cites this paper.

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:17.745300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T03:00:34.728880Z digest=sha256:5a4ca18fe290e5a93c1070743cc79f72a99191bc6b06ad69b53dfb2da4f1a8e2

Observation 28611be7-b521-4a79-b59f-77353832f72c · inbound

StreamOV: Streaming Omni-Video Understanding via Evidence-Guided Memory and Response Triggering cites this paper.

StreamOV: Streaming Omni-Video Understanding via Evidence-Guided Memory and Response Triggering Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:34:02.045071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:27:17.092553Z digest=sha256:d0e3f346b7ee0ba4352730ef8fa8de42d3a9882e599c22a7517d29c0e3642ad7

Observation aa4d9b63-b76a-4ffe-8f0d-c3c75a2c8f0d · inbound

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams cites this paper.

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:51.046724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T18:24:57.881644Z digest=sha256:28a48423acd8716df261a8bdef2ec00d53a9bde6a57268905f23bf6f7a5bfb00

Observation 8c46cb1b-9174-40ba-9407-3e7327970d1a · inbound

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning cites this paper.

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.977490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T18:15:28.261185Z digest=sha256:7b410745119dfe43c4e24cbb576efe0d861a778840dfd92fb4c879cfafa81eac

Observation 568956ce-6d85-4e03-9139-ae656641ebf4 · inbound

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning cites this paper.

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T15:54:12.909974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T15:54:12.909974Z digest=sha256:e6c1bea6267169ff29d49511d278e589e9ce48c338850578003453922ea7168e

Observation fc8ede71-d296-4845-9683-09091ca4b6a9 · inbound

EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision cites this paper.

EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.375881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:40:13.720341Z digest=sha256:78f1a90ca55ead8d8dbf1e6a3634e9f97c68d40a22ec0764d28e4bc487772a5d

Observation b6fdec21-3bea-44d9-bb3a-408d96c68408 · inbound

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks cites this paper.

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.799122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T06:16:07.090870Z digest=sha256:e0a0d70d71cc53762efa8164a63f22cce25fd6477fb4ab8c3070018d271a2cb9

Observation cc8386c5-477b-468e-8966-e97c33317c4c · inbound

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding cites this paper.

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.826651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:11:01.690237Z digest=sha256:f26350d31d8c3d2d5a0dd51a1bb7e6b74ca8e07a7de6079dd756d99705c24acd

Observation a9e49ad1-317e-4670-b9c3-23633883a83e · inbound

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams cites this paper.

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:56.211274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:12:46.295455Z digest=sha256:6d198551a7d4019d5f0a38ddd405d260a4f9af9820b3fedb38ef4ecb633a845e

Observation ddcad86d-fe93-4418-b99d-a8e7fa07a6d4 · inbound

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference cites this paper.

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:19:29.882474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T18:17:53.013043Z digest=sha256:ec315fa2e7c416153ace82593f00b7414bdfcebbadbec1a4028259391e056d71

Observation a35e0a5b-0773-466f-8269-8709cf983134 · inbound

FOLIO: Focused Semantic Memory for Streaming Video Understanding cites this paper.

FOLIO: Focused Semantic Memory for Streaming Video Understanding Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T05:40:48.131445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:40:48.131445Z digest=sha256:1f498f8ac826cf9e8580d1175ec6b583a75cf53b337fdc3fcfd813f63d7591e7