Pith. sign in

Paper Citation Record · LEDGER

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation

As of 9 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 3 inbound Pith citation observations for arXiv:2501.19098.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.19098 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T21:21:45.825228Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:24:10.600569Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T18:16:30.548575Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact7
  • verified fuzzy24
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2d420fea-927a-4bc3-ba40-d78db2b0997c · outbound

This paper cites write newline.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.521644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.521644Z digest=sha256:852111523bba1d127273674711d89087ef29706f4fa64b00b16057b515b85726

Observation 7e276d2c-9384-4bc3-ba82-8ee8c582578a · outbound

This paper cites Prompt Design Matters for Computational Social Science Tasks but in Unpredictable Ways.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Prompt Design Matters for Computational Social Science Tasks but in Unpredictable Ways

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.528657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.528657Z digest=sha256:f0e353078b138b4acb3fde9557f7ba82c713817739e0b9cdfe9ee5e0f83b81e9

Observation 6ac8c100-f255-4f77-9e94-f0820ad3f37a · outbound

This paper cites Neural machine translation by jointly learning to align and translate.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Neural machine translation by jointly learning to align and translate

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.341786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.534256Z digest=sha256:28560aa885d8380894aa284e0b298c1459c39bd7c4fd5d606a1df238573ebe0a

Observation cce2e19a-069c-4cc3-82db-e57f6af2fbeb · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.539843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.539843Z digest=sha256:65a34c1754eb0255b342cbf79896648d4e6ec2a140512bd76e90ac16694a8f88

Observation ab570d1f-42c7-4d8d-88b2-10f9c56e9479 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:47.325118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.545530Z digest=sha256:11057ceb22c4fd38b72ee2450dde741366cf5e3e722b3314e6e318f55c43c19a

Observation 8c16ccaf-a6a8-4f1b-86db-8013ff1a3383 · outbound

This paper cites F., Konkle, T., Alvarez, G.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation F., Konkle, T., Alvarez, G

Reference 6

Resolution
verified exact
doi, observed 2026-08-09T21:21:45.951492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.550373Z digest=sha256:1a69d9cc2065a01725b207b4e4531bcd59b4893594005168bde181c8049a63c5

Observation a12d9185-0fc4-4f21-9b09-a646fc35d194 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:47.309315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.555577Z digest=sha256:73bd3045a518ad66deb0bbdb4283328cd97bd9437f2560e6161f6a55b1ac9e5f

Observation e4da986d-4097-4bac-95b7-9598370d1349 · outbound

This paper cites Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.560977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.560977Z digest=sha256:1869304805191bade06890e16b12319692c906d64cc62b6e444b5544c4969c74

Observation d50776a5-9d56-4f1f-bb01-b05c14d94e2a · outbound

This paper cites F., Jadhav, S.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation F., Jadhav, S

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.293705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.566044Z digest=sha256:cdd1ce7e542528eafbd060f5d7c811302a0a6679f8d200257d1fdbdf07d240e9

Observation e7e4ae0e-789b-4287-8a37-dab984907c8a · outbound

This paper cites Episodic and associative memory from spatial scaffolds in the hippocampus.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Episodic and associative memory from spatial scaffolds in the hippocampus

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.570797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.570797Z digest=sha256:74511a95b6cb64f36c583c02310b1921f2a5a2f42d4fd88bad289a764de3e95a

Observation d94dd3c6-0fe4-4d42-9c67-1d61c89559f3 · outbound

This paper cites ShareGPT4Video: Improving Video Understanding and Generation with Better Captions.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation ShareGPT4Video: Improving Video Understanding and Generation with Better Captions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.575570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.575570Z digest=sha256:2fbb5c31ab9e7a6e6e558581455606a5015e46cfb4ef601aefc2a167db4443a9

Observation 117bf029-02f5-4f29-9bde-d97cc264a26d · outbound

This paper cites Longvila: Scaling long-context visual language models for long videos, 2024 b.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Longvila: Scaling long-context visual language models for long videos, 2024 b

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.278181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.580267Z digest=sha256:f9dd1ab711abf3109ca7500e6b10f1a0391675438ca9e2548979f571ec34d37b

Observation 0d7b4423-aca4-4326-b7c3-b8c1d828d0b4 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.584677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.584677Z digest=sha256:5ef78877f2a2d98c895ea8d1680fea48b40bbe43eed31b5e1d88a650b71a5755

Observation badaeec9-1d75-4268-8686-65d57d529503 · outbound

This paper cites E., and et al.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation E., and et al

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.262409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.589277Z digest=sha256:816ad5621275041365ad74777bb28cfc70c0e2a3268e0efb5c4d98c151077d02

Observation 723d6768-9c46-4a7e-b2f5-0501d8ce85ca · outbound

This paper cites T., Schapiro, A.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation T., Schapiro, A

Reference 15

Resolution
verified exact
doi, observed 2026-08-09T21:21:45.923638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.593938Z digest=sha256:8231185a6023daeaef60a491f71d95575f43dc91e7577fc34a948157efeb1f0f

Observation f7af596a-846d-4370-b4f7-c708b7838e9a · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation An image is worth 16x16 words: Transformers for image recognition at scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.598511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.598511Z digest=sha256:8e4d8448684d3a0a9f15c96eb97ef25594370e0d05ddb51b8e1c6e388dc3873f

Observation b151fde3-fa52-4e42-8d60-907fda469ffc · outbound

This paper cites The consolidation and transformation of memory.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation The consolidation and transformation of memory

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.236680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.602838Z digest=sha256:299f0140396a13dc6ad363809f47219c09f8e557097d71d9f7f844e8b463b18d

Observation c519e660-f620-4b1c-8090-8f697a1dac0b · outbound

This paper cites Eva: Exploring the limits of masked visual representation learning at scale.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Eva: Exploring the limits of masked visual representation learning at scale

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.219458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.607713Z digest=sha256:615570e366eb411efdb66a09570d0ab0782d4833cc7bf727026f10689db486c4

Observation 86b60a2b-0343-4ecd-8722-d02aaff42d73 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:47.204297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.612195Z digest=sha256:b68996081059617cb2ad53833462786598071102d89ab23e7fcf964621d4b535

Observation 3b69a703-fd94-439e-be3c-a3366ec59cd3 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:47.187660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.616934Z digest=sha256:2423bd275e48f67d3fade3524f3bf156e412f87d2ffe1f2315bd8f8a84859642

Observation 197bf961-51c8-4dee-a96a-20a6d494e3f5 · outbound

This paper cites T., Norman, K.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation T., Norman, K

Reference 21

Resolution
verified exact
doi, observed 2026-08-09T21:21:45.906727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.621485Z digest=sha256:50e95153a400ca9ab0b4782a7086bb179eac3dc38d29d974dec3928f54b2e4ee

Observation 92067a1a-3d09-490f-9f10-32406ea65b14 · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis, 2024.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.172410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.625992Z digest=sha256:e2c7b0fdb58e20bf248fa11de81c3be84b9e4ab53c068ee92e107c74b8ba2ace

Observation ceb5b2e6-d6af-40cc-83aa-6351042f1f71 · outbound

This paper cites O., and Nader, K.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation O., and Nader, K

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T21:21:46.683811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.630762Z digest=sha256:9b9b9a8e8a5ee9d8dfe9bb5ab893ddeaf09360bbae7ca3588cc6e317a643109f

Observation 2b26a268-b01c-4b6b-aa50-014d0a093f5a · outbound

This paper cites Langchain, 2023.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Langchain, 2023

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.156465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.635090Z digest=sha256:02874269103cf2753287bfb4688b91cbd6ce33a81841a422103327bb8930dacc

Observation 68038fef-b409-4940-ba8a-46ce63c6eae0 · outbound

This paper cites Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.639522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.639522Z digest=sha256:58a69647482d9f4ed982ee6a1cc210d9c8566090c4a846dba4feb0cdc9a4d4b5

Observation c524dcd0-042a-4a74-9da4-f6af7d753944 · outbound

This paper cites Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.644388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.644388Z digest=sha256:92c00fbfdc7284a00d3e407830c6cd8ae74d2be8f4b33558b4df193e84c4b257

Observation beb64fa5-27e6-4595-a4ea-20ca8a82ee70 · outbound

This paper cites BLIP -2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation BLIP -2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.130359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.650082Z digest=sha256:c902d626567dfc89c4fef89fb321bb115cbbe035fad9388a91ddc536bf53087f

Observation a29aa8fa-f608-47fe-bd84-ffa7172425ed · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation VideoChat: Chat-Centric Video Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.654533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.654533Z digest=sha256:fcbb7a8f56775d010b72fc40375d01f91172e9c0108f1d2f7bfece517fd2cf22

Observation 62880aad-12d4-44c3-9e85-67cdf087f958 · outbound

This paper cites Unmasked teacher: Towards training-efficient video foundation models.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unmasked teacher: Towards training-efficient video foundation models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.114348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.659809Z digest=sha256:eb89d2cd5e16e6ed28f9d1d5edfd2684bd7507d48f67734a1a167a6ed1366c2d

Observation 2806aa15-9fd4-43c2-a6cc-a52db8a1468e · outbound

This paper cites MVBench: A Comprehensive Multi-modal Video Understanding Benchmark.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation MVBench: A Comprehensive Multi-modal Video Understanding Benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.664685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.664685Z digest=sha256:c60b9fac973d880236c6838fe14e7159ddc8dafd292fb4258c977fb7a4b693f5

Observation d839a870-bf8f-49c5-b0ad-66350725a192 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models, 2023 d.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Llama-vid: An image is worth 2 tokens in large language models, 2023 d

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.097929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.670166Z digest=sha256:47ad53e481e7133cb547394e691612759ef599490a3ec23dd92345ced97c4d47

Observation bd08e1fb-311a-48a1-9128-d71bc60c3250 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.675211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.675211Z digest=sha256:dbd014f36c0790aae2d8b0797809c56b2bc93b868cca2f14637c24283b9b702b

Observation 559b63bd-5f9f-4fb8-adf3-764720fca6fa · outbound

This paper cites Kangaroo: A powerful video-language model supporting long-context video input, 2024 a.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Kangaroo: A powerful video-language model supporting long-context video input, 2024 a

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.071244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.680368Z digest=sha256:348bb5da39b8c4c20a76777f27fbc5352088bb11f3c7b55ba996919addca39c4

Observation 623c2a95-14ac-44a1-974f-138f6c9b5feb · outbound

This paper cites St-llm: Large language models are effective temporal learners, 2024 b.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation St-llm: Large language models are effective temporal learners, 2024 b

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.055866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.685113Z digest=sha256:eca22aec6e6feb945446ea86d0ea5dad6aaefff9808234cb404aceb89773d6af

Observation 61483408-a88a-441f-bdb9-91b5c2cf9cf4 · outbound

This paper cites Nvila: Efficient frontier visual language models, 2024 c.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Nvila: Efficient frontier visual language models, 2024 c

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.040305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.689976Z digest=sha256:c8e3a3d8ae06343481cbdf857ae9b5ef0976eccccca8d8e443246a81ef358b36

Observation ecc6e15f-21d8-4830-a474-16facfd4ff2d · outbound

This paper cites Valley: Video assistant with large language model enhanced ability, 2023.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Valley: Video assistant with large language model enhanced ability, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:47.023115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.695013Z digest=sha256:264ad0e470c6a9e9f40e4c2c5fb1dbb50aafb747807be7eb52c7510ac7864067

Observation e1dac759-2f91-4eae-a433-d4dcc8788259 · outbound

This paper cites Changing concepts of working memory.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Changing concepts of working memory

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.700138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.700138Z digest=sha256:881d3fc0ccfaab5d75fae770e86917bfe11e162cdbf14a13e498279a0999208c

Observation 9831d189-7541-4bc5-b30c-6499c075e535 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:47.007176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.705364Z digest=sha256:7f365186c8840121cd009f147c18e2af279e65ec39c84a4f2fff90a7b529a52b

Observation 19a7a086-642e-4a1c-8375-70f6a74947f8 · outbound

This paper cites EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.710426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.710426Z digest=sha256:abd108467a7e4f1176c9748a256b155c3e68d959611e8e55f145bfbf44c3ddde

Observation 679aea19-deb3-449d-8fd6-2afafd4dc90e · outbound

This paper cites Sparse and continuous attention mechanisms.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Sparse and continuous attention mechanisms

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.991311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.715584Z digest=sha256:7c23db069e008127326a880257704854c18afe8f238925f72722c6fb24c90d55

Observation 2830ab88-04bb-4455-bcc3-d96523499eb3 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:46.974607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.720563Z digest=sha256:2537c1618566a59fabd864693a61036658cf693667a195c2d7e0b9d39f904d3e

Observation 450d32fa-3113-4c29-ad96-55b4479442f4 · outbound

This paper cites H., Marinho, Z., and Martins, A.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation H., Marinho, Z., and Martins, A

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.956790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.727117Z digest=sha256:120723149a31f2257d24e1eff9fdb6ba2f27519a390a8c174d9dacc191a90918

Observation fdc6d630-019e-404f-8251-1640d5361709 · outbound

This paper cites Making lasting memories: Remembering the significant.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Making lasting memories: Remembering the significant

Reference 43

Resolution
verified exact
doi, observed 2026-08-09T21:21:45.877413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.732243Z digest=sha256:92b928d361f215a664425ab7eea572f559e0880030d36e3be4f8840a11a8ce11

Observation aceeafcd-2d95-423a-8c20-37eba0db165e · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 44

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T21:21:46.434914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.737547Z digest=sha256:2582a6fcf075436f5a53ea0dc55345e0c4859ba7685933872112e9b0a2220074

Observation d505ae1e-7748-47fa-8ef1-3068b1045481 · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:46.940319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.742248Z digest=sha256:5b9beb888a1ab40a4258391d4bd6e2cf07f0d9510c1faba0a5fdfcc78cb1b359

Observation 25d892b2-064d-4607-9168-7f6354bcb8df · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:46.924239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.747194Z digest=sha256:a8636bd085e9575af1816adb967edef28d31e393999780a58c20b3c21aebf9bc

Observation d909b9ed-4cae-4131-8ef2-d20d47372f2d · outbound

This paper cites an unresolved cited work.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T21:21:46.908406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.751838Z digest=sha256:2fc111014647b29baab5d1033e811cf408157b314414832add080280e6d4b289

Observation 7bb2bcc8-933b-4210-bcc4-1530b21798f0 · outbound

This paper cites Video-xl: Extra-long vision language model for hour-scale video understanding, 2024.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Video-xl: Extra-long vision language model for hour-scale video understanding, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.890034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.756908Z digest=sha256:4841b96b352b7281fc3081ae0ec666fad7a7e6b23d1a4e4b9f0ae124c0abf207

Observation 8c89f6b3-4fb0-4956-981a-daa71ace036f · outbound

This paper cites MovieChat: From Dense Token to Sparse Memory for Long Video Understanding.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation MovieChat: From Dense Token to Sparse Memory for Long Video Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.761448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.761448Z digest=sha256:578d7000709946869aa73ebd61ecdd37b55a6cc3cf96bb6a809ac40d5a702925

Observation 38282128-a542-43ec-880c-c90d314deca6 · outbound

This paper cites MovieChat+: Question-aware Sparse Memory for Long Video Question Answering.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.766382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.766382Z digest=sha256:16732a538c5292e012ecc2d7808b4d31b084d1d898f47b0ca08d02a3992c9706

Observation 439559c2-9c81-41ff-b6b2-dcdb8118818c · outbound

This paper cites Energy and policy considerations for deep learning in NLP.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Energy and policy considerations for deep learning in NLP

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.771465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.771465Z digest=sha256:b62273e372e47df19edf709a38925bff3672d494285a9c339deafc8bc5d0030f

Observation 2a54d809-0b8e-4b64-b7bd-148bb3d6122b · outbound

This paper cites Episodic memory: From mind to brain.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Episodic memory: From mind to brain

Reference 52

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T21:21:46.209806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.776568Z digest=sha256:42c5f577c8bbb6540f178836dfa98d527eb8a0f891bf3ad446a831c09367791d

Observation 3b29d028-d6c0-4638-b099-c48a15e94cc1 · outbound

This paper cites N., Kaiser, L.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation N., Kaiser, L

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.781296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.781296Z digest=sha256:6c94d56b02e83f4e44cc9ec4d069cc8c60d837c126fc35a72f2753fa017e5ab6

Observation a063eafb-c88d-4123-b770-ac8ce510ff7a · outbound

This paper cites Videoagent: Long-form video understanding with large language model as agent.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Videoagent: Long-form video understanding with large language model as agent

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.862809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.786066Z digest=sha256:467032470d088e4d4a4b44567a5630a39b8060138d3d31debf7c19d2cb279c3c

Observation fac21ca3-234f-4c9c-b8d9-4097f86b0294 · outbound

This paper cites Videollamb: Long video understanding with recurrent memory bridges.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Videollamb: Long video understanding with recurrent memory bridges

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.845064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.790718Z digest=sha256:38d3a8f60d64112b251cdf86fdca4b9a6830867ae8e397b7588abae1e1feaed3

Observation 66e6dcd5-d004-49f6-b368-7eaf9ffc69f5 · outbound

This paper cites VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.795412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.795412Z digest=sha256:40c7e687e3820ac17f5d8ca64f9ca57ae3134e0f30839a5c214530976c3293c6

Observation 439ac354-ee29-4557-b193-43f3ddcc030b · outbound

This paper cites Next-qa: Next phase of question-answering to explaining temporal actions.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Next-qa: Next phase of question-answering to explaining temporal actions

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.828655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.800286Z digest=sha256:e5e876542629d3ccff669626c8e9e5869abb967e306248090347db790f239032

Observation 65fc1f41-8b89-4349-98e2-7a520a9b6066 · outbound

This paper cites mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.812016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.805494Z digest=sha256:96033d9842b9f95aee2da6b60f4d0355e692bf3a060d6a8499803f9969bed4e8

Observation 43482614-b7e0-4649-b08e-c5d72641b19c · outbound

This paper cites M., Wang, Z., Yu, S., Bansal, M., and Bertasius, G.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation M., Wang, Z., Yu, S., Bansal, M., and Bertasius, G

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.796123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.810123Z digest=sha256:9b0698378d240d94898ad212cf4b68321bb66f79d6d63299a71f7ca2319b62ca

Observation 0245ea5f-ed5c-43f2-9335-3b34f2c48e7b · outbound

This paper cites Videollama: An instruction-tuned audio-visual language model for video understanding.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Videollama: An instruction-tuned audio-visual language model for video understanding

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T21:21:46.780211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T21:21:45.815159Z digest=sha256:79e4443b7ad69d6fbd57102d3c99f0fe0ca6de67554d1bc2062d5f2caed9a113

Observation ccba3e18-c063-432f-a6f1-204bd7e412ce · outbound

This paper cites Long Context Transfer from Language to Vision.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation Long Context Transfer from Language to Vision

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.820352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.820352Z digest=sha256:de70e86855c5d02967b37a12ef34a3ffc7da00e38c60d46b51cb8a5270bc51d9

Observation 675c6d87-f729-4e77-b915-02cfef1eb399 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.825228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.825228Z digest=sha256:ac8b8e2dcf4bf778e59018fa2fe3dc7cddc7ad14d71d93ca7e012fb4b61838d1

Pith citing papers

Observation 4304ae1e-f874-4465-bbc2-4104052c62db · inbound

Modern Hopfield Networks with Continuous-Time Memories cites this paper.

Modern Hopfield Networks with Continuous-Time Memories $\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T19:24:10.600569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:24:10.600569Z digest=sha256:47bf1721cc74dfc675c4af7110f237a97b85d46f8f2971564afa00a31ed0ed03

Observation 6f61abf0-2a63-4875-8b84-cfeeb0d97834 · inbound

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding cites this paper.

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding $\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:56.957099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:56.957099Z digest=sha256:5867b310845ab43a558d8fc5fb446c93cf8328bbee91636f775f76f5f9a4305b

Observation edc308c6-cf1f-431d-b0d1-b121878aafd8 · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding $\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:10:15.807015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:10:13.200174Z digest=sha256:ccaf3d13a2b98fb2a4f0dd465dab9622c25f39ec599e0ec75b6d1019a9c1cc93