Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:56:05.432793Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 4 inbound Pith citation observations for arXiv:2506.03990.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:56:05.432793Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T23:46:51.641946Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:19:29.938511Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation aa61887c-8890-4ecc-bfdf-96752c60cec1 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc61cc26-beb8-430d-85f6-82c0ca483762 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c79db32a-ddd9-407a-a895-d22bc85a8470 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e9bc4e7f-5737-47b7-bc18-3eda5692741f · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2a779ba7-a508-414c-a14a-7e710aa26f10 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding An Introduction to Vision-Language Modeling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4c6e30f-30cf-4d31-b3da-6649ec251dc0 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33af4884-a719-4c17-889e-bf36fe9aa410 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dc5e770-71cc-4304-8aad-4f9b10a4b007 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26f4b637-9060-4488-80eb-9c2a838b3e8f · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a916a4f9-d205-473b-b4f7-13df30d1a489 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5b5039e-25b6-42c4-90b9-336ed569aca1 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Matryoshka Query Transformer for Large Vision-Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6f80ed6-3d9b-4dc6-b009-6f5c563b7436 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation da39ac70-4056-459d-bb73-49d2f7a73fe0 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 264fb417-cc39-45a4-a218-e19a0cc9fa8d · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Aria: An Open Multimodal Native Mixture-of-Experts Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfbdef21-7869-468d-8320-59bfb76e569e · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb72792c-376e-40b7-80b9-4ed91c6044b2 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding VideoChat: Chat-Centric Video Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13319851-0f2b-40b7-8362-c5ffbd386569 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 81818563-8738-439b-a096-f401ce88c5b9 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding TokenPacker: Efficient Visual Projector for Multimodal LLM
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ac41062-a81f-4c23-8768-e9838f583c2a · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac04eb9-4d82-49b6-bc64-11e7635497c8 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f293930-0977-4141-a852-9528cbacc60f · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b21742c1-cb48-48f8-9e18-72ff9f469fec · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20152ca6-858c-4b66-ad82-e400bf1af8e4 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d1ff616-b72e-42c1-8406-557aaa997bec · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 53677ca5-1ac8-42b3-9335-e3510cf0ee6d · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ec724f-b4c6-4fcc-a125-e0bdab7eedd9 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2792e61-63a5-472a-8d51-e05cbbedfa04 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106473ca-fc15-413e-9740-5656892a4c6f · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64346559-2710-49cb-890b-c1fc11bfaf92 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559084cc-f064-4cc9-8f56-f56b9d8f608a · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Dynamic-VLM: Simple Dynamic Visual Token Compression for VideoLLM
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c33537-a5e3-4dc3-b645-08885b5d2b8c · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f45281d-0bfb-4c17-81bc-01fc47dc6ea7 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe533a48-10ce-4d8c-beb2-38ac03429af0 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca11976d-f9ad-4707-a57a-bbee01e75148 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae79a53-3939-4976-a6e4-916cb7ee8073 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b80eb3c8-54d4-45d1-975d-01eed4caa809 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a589e1a-66ac-416d-80f8-67ffa2397b8b · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Qwen2.5 Technical Report
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c81d325b-2bb5-4b51-8784-28bdf7bd4a62 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d1ab832-3f2b-4656-a918-af0566f5aa48 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e6a6b50-4c14-491e-9a18-6c60e07f2cc0 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Sigmoid Loss for Language Image Pre-Training
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29db3ad9-d287-4e59-8754-7cbb96c444d0 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1f0c0cb-268c-402c-aad2-c20d039f92d4 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 500a1ece-fc59-4481-a16d-988122648280 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b9f381-ede6-48c1-903c-f0c0bcad59ca · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding MLVU: Benchmarking Multi-task Long Video Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ad1be84-66b9-4d36-9bc9-d868946137e9 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cc700f1-1743-4f53-81f9-3d7a5be55d0b · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding online" 'onlinestring :=
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b331cc54-46dd-4cba-be33-02492310fb91 · outbound
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding write newline
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b063c2e-aa4d-4da7-973c-621e23709d06 · inbound
Stateful Token Reduction for Long-Video Hybrid VLMs DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bcde7a0-8149-44a2-85f6-c3ce2e762161 · inbound
ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 62bd9af1-87b9-4723-8fd0-00f2d8d2e4a5 · inbound
FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
Reference 248
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 84d2da63-359b-4e4c-9943-42b6095c3fb3 · inbound
CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.