Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:00:10.950717Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2608.03812.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:00:10.950717Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c4ff1c81-251e-4035-8a6d-98e3d0f22594 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2a2490a-d605-4a04-b8aa-288cb3f7a3f2 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Token Merging: Your ViT But Faster
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85a71cfc-c61f-427f-bb6c-9328324d3d24 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b8993aa-07a6-4baf-91de-9cc073e2eb77 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Avocado: An audiovisual video cap- tioner driven by temporal orchestration.arXiv preprint arXiv:2510.10395, 2025
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc11dee5-e447-40a7-bf99-f86c01531b48 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf782b0-604e-4e06-ac09-592868da529c · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models OmniRefine: Alignment-Aware Cooperative Compression for Efficient Omnimodal Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4c6e423-058f-4fc6-a23b-65cb6784fea9 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6767fe0-689d-4823-a9eb-22133b8c728e · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Unified spatiotemporal token compression for video-llms at ultra-low retention
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0aa1385b-96f6-4d00-897b-e93558c66bde · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Study on density peaks clustering based on k-nearest neighbors and principal component analysis.Knowledge-Based Systems, 2016
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7229e11-888c-461e-bb5b-e131dff37ef1 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Flashvid: Efficient video large lan- guage models via training-free tree-based spatiotemporal to- ken merging.arXiv preprint arXiv:2602.08024, 2026
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1563059e-d21b-4e68-84a5-fffc1a5b873d · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models VITA: Towards Open-Source Interactive Omni Multimodal LLM
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85f39ece-efb9-47a9-8589-866d897e697f · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df338285-baa2-473b-a284-b677b5a98a8e · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8724ec8e-e5ea-457b-b42d-53d63a4e1e83 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e80f3558-050b-468e-920d-7cfe665d1e4b · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Echoingpixels: Aliasing-resistant joint token reduction for audio-visual llms
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9e4c7575-aca0-494a-8d10-8be0b01266c7 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Gemini 3.1 pro model card.https: / / deepmind
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f691b94f-fdd3-45d0-82b7-fb384736a7ad · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85fb0f3a-968e-4add-bd8b-17f03ba460b7 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models GPT-4o System Card
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 669715a8-0864-4eda-a1c1-70f44e032344 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models ContextGuard: Structured Self-Auditing for Context Learning in Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb040714-fcd5-48b0-a191-969f1ff2c6ce · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Token pruning in audio trans- formers: Optimizing performance and decoding patch im- portance.arXiv preprint arXiv:2504.01690, 2025
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 969fe028-3840-4a74-9ff4-ab0d84dfadbc · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Omnivideobench: Towards audio-visual understanding evaluation for omni mllms.arXiv preprint arXiv:2510.10689, 2025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a535e23-31cd-43b0-b60a-cd5745492821 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models OmniGAIA: Towards Native Omni-Modal AI Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02549923-0200-415b-82f1-2e61ecf47e87 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Speech- prune: Context-aware token pruning for speech information retrieval
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec615e69-17b1-4750-9904-d44b2db9a144 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Video compression commander: Plug-and-play inference ac- celeration for video large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e67b4fa-d291-4650-8def-ea6526ef4f2d · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Gpt-5.5 system card.https://openai.com/ index/gpt- 5- 5- system- card/, 2026
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0be0490-6173-4d6f-bfd8-bde6e8008139 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models OmniDrop: Layer-wise Token Pruning for Omni-modal LLMs via Query-Guidance
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 01735dad-625d-4a25-ba0b-57e432753fd7 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Clustering by fast search-and-find of density peaks.science, 2014
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7e7f61a-fc50-4892-8ba9-396f3c1eacd7 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Fastvid: Dynamic den- sity pruning for fast video large language models.Advances in Neural Information Processing Systems, 2026
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a797368c-461c-4461-b661-4f748b721c06 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Mavors: Multi-granularity video representation for multimodal large language model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dad8e06-66ef-42ec-b8ef-8d8146649fb1 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Audio- visual llm for video understanding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98885e3e-5829-4942-9e14-7b991be14341 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models TokenCarve: Information-Preserving Visual Token Compression in Multimodal Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca6b901-294a-4feb-ad39-e4f4a9168434 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models video- salmonn 2: Caption-enhanced audio-visual large language models.arXiv preprint arXiv:2506.15220, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58d07e37-2acf-49f2-b954-dc6c6718216e · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Dycoke: Dynamic compression of tokens for fast 9 video large language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7911a0d-46d9-46a3-a5ed-b0260e4b4a63 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Omnizip: Audio-guided dynamic token compression for fast omnimodal large language models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f599502c-6604-45db-8365-c1e444359355 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Lvomnibench: Pioneering long audio-video un- derstanding evaluation for omnimodal llms.arXiv preprint arXiv:2603.19217, 2026
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f9388e2-8b93-4c69-b335-62a917ab6e8f · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Qwen3.5-Omni Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b29d5af-d0d7-4469-8ab1-0f1a0e14ed1c · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Monet: Reasoning in latent visual space beyond images and language.arXiv preprint arXiv:2511.21395, 2025
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ac6f9ca-05d6-47f1-8765-54db3d9cdeb0 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Beacon: Knowing when and how to perform agentic visual reasoning, 2026
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8282b130-6e75-4184-bba7-5950cd3130ce · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 701a70f1-3fc6-4030-b6a8-db64f3001a8d · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Varcmp: Adapting cross-modal pre-training models for video anomaly retrieval
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f595ffb2-ac58-4259-9c6b-752dcc24693c · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Avadclip: Audio- visual collaboration for robust video anomaly detection
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9aa36018-c1f7-4156-8b9e-ad6fa986daaa · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Stage-adaptive Token Selection for Efficient Omni-modal LLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7cfe3c4-3dc9-4fa6-a3b7-e8a13743fa6b · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06452d4f-c1a8-44f1-95eb-50b1c332fe3c · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Qwen2.5-Omni Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8895f65-c30c-4837-8fe5-590202e4c06b · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Qwen3-Omni Technical Report
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ff0a44c-6bed-496d-a2cd-b373b08cf750 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba8c6a1-d9af-4899-9b7e-b884074ff5dc · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Visionzip: Longer is better but not necessary in vision language models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9641e354-286a-4a75-9b54-1649e8272a96 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Audio-centric video understanding benchmark without text shortcut
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation efb0d17f-458b-40e1-ab02-c7e005dbd68c · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6589d0ef-0127-47bb-b8c3-cffbb2fe0e68 · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Lmms-eval: Re- ality check on the evaluation of large multimodal models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab4978fd-8478-4d6f-b98e-8b08c42cb70e · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Debiasing multimodal large language models via penal- ization of language priors
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b95c888-b354-4cab-b365-f2adae77069a · outbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models What is this?
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.