Pith. sign in

Paper Citation Record · LEDGER

LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2406.18139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.18139 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T11:11:17.788093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.109589Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97fa194d-5d25-4a7e-8d84-69069a3bbfd9 · inbound

When Attention Sink Emerges in Language Models: An Empirical View cites this paper.

When Attention Sink Emerges in Language Models: An Empirical View LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:41:03.758200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T17:41:03.674759Z digest=sha256:596deb55712f2011053e5800bf64060ffd3cfbf5287d39a51d729b82997cc101

Observation 73bf3f2f-7191-4dfb-b041-4b4823b62010 · inbound

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation cites this paper.

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T11:11:17.788093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:11:17.788093Z digest=sha256:26364d2e7d353e8d00543c8e0cbb5f5bca7a8eec14340598f9c8e18db1a7c392

Observation e0eeda1f-9f52-4ae6-9519-2ecb36638b25 · inbound

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM cites this paper.

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:11.922497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:11.922497Z digest=sha256:9e425c4cbdfbb687b4fd95c45ec4b7f8c54551e10c61961be6b81fa4590c4c11

Observation 1cabc8b2-404d-4115-8940-3fe58a98ed21 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:18.606283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:18.606283Z digest=sha256:476047c55700e59563e28e790e6b8b1c31a9ecddcef054d017171bd3676a8e8c

Observation eaf5978a-e6b8-4f6c-928f-420226815c55 · inbound

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs cites this paper.

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:05.941977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:14:05.941977Z digest=sha256:5010ec60f99f5851532f2d542621cf4a2a485a34e844e5ad0fbe7e053245bf1d

Observation b3ee3db6-3c52-4b84-865b-109953407c42 · inbound

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference cites this paper.

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:43.864806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:43.864806Z digest=sha256:aca7e84a3be0073ebc3bfcad337ac6859cd299f5227faf09675f00d31840f22a

Observation fdea2b54-0b7f-49e1-965a-c3998b9c5e04 · inbound

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation cites this paper.

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:13.894002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:13.894002Z digest=sha256:21ed38f8d4ac073c303f768182cc16557ca454d5255ea1b6b6dc9ab6cdd4e3e7

Observation b265b223-eb3a-43b3-b1eb-2b0e57b0a087 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.193142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.193142Z digest=sha256:697641307e5443d2da0f289ae5194db40dce1ec3b8161290de4cd09237b4cf26

Observation 50d62950-433e-436a-9662-4ce9c855f577 · inbound

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs cites this paper.

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:32:49.280218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:32:49.280218Z digest=sha256:c7c8689cb63798952e577ff7c2c19e058851f0d04516ea2d1f78e1f6db13a644

Observation 288fe766-d424-4dab-bd1f-25d092fcba27 · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.814093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:c826c7ba45c724b4c185244ed8816b501db3df946da8937f70ce8556ba337aaf

Observation 21c60f4b-2a68-41c3-ab1a-f1f0fae2b42c · inbound

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines cites this paper.

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:17:37.412398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T08:16:02.102975Z digest=sha256:56a44506f6de79c7aec7b672e4e43823ec39741b7ef23e898f78342bbbcbe629

Observation 877a345f-8fb9-460f-86e1-7b63debec640 · inbound

Geometry-Guided 3D Visual Token Pruning for Video-Language Models cites this paper.

Geometry-Guided 3D Visual Token Pruning for Video-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:09.954656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T05:49:38.346274Z digest=sha256:652b5dc9c5f62e5125a0a4c8b4340173d725169b96ddcd6f533e85295eb133af

Observation 8fc9dade-a558-48f2-aa43-ee39443b6fcb · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.285028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:e7a7a807f63027715aed011144a0c4847cf32223b937b32adc3cd970d5ce6fb4

Observation b76fe186-f0f4-48b4-8d16-e9c3f2ee60b0 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.116725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:a28c65f51ae8b87b4ccbfffed34a1d275b4746ba4fdaae3733aa7c031f75e607

Observation 83b6d075-edef-44c0-aa23-0f1e797cb643 · inbound

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity cites this paper.

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:08.583429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:11:04.146711Z digest=sha256:985a4fe3786b663068ca1b959e7318c310ef90e68c89de351e7f70d90f457097

Observation 420c0c10-6f9d-40d2-b6af-f9b97f0c0fc8 · inbound

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy cites this paper.

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:58:59.074345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T19:55:35.051834Z digest=sha256:5b9d07829c02c5feeb15764f0c034b7d2d4432914e1790583e9524f0322633bc

Observation 1c2908e8-a975-435b-a5e8-182033b85bd2 · inbound

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning cites this paper.

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.575478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T23:13:15.593994Z digest=sha256:3e7f828b4850532dba91c186182763a82ae7314bfc2ff5ec79909bbca42ccca2

Observation 3c964e99-c6c0-42ff-9c93-a4a254a53dc9 · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.429651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:2b7f7b782f1fb1f45aad05c89aa008657dfb88b3ee07728027f50781f050a026

Observation bd0d4249-644e-4203-951f-b49b07767c22 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 118

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.110744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:e2f11c0ff39d83adf76c535675ccd0ea07323a3febd988f1d58744104b45c4d5

Observation f4c52329-4ebd-4876-957d-ec47b3ddc319 · inbound

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models cites this paper.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.746228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.746228Z digest=sha256:a3ec567518c82fdf1707a5216e2168d9776ef5e8eb125070952290ac4ddadefa

Observation 88d04430-1dc0-4359-9dc8-82ebca331afd · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:57.570458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:57.570458Z digest=sha256:a7de0e7937af724f87f82372e5a4e3e97123817e1ac50face0fad1abfb0870f8

Observation dda273cf-5b28-49da-b79b-5437f6da7fce · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.173076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.173076Z digest=sha256:1760f0bb8b95fef6519c68c6e10a00a234d9e7cf835b3497c298a2e95ab5431f