Pith. sign in

Paper Citation Record · LEDGER

LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.18139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.18139 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:15:11.922497Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.109589Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97fa194d-5d25-4a7e-8d84-69069a3bbfd9 · inbound

When Attention Sink Emerges in Language Models: An Empirical View cites this paper.

When Attention Sink Emerges in Language Models: An Empirical View LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:41:03.758200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T17:41:03.674759Z digest=sha256:e4d96ad1d904613e4f4f3dcfe047c472f438ff09ec1a957d378ebce72acd5f0f

Observation e0eeda1f-9f52-4ae6-9519-2ecb36638b25 · inbound

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM cites this paper.

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:11.922497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:11.922497Z digest=sha256:588db221c5994ed657ffb63b97028dd83722d610a318498a472738b2cb5e84a9

Observation 1cabc8b2-404d-4115-8940-3fe58a98ed21 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:18.606283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:18.606283Z digest=sha256:2a43241de6e75a21f625bb96f60414ae7ecf3170ae04bbd90c2afbce2ebb4d81

Observation eaf5978a-e6b8-4f6c-928f-420226815c55 · inbound

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs cites this paper.

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:05.941977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:14:05.941977Z digest=sha256:d9ac94e4ece70fd74916996b213d8787d29fbb86c54ceb06e57d81f1ebb604a4

Observation b3ee3db6-3c52-4b84-865b-109953407c42 · inbound

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference cites this paper.

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:43.864806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:43.864806Z digest=sha256:34922fcc78d5d49074b0b0599b7a1f5148094eb2e1e9fd096c2a93b568d34786

Observation fdea2b54-0b7f-49e1-965a-c3998b9c5e04 · inbound

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation cites this paper.

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:13.894002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:13.894002Z digest=sha256:08d19587501f3d0f6e1cae3ef85f0b25f7e5d290f9acf1f631041e948c75d291

Observation b265b223-eb3a-43b3-b1eb-2b0e57b0a087 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.193142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.193142Z digest=sha256:697641307e5443d2da0f289ae5194db40dce1ec3b8161290de4cd09237b4cf26

Observation 50d62950-433e-436a-9662-4ce9c855f577 · inbound

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs cites this paper.

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:32:49.280218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:32:49.280218Z digest=sha256:0dfc0009d872fab559d3056b4d2a69be5f36bb1b4e830d705c812b0ef0b61754

Observation 288fe766-d424-4dab-bd1f-25d092fcba27 · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.814093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:4abc2f1604bc9e32581f16019b74eb3b995c75144fb678489b7f353e74346593

Observation 21c60f4b-2a68-41c3-ab1a-f1f0fae2b42c · inbound

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines cites this paper.

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:17:37.412398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:16:02.102975Z digest=sha256:0c97faeedc2b95720cbdcb08c86d672aedc4e450102553bcd26cb9a0e46003b0

Observation 877a345f-8fb9-460f-86e1-7b63debec640 · inbound

Geometry-Guided 3D Visual Token Pruning for Video-Language Models cites this paper.

Geometry-Guided 3D Visual Token Pruning for Video-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:09.954656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T05:49:38.346274Z digest=sha256:58f3bbcf4665c7d2f3ac20f881f0e0a7c84a33fd70546ff647a3f4ab8fff511e

Observation 8fc9dade-a558-48f2-aa43-ee39443b6fcb · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.285028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:f283063a2c71a51a93060422715318b06012600eb37b65e8dbe66a852a43c3af

Observation b76fe186-f0f4-48b4-8d16-e9c3f2ee60b0 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.116725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:23cf024b3799fbbf850d41ab395420cebe38d2a1816a22cecdc28b4895b3a34e

Observation 83b6d075-edef-44c0-aa23-0f1e797cb643 · inbound

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity cites this paper.

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:08.583429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T12:11:04.146711Z digest=sha256:016170effd0f52f4ac1748f5726c588e5794df90a496b3bb11a609dc517961e7

Observation 420c0c10-6f9d-40d2-b6af-f9b97f0c0fc8 · inbound

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy cites this paper.

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:58:59.074345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T19:55:35.051834Z digest=sha256:3ced569f7a9817c4c25613147ec6367bab64f2b5fb70c407d61d7c00aa8ad65b

Observation 1c2908e8-a975-435b-a5e8-182033b85bd2 · inbound

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning cites this paper.

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.575478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:13:15.593994Z digest=sha256:4b3a83d499ddc90b73b49356ffaebd4a6eaaf7a28eacd69595549eef693638e4

Observation 3c964e99-c6c0-42ff-9c93-a4a254a53dc9 · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.429651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:10e5126ee811a4fe488e29787456785b0dffc8d9ac923b72c086636970032fbf

Observation bd0d4249-644e-4203-951f-b49b07767c22 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 118

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.110744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:3a774af64d45b9f5933f79c8b1d7b5f271455e3ede4efc9d02801d042645b2a7

Observation f4c52329-4ebd-4876-957d-ec47b3ddc319 · inbound

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models cites this paper.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.746228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.746228Z digest=sha256:a3ec567518c82fdf1707a5216e2168d9776ef5e8eb125070952290ac4ddadefa

Observation 88d04430-1dc0-4359-9dc8-82ebca331afd · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:57.570458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:57.570458Z digest=sha256:449ee33763fe5d67e1e30cc8e8893a40b4d0442ded55cdb542c863d291230667

Observation dda273cf-5b28-49da-b79b-5437f6da7fce · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.173076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.173076Z digest=sha256:c691acc9598aea1d03da2ee04498be8471bd7f9961f1047e531a07a14c93b295