Pith. sign in

Paper Citation Record · LEDGER

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 0 inbound Pith citation observations for arXiv:2607.22586.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.22586 v1

Coverage vector

measured 84 of 84 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T11:58:32.095426Z

measured 84 of 84 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

84 of 84 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved84
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4c9fa41f-ceba-4ddf-bf19-fd591d23e4c7 · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:19.894867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:19.894867Z digest=sha256:a15d030d903378af87e6dc18a0555b8b90fadc123dd9e90697871b22b86d0dfd

Observation 610b6963-6a14-426d-9b6a-0b323fa6235a · outbound

This paper cites Nikolopoulos, Hans Vandierendonck, Deepu John, and Bo Ji.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Nikolopoulos, Hans Vandierendonck, Deepu John, and Bo Ji

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:20.054978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:20.054978Z digest=sha256:e764d376cd95ddccc58239e371edc1e07f8b3c33979c6cabf14e0c8574dc93a4

Observation d5560da7-6839-4c85-a888-927562fbf26b · outbound

This paper cites Qwen2.5-VL Technical Report.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:20.222339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:20.222339Z digest=sha256:6a8a36213afd130feddffa68084a0cad76e8f6918aaf7d6390684866ce0e131e

Observation 1403ce93-0877-46f5-8cd4-6d4cd083be51 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:20.585480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:20.585480Z digest=sha256:0fc96d250e49a31915d7d11cdc8a5cf7ec364a08936fdaafc9b7e25952dbb30c

Observation 1ab748ed-05c8-4f8d-8d04-4e7a709034e8 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:20.913521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:20.913521Z digest=sha256:da0d74cc41dc814e9cf0047939ca718016bf3d3124d5268494ddd37f07350279

Observation 4ed5212b-863e-4cd8-b443-04f25aa6187c · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:21.026569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:21.026569Z digest=sha256:0da977d56b89b3ed5fc7da29c1aaa6c60864bc365b4b43821034138f7939956f

Observation 04a2b01f-4983-4b17-80b9-b58b4dbf6f4c · outbound

This paper cites See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video LLMs.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:21.450728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:21.450728Z digest=sha256:e7239ad7aae565233298f2f7f6e6ae5a52ba051d4a7ab4eaba06058b50e2c265

Observation ad466646-dfd4-4c2e-bfbf-33543f930347 · outbound

This paper cites Abdi, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Abdi, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:21.569466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:21.569466Z digest=sha256:36484abe1d369de632bf325cf39cd82fda81a61848d8bc513d542b0cbdefd757

Observation ea860bbb-30e8-454f-b31c-de52ed70f3d1 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:21.770610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:21.770610Z digest=sha256:f9ad54b1d0f9a5df333fe8db488b559c79773d40f90d4712134e3acbaeff1971

Observation 3df26648-73ab-4607-b379-44f33c18f53f · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:21.891492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:21.891492Z digest=sha256:3c06dd2c2914cf1414687c8c74dc98307f1ff858398b161a431d7c25d7d61aa1

Observation 672026b0-2428-4a96-852a-a19364078c1e · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.063778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.063778Z digest=sha256:e1afcb674bf849508780f8fe3d77344b2b7aec871b6a9b52210a02b06376c551

Observation 40a1932f-c375-43cf-892d-0df718559623 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.181795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.181795Z digest=sha256:4e642e5bfeb7c80608dd0ef055a6b932968f0be6632ffba28db451c03d79dc6d

Observation 713147ac-3da5-4830-9770-8dfa6e57943e · outbound

This paper cites Training-Free Activation Sparsity in Large Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Training-Free Activation Sparsity in Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.293241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.293241Z digest=sha256:d9ac7aeaddda2fef8d69744b3999219b249a06a970fba87bfd99257b4f1121ff

Observation b2be34b0-341b-4c68-b7b5-37c4b0c39e49 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.698273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.698273Z digest=sha256:302c0a8be9ae88d4b6102260b40ea98ba74e4ff3541522242a0516928fe8adda

Observation 7fe856aa-0215-44ef-9a08-e0b5b3992028 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.839506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.839506Z digest=sha256:a2f71eec1623a36a4e7beb3889d157542cf154db6b966e27899142277e78437c

Observation 50b5f7a1-d2c3-4162-a318-e3c84d2dc512 · outbound

This paper cites Morse, Raghavv Goel, Mingu Lee, and Chris Lott.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Morse, Raghavv Goel, Mingu Lee, and Chris Lott

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:22.959713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:22.959713Z digest=sha256:26580268b8dcb9c04faf33bad7469b41f5ad485c0f6b796d7b2f35c7dc39f923

Observation b00a8e43-9164-48e3-8c1a-d9f65f3bad6d · outbound

This paper cites ANLS* -- A Universal Document Processing Metric for Generative Large Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models ANLS* -- A Universal Document Processing Metric for Generative Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:23.108518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:23.108518Z digest=sha256:9e328281bb2dc8316823fe2678ad6ecc0993e1ca1f91954c92102eee17f01f5b

Observation c9c0e047-deeb-4640-adc3-d986e28eeb9e · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:23.423299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:23.423299Z digest=sha256:b16b98009f4fa98303dcf0d26aab2ac654c5a9d26fe26809ff20158f63679a1e

Observation b76d9ec2-e6a8-4c75-bb4b-7485e2118a70 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:23.571629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:23.571629Z digest=sha256:fa65821e1e1cbcd6af4b38309de663b44e9c4c6300ea8754b761b745a00275a1

Observation 30e33678-916c-476e-9e8a-73626e926fc7 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:23.810534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:23.810534Z digest=sha256:e57a7ab954bfb5195968f31db2ae0a494c4bd7224ac6f58b9ec3d31c9d17e67e

Observation 20c69d2f-748e-4b1f-8a59-577e5f5e0335 · outbound

This paper cites SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.192980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.192980Z digest=sha256:c363e6cd9b035c60bc078576f7dfe18b19abeaa4690105648b02fc71e79e8621

Observation d24cf86a-d86d-4779-af46-ecac5407352a · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.335787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.335787Z digest=sha256:c02e006ad4feb32998f781b42d9315da25d8574bb8cb0d9c2d8a9c8a3b944457

Observation f2bd8cea-7ddb-4931-ba17-f052837b67fb · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Efficient Streaming Language Models with Attention Sinks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.448018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.448018Z digest=sha256:f39bfb7d0780358615816af5837bdbb6270f2f300bc0792b810b62123dbfda90

Observation 3459f733-da4e-49d3-b874-c26b79a2e42f · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.611574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.611574Z digest=sha256:d17f4c492f4c6e989dbf2f20b898eeafca88a408690e8124faafd67293026b03

Observation 7c186425-9c29-498c-8f4c-603e8ba516ab · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.737749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.737749Z digest=sha256:4b873f2d0b7eeeacc072e44b4f5ffe963a0514e5d5d0a0f7b43233c50b2beed6

Observation 90b9b16c-7779-4e73-94e4-1064694c1459 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.851839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.851839Z digest=sha256:bca5e2c257372a01eb5a58ac6c5704ebf983e553368efde3fb779606e488d5f3

Observation 5dda8ded-2289-44d8-a403-4508c5cb6fa6 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:24.963171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:24.963171Z digest=sha256:98e05b9a592d1242c3ef051514d342fa978cd7aa73343d4ddfe04f17076638e9

Observation f811e027-85bb-4c6a-be33-3662a4401d3c · outbound

This paper cites HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.085915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.085915Z digest=sha256:ccb423373fbf2e8ee65207dcf8cd871a6268cc1375fc3db66438da70eda3bdb6

Observation 82653632-2db3-4d9f-a680-d6a8e1c94ad2 · outbound

This paper cites Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.244892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.244892Z digest=sha256:abadfc1af88256c40621a8a27363c421d0e3aa682e15161996847d31f5b94f8e

Observation 36df42ca-8cab-4db5-bc92-6a6657fc4cb7 · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.477725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.477725Z digest=sha256:6e1eb73e240810ed3ad8f9211dabbaba1b270d49330b0431852f67f41dd12441

Observation 6922c32b-0f07-417d-a96a-34fa06eff11b · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.572692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.572692Z digest=sha256:7a86982de2a30a062c8cd4d96b2dc47d8fe81f48ab51b65a065e88676e2dc913

Observation fc4a7285-a518-4125-97db-8435ee05097a · outbound

This paper cites an unresolved cited work.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.702041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.702041Z digest=sha256:a48d1a7fbf1a71b1ae0904fc39b34cf2ec25bb9bfed959b7726455076f1c9dc6

Observation 5d211f02-76b7-4335-b5a0-4b146a7d472a · outbound

This paper cites 2025 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2025 , eprint =

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.791801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.791801Z digest=sha256:4af4b84ce7e0d93ec11ce56fa959b2c24f4d2526f499f3999516dd8be0a20f33

Observation 4e6a33a1-3ea4-49e2-9cc8-d27a4e75c50e · outbound

This paper cites 2025 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2025 , eprint =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:25.922021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:25.922021Z digest=sha256:0740a78f280cf42c45896cd9bdc7dbf6be71affad1bb6684b6f03164d3566eec

Observation a6d8d12c-f06b-40f5-bcf6-2c953812cad8 · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Advances in Neural Information Processing Systems , year =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.082011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.082011Z digest=sha256:d8eb1b8a9d3a72f9160430fdb34e2d97fb41eae124cdd63082349fd668310cc0

Observation f830f334-e753-4852-be3c-3b67bbe3877e · outbound

This paper cites 2025 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2025 , eprint =

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.168724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.168724Z digest=sha256:533e3bc6bdb9083c280938451e96ac9a8657200ba6aa28c28438e7dc6e91b2a8

Observation 246fab55-4d51-422b-b25b-f32212782cfa · outbound

This paper cites 2023 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2023 , eprint =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.270260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.270260Z digest=sha256:437faaf9db560a31dcdd3bc526a0fc269a6f64576ea33fc0d547c8b1cdf10952

Observation f0cb8b69-e774-4754-89e0-325ae7162404 · outbound

This paper cites 2023 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2023 , eprint =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.385348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.385348Z digest=sha256:e21bdb0752f16e7e1799c576fe90c14cc52f4949f6dc31a425d01bf8de896370

Observation ad1ef3bc-18b5-42d1-98d5-ea8223b8c366 · outbound

This paper cites 2025 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2025 , eprint =

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.512769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.512769Z digest=sha256:3f36baa5f8ebe5c54fbef00ad454292073cfc5a99cee0051fcb8b2d02526162f

Observation 973b3d56-208c-4b95-ae63-42e351e05ff4 · outbound

This paper cites 2024 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2024 , eprint =

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.635446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.635446Z digest=sha256:735b720b95463fd7ff82816278765d5a2f4addc8f6bd0dde871b1319a3ecdfa0

Observation ba523d12-1836-4ff9-8495-eec731bfa467 · outbound

This paper cites 2024 , eprint =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2024 , eprint =

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.768096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.768096Z digest=sha256:ffd3df90a7200675c93084a615b5129f882f178a0f017132832827e6421e494d

Observation cfb5dd04-0ff9-4297-ad0c-19a70376f45d · outbound

This paper cites OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:26.902599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:26.902599Z digest=sha256:023445ae58d5af20ca04664fb8827c8d03d1e1279354974abed6651d6dc59243

Observation 70e8483f-bc4e-4b6e-afd4-007a2ece93d4 · outbound

This paper cites TextCaps : A Dataset for Image Captioning with Reading Comprehension.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models TextCaps : A Dataset for Image Captioning with Reading Comprehension

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.039296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.039296Z digest=sha256:6e8eba5ab57bdb1ed194bb33da9317c2b4c110ea56faf8a29fba42c60179438a

Observation 1029a303-26d1-49e1-a0cf-5863cb5e4817 · outbound

This paper cites MMMU : A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models MMMU : A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.150104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.150104Z digest=sha256:a98f04f67bf5d71f20fcfb2f830e42f5fea2e2859828d520f55700f8e2467351

Observation 956b0500-334f-4eb6-995b-dd6c947af03b · outbound

This paper cites Findings of the Association for Computational Linguistics:. ChartQA : A Benchmark for Question Answering about Charts with Visual and Logical Reasoning. 2022 , address =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Findings of the Association for Computational Linguistics:. ChartQA : A Benchmark for Question Answering about Charts with Visual and Logical Reasoning. 2022 , address =

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.289634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.289634Z digest=sha256:3a091c33546a8041bce9e3c5cecaca2031eecf1ca191f2c2c78cc99539dfb800

Observation 636132c9-4b37-4837-8a89-f7b94ec292b3 · outbound

This paper cites Towards VQA Models That Can Read.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Towards VQA Models That Can Read

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.422148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.422148Z digest=sha256:f2ac6a03d6d4efb1199260df9ff91a2a716c09a011c4f3fed04b05db2bcf7b06

Observation bb64e5c9-fa21-4552-bf88-666ffc84c5f9 · outbound

This paper cites , booktitle = "Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models , booktitle = "Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.549270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.549270Z digest=sha256:8681f3b7a6a3278e4b488b8a9097a802dab2fa64f8efc0aa14e171e2d0137362

Observation 8a43d895-8e4e-49a8-a5fd-6d3a8be1cee5 · outbound

This paper cites KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.683696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.683696Z digest=sha256:b1c3de50a66e57c43a7956643f42021652222729d002cc02e691b4b256794c54

Observation d18f3edf-a0dc-4c1a-bff2-65ae682dd6fb · outbound

This paper cites AWQ : Activation-aware Weight Quantization for On-Device LLM Compression and Acceleration.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models AWQ : Activation-aware Weight Quantization for On-Device LLM Compression and Acceleration

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.842372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.842372Z digest=sha256:eb4e10f586e7218fec4420e0ceb95ffa28786c13d5151256c740b394a0c752ea

Observation 1a51ac52-0af4-4c70-a11a-e24e0bd6f9c3 · outbound

This paper cites Computer Vision --. An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models. 2025 , publisher =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Computer Vision --. An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models. 2025 , publisher =

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:27.962962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:27.962962Z digest=sha256:d3b738ddba94c6c29a1ef51558303fb13163b1f8fddef06e37a92673a73df84a

Observation 618513c7-b8a9-49ed-a6c3-48315be267e7 · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.105753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.105753Z digest=sha256:eb6816a0a2b4e121be756c072549cccf01b30206c3c70e1992fbb8109f586ffc

Observation 94e7f057-f949-40b9-897e-1162f22cc9e3 · outbound

This paper cites Association for Computing Machinery.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Association for Computing Machinery

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.185604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.185604Z digest=sha256:f93e10130ff0ba6e10d3aeec13c43032fd22849e83eec70b28ee8ef140109342

Observation 9027bd55-13a7-4f06-9b86-e558141dad8e · outbound

This paper cites InfLLM : Training-free Long-context Extrapolation for LLMs with an Efficient Context Memory.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models InfLLM : Training-free Long-context Extrapolation for LLMs with an Efficient Context Memory

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.334271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.334271Z digest=sha256:8dcbe29b87fa5158c4ec67f6e54529c96385cbd728bddb97d5cb10e718555af0

Observation ae356680-5cb6-4441-89e3-60a5f58b5277 · outbound

This paper cites H2O : Heavy-hitter Oracle for Efficient Generative Inference of Large Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models H2O : Heavy-hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.466982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.466982Z digest=sha256:048cf43ae052dd50f299f14642abacc5a80c601fb412db70f45215ab63db70f0

Observation 15fbe337-d973-4b09-b47e-e986514b5c09 · outbound

This paper cites RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.560458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.560458Z digest=sha256:95e487b732430ef794d05f3afa49ab6c647aaf2403b6e7476d58f8c11a81fee8

Observation 3a451ae3-ee0e-4995-b1fe-ff143af95e43 · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Generating Long Sequences with Sparse Transformers

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.674400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.674400Z digest=sha256:64eb0031e51e04eac4e812cc4c232927826edbd864da98d19d64b2a37c0dddcb

Observation 1429e6f6-af11-466b-b455-4c7253c6d26b · outbound

This paper cites Training-free and Adaptive Sparse Attention for Efficient Long Video Generation.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Training-free and Adaptive Sparse Attention for Efficient Long Video Generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.800958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.800958Z digest=sha256:4b339931c0e17f23ed420ddb6564f0bf3d301476d05126f725dadc0e92d9f238

Observation 5d9ccd38-9620-42e3-8271-7f82ef51b2a7 · outbound

This paper cites A Simple and Effective $L_2$ Norm-Based Strategy for KV Cache Compression.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models A Simple and Effective $L_2$ Norm-Based Strategy for KV Cache Compression

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.898360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.898360Z digest=sha256:0fcb799c752dfa0eb1fb45ed90a55d0f61c46e5998f03dec8e5be131fe0180a9

Observation 8ff85135-0005-4cd1-a816-5f5b2afa60ba · outbound

This paper cites , journal =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models , journal =

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:28.992900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:28.992900Z digest=sha256:202b54c517219fb4ee1b849338a05f1c0057013965e502796f4992971ea50100

Observation 01fccef6-38ff-457f-ba8f-bfd7f0bb4477 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.142425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.142425Z digest=sha256:9ca31e50e038971d3bcd321bc8fe7a432023efa0bd8f53bedd837a9ab7395bf9

Observation f69adfc5-62c1-4796-8bab-c8adb5afc3a1 · outbound

This paper cites VisionZip : Longer is Better but Not Necessary in Vision-Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models VisionZip : Longer is Better but Not Necessary in Vision-Language Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.256668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.256668Z digest=sha256:d87b4a9742a05bbf26b4d384dbe3236a00311f7bc30faa4c04f5c3f764c89483

Observation bac81452-c74b-4091-9a4b-3ade95250148 · outbound

This paper cites DyCoke : Dynamic Compression of Tokens for Fast Video Large Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models DyCoke : Dynamic Compression of Tokens for Fast Video Large Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.361640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.361640Z digest=sha256:0ded8dc68dc0128d160b1efa6c730b22fd0453393cb123bf0911859b0475a9ed

Observation da5de966-860a-4fa6-8ea5-f2d30ba177a6 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.497400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.497400Z digest=sha256:e9890a66cfb914c3f6ea3ed08b4481cbffed36a188ee6389387df1d336f36bc4

Observation eedee17c-53af-4b77-81d6-c8aafc2a1471 · outbound

This paper cites TopV : Compatible Token Pruning with Inference Time Optimization for Fast and Low-Memory Multimodal Vision-Language Model.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models TopV : Compatible Token Pruning with Inference Time Optimization for Fast and Low-Memory Multimodal Vision-Language Model

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.609514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.609514Z digest=sha256:e4252d890eca3e643da07bbe9fc47772052afaa22a50bf4be48b0e175600ed23

Observation f4c52329-4ebd-4876-957d-ec47b3ddc319 · outbound

This paper cites LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.746228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.746228Z digest=sha256:a3ec567518c82fdf1707a5216e2168d9776ef5e8eb125070952290ac4ddadefa

Observation 95ed4cdf-9601-4f9c-a19a-5b5fe7b61733 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers).

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.845572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.845572Z digest=sha256:cafc9ce2e9d78db26c53049a56d5bbd22cdf4106d41fa94b48902eecfb66aa97

Observation 5348cdde-7e76-4659-8e7e-6d942dd84429 · outbound

This paper cites DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.935195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.935195Z digest=sha256:9f28024720105d3589eb25733089e7b1843c88eebfadbc81759fddf65f59fe94

Observation 0e0d93f9-2ca8-4c7a-8d0c-75e38cda5bc0 · outbound

This paper cites Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.056509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.056509Z digest=sha256:87c5c5cba084ed527779bf12804600bd878b6dc60e1d9e01de43f8f1edee77ec

Observation 9ddf2ab0-fcee-4955-aa19-8f1645883cd3 · outbound

This paper cites and Li, Dongsheng and Lin, Chin-Yew and Yang, Yuqing and Qiu, Lili , booktitle = "Advances in Neural Information Processing Systems (.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models and Li, Dongsheng and Lin, Chin-Yew and Yang, Yuqing and Qiu, Lili , booktitle = "Advances in Neural Information Processing Systems (

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.166806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.166806Z digest=sha256:06f0b9fc1132a59b4aee9c373713a7722967b6d7ee005538afa75410b90cc15f

Observation c6446cc8-25ce-48e3-b393-e3363a4e9e7e · outbound

This paper cites SparseVLM : Visual Token Sparsification for Efficient Vision-Language Models Inference.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models SparseVLM : Visual Token Sparsification for Efficient Vision-Language Models Inference

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.314903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.314903Z digest=sha256:7a3afce48a84fb15b93a70406b6480b45b47359ece592803704c075ef59851c6

Observation cf8bd247-429e-4259-a8fa-97fc5e1019eb · outbound

This paper cites LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.462516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.462516Z digest=sha256:91b522d0e732b74af95cfa1e13edd1cf72e57c210f6dfd2aa2edc134b8392fbd

Observation e7840632-cd37-4724-a617-bf44c1d6517b · outbound

This paper cites GQA : Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models GQA : Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.589649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.589649Z digest=sha256:c764c4a4854cb01a58bff80679b62a25284be4155b990ffea80118884a8099ba

Observation b7d7e243-73ff-4194-a052-a263e3d8d0a1 · outbound

This paper cites Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.699815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.699815Z digest=sha256:a6fc97107f9d1a2845fa8106183b4b83dbb2b4a11743f6606d2d8696d64cd86c

Observation d4b9599c-978b-4803-9284-d698d4c5d12c · outbound

This paper cites Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.811138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.811138Z digest=sha256:5949884982d48e6f52a51040fc929e570dfdb7b6929da3beb97cef0f4e686ae6

Observation 41e47fe2-a1c4-48a4-b890-0c33ada6060d · outbound

This paper cites Proceedings of the Thirty-Ninth AAAI Conference on Artificial Intelligence.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Proceedings of the Thirty-Ninth AAAI Conference on Artificial Intelligence

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:30.936965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:30.936965Z digest=sha256:bc1ce9c326ab7f466badf4c7567945a47d03b1ecd3ab5e3ef791ab8effac6ead

Observation 4ae0194c-339d-47ff-bdd1-b83036f7ba99 · outbound

This paper cites Multi-Layer Visual Feature Fusion in Multimodal LLMs : Methods, Analysis, and Best Practices.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Multi-Layer Visual Feature Fusion in Multimodal LLMs : Methods, Analysis, and Best Practices

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.077641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.077641Z digest=sha256:e4eb1d59dd93063432e18d3282e3d1d09c5abcd36ee4c3772092d70a2684055f

Observation 40bbf5e9-7530-4ceb-950b-f498f214554d · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.210800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.210800Z digest=sha256:d9716a4182ff52d93851600f66e93578763baad14b90b31d79eed6e81268027d

Observation c389fac6-4e4d-4d9e-944c-cea45b386618 · outbound

This paper cites Exact Matching: Algorithms and Related Problems.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Exact Matching: Algorithms and Related Problems

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.358219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.358219Z digest=sha256:9f82a9602b1aa333fb490bdf7b8037c4590fdc2d2322e3ea592313102d3311cc

Observation b7eee086-72ee-43e4-b1d7-6ec22f130b18 · outbound

This paper cites CIDEr: Consensus-based Image Description Evaluation.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models CIDEr: Consensus-based Image Description Evaluation

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.477570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.477570Z digest=sha256:20210231bcec2df8a5386ff7da36b4cf71ec64885bd0892957c47ef49f7b8d33

Observation 71d68937-354d-4e03-811b-8792b63dca2a · outbound

This paper cites ANLS* -- A Universal Document Processing Metric for Generative Large Language Models.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models ANLS* -- A Universal Document Processing Metric for Generative Large Language Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.594074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.594074Z digest=sha256:bbaac058f690aeba78267c2652a4fcf605b7104b4064a9da33be1776dcba8ea0

Observation 6631fa7b-dd24-4a0c-9655-bf93d45333d1 · outbound

This paper cites 2026 , eprint=.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2026 , eprint=

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.690860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.690860Z digest=sha256:92cdb165521c8f81b21b52a6458a51387b43798e7a1e518e2c37b3ec5e593c53

Observation bcc1a2f5-baa6-4f6d-bce4-865a6f4ac7a9 · outbound

This paper cites 2026 , eprint=.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2026 , eprint=

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.854901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.854901Z digest=sha256:7298cbb33c04105d9d41a3b2d67759949331e25ded08809d806f1ba93a85dd16

Observation adebac80-f026-4807-85a5-c17826c9a6a9 · outbound

This paper cites 2026 , eprint=.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models 2026 , eprint=

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:31.995756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:31.995756Z digest=sha256:6626b07845ff1cbadf23d0ef936042a438e3045286a70660238e40cbcf74ef70

Observation eade1ae3-54f1-4bb1-b807-0b5d172e1d6e · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , series =.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models Proceedings of the 42nd International Conference on Machine Learning , series =

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:32.095426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:32.095426Z digest=sha256:57b04488d7d1d9f39afd8c6734653ab8f2cc383f6dab4964d8d41c6f65bd1c75

Pith citing papers

No inbound Pith citation observations are available.