Pith. sign in

Paper Citation Record · LEDGER

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models

As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2506.13166.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13166 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:42:53.748135Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T08:36:15.783708Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1937c3fe-d399-4306-8c5a-d8a909b7d97c · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.341804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.583031Z digest=sha256:db9c1b318dba14cc7f3a32649943b9508c8175d29321171606a29bd462191f6f

Observation b7a2c712-d4dd-404c-aaca-61c8123a70ec · outbound

This paper cites DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.586485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.586485Z digest=sha256:6fec304522c9df86641ba617efe2707c645443e4c845257d8a2eed6766fd714c

Observation 25f712e8-e4a4-4b03-94a6-4033a5b2d9e5 · outbound

This paper cites HiRED: Attention-Guided Token Dropping for Efficient Inference of High-Resolution Vision-Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models HiRED: Attention-Guided Token Dropping for Efficient Inference of High-Resolution Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.590496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.590496Z digest=sha256:7967273eb5c9bff14c4415fa9d001b85c2de9663fccfff77e6ace56ef6b645b5

Observation cb2f33b1-9778-4dd2-bdce-2e20ed9ffc9d · outbound

This paper cites Qwen2.5-VL Technical Report.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.593771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.593771Z digest=sha256:842bfff8e66f56cae9b7261f00ca24a95293d918187276bcc3936b080c665c78

Observation b6af6db1-4955-41d8-bb7d-62ddb7bbf1a5 · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.333814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.596977Z digest=sha256:3eb0a9c6d5bb15ac680113b7873e2ca82f3dbf1dbc9e0c309af348c80d9f855b

Observation a006d69a-51f8-4768-aba8-2d08ec201e87 · outbound

This paper cites Token Merging: Your ViT But Faster.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Token Merging: Your ViT But Faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.600247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.600247Z digest=sha256:6b6254109f8d2d77480d0ffc8e8fda41107fb60cf0f0524276a591a7cf10ee1d

Observation 28c77754-5715-4cfa-a80a-afb162c9921b · outbound

This paper cites Efficient Large Multi-modal Models via Visual Context Compression.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Efficient Large Multi-modal Models via Visual Context Compression

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.603924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.603924Z digest=sha256:e8052a0339d7822bf48113ed678bb4c81b9731b1d68ead2260d0fdeafc1ef011

Observation c4466222-5db8-4ae8-bf44-17d01ce99564 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.607094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.607094Z digest=sha256:3855b84b010ddd29d7eb5ba6bc266a3b734fcc4d7f988926aed61d6dcbda9e55

Observation 6dc287d1-f569-467c-a16f-9564fbd18c74 · outbound

This paper cites An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.610866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.610866Z digest=sha256:e27a218aa43b41d5dec86be63fd50b68e8932c08c259a5e56d131a4b3adf6225

Observation f8b1262f-5c77-43e0-b56e-ef471deb17f0 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.613762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.613762Z digest=sha256:7ff67df0b2fd58423df7e9679652553f630f157d91940d86d5cf5b0c2b9e7bb4

Observation 86f8dce3-1e00-4012-bb75-02b7c0513134 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.617098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.617098Z digest=sha256:57e3553b1061bd0d547b00e3ecd3deade4244bf2bc2fc05d7de013a3bef279b9

Observation b77225b5-e6dc-408b-9b9e-a14f9a75ec30 · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.326120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.619924Z digest=sha256:fc72eae7bf3b2e97c7685ec04ac946b12a7724aaae2badb8958270c5b8364ac5

Observation 5736cb6a-0961-40dc-87df-3b4d5537d474 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.622519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.622519Z digest=sha256:f53b277ccd570728437186498b2dac424d683a7dc9b296ae5997094812fad344

Observation e47cb805-1a8b-4e5e-86a5-f734ea296524 · outbound

This paper cites Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.625121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.625121Z digest=sha256:9e1e78993ef7cb9125ec137299e5fb46accaefc2de6a3753e1786c4feba8388e

Observation 31aca71c-f40a-4e95-a2d4-9898cc5aa111 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.628108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.628108Z digest=sha256:4cf2fb84ac8cff3a959b6a738981a6eebb18ec9e6f5c1af9c15bfa1879aa2edc

Observation 13d89cbc-3c5e-4e48-b36c-be6f947ea3dd · outbound

This paper cites On Speculative Decoding for Multimodal Large Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models On Speculative Decoding for Multimodal Large Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:42:54.131638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.630964Z digest=sha256:4e9b4bba285cd49e08eaf0b83fb748e66942797cf39c38a09d307c1d109391ce

Observation 4903a790-14c1-4a4a-9ef7-034902d90dfc · outbound

This paper cites AdaFV: Rethinking of Visual-Language alignment for VLM acceleration.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models AdaFV: Rethinking of Visual-Language alignment for VLM acceleration

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.633999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.633999Z digest=sha256:2d21abdee39da5a012cee189c97a8775f34ec420ebf5b4e63c2e73899979e449

Observation b27d5629-613c-475d-8d83-4df6dcf03fc1 · outbound

This paper cites A.; and Manning, C.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models A.; and Manning, C

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:54.317632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.636916Z digest=sha256:651a6a7246fc9902de30db963db62b8f3c0f66af206a444d6acbb607ff10560c

Observation 211f7554-d804-4ebd-a940-a28672649b2b · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.640084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.640084Z digest=sha256:e807521cf2255b7f0b2f3177b3a591a075539b64ceba2a925eb8b1b766249287

Observation 4377c1ce-e8cd-4a1c-b578-336be540c6ea · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.309726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.643357Z digest=sha256:1ed89a1fdaca4a7f7e6a42a6c2c14cf638ae9f90d6aad6d56b73c6d2e28e405a

Observation b87afc1c-3f1c-41d0-9cba-d4e7caf476b7 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.646688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.646688Z digest=sha256:ee51ff97e8625d1ffab0d1342ad399020ada963bdd887841d7fd7c0ce5e5fb25

Observation 5395214e-ad08-4edd-b7fd-34f8c487b2ad · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.300470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.649968Z digest=sha256:ad33def821b3bab2d9bac989725a3c4fdbf392b5298f40336540ad3e9fd21380

Observation d7c2a526-d729-4429-b403-e74d6e6bf84a · outbound

This paper cites Dynamic Token Reduction during Generation for Vision Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Dynamic Token Reduction during Generation for Vision Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.653335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.653335Z digest=sha256:bec3196eaab10f3f21f7f888ce03d5c902c9dec840b199bb7736f231110a1861

Observation 2aa64794-545e-4a05-a1dd-69b1fc00e808 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.656785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.656785Z digest=sha256:bb191b6b5d6ce317f5c89dc96cf7679fab331127c3ff94d69f4ae13bc2b9919e

Observation b00a0740-01fd-48c8-a130-bbcfc683fd6c · outbound

This paper cites Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.659544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.659544Z digest=sha256:ec59e02b3d68d84f6d7960d94122b6ad00d4556c45b42351f8d2f8301fa1e92a

Observation d903c29d-950d-4f37-b42d-5ca3edd2571f · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.290929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.662428Z digest=sha256:adca7fc45f219326c7d081efe0bba72bfbccbda0382aa19e87b3de21eeaa3f8a

Observation d592996a-6d66-4233-b2b1-4689b6da923e · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.280890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.665273Z digest=sha256:d45a47e5cb2c7185d82adb63bf80eefdb09f4cd756a7e34745610c5c1944eedf

Observation 762b43ab-b672-4b1b-b4ac-cbd32c5250f7 · outbound

This paper cites Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.668156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.668156Z digest=sha256:d3bff9aa2e9d20f0d073e7423c5fb06be94d76d91b82e293e045dc756e7b1258

Observation 5c3e1455-e133-416e-b8ee-00db0d9c04c0 · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player?.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models MMBench: Is Your Multi-modal Model an All-around Player?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.670940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.670940Z digest=sha256:76798381ef20ee29437e66240854403e92a3897f3ce2f6c84e8dde03d958200f

Observation 68bbc9ad-3ecf-4f7f-897f-e87bf62e9ca5 · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.270686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.673780Z digest=sha256:2a5bd47261647c976e76fb9cbbadd0ff0c351eb97ba936c0eabd71f76d2247f5

Observation e2a34564-623d-42b2-9d2e-aac72c0e50f1 · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.261322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.677586Z digest=sha256:169616b44888941097eb21f95fe4b70f04a7576182a01da00eceeff520152062

Observation ef92b47f-c261-402f-b269-3b1485509843 · outbound

This paper cites PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.680275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.680275Z digest=sha256:e2feb7ff8f94cf3807e1229580c58814a9fee1c7fda0cb1a957080b57b354009

Observation ec33be87-522c-480f-932f-c7ba9307b7fb · outbound

This paper cites Cross-Self KV Cache Pruning for Efficient Vision-Language Inference.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Cross-Self KV Cache Pruning for Efficient Vision-Language Inference

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.683212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.683212Z digest=sha256:66403167685b5a694d65cd09eef7c5548edec6a6871bc9efc41dcd7c125e192a

Observation d1e6a688-4d00-4f77-913b-2b58ee3d0394 · outbound

This paper cites Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.686184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.686184Z digest=sha256:8a7ddaee7fd399202a5eeef1118260367e73e0def4d7574f097105ff62ff2e68

Observation 96e55dfa-522f-4d73-8f1a-764ff736e1fd · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.252955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.689290Z digest=sha256:069195f84e02c84b7d44fc479ee4bfbcbc3d07cbbbe722543ba7142b1754343c

Observation 1c913274-7dc2-43a5-8a5b-cfd530c8c3c0 · outbound

This paper cites J.; and Yan, Y.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models J.; and Yan, Y

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.691986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.691986Z digest=sha256:f0dd9d736f9aab26603ed79d8dd8614ff89584e3039c308f75178e1389c327a8

Observation b1584c28-ad25-463e-8288-d5675992814b · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:54.243883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T00:42:53.694891Z digest=sha256:820b048ffaf39af4ef48bed783798c8c79ae2b27a59f4ea49aa80addafd17349

Observation c0d316f9-e94d-480b-ace7-8ddb0eb2d774 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.697593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.697593Z digest=sha256:453f94fd8d3278613724c282d99e2a44661ce2176e3ab10dbe72908e85b496a2

Observation fc21fc7c-29c3-408e-acfa-5ccb1351c664 · outbound

This paper cites Attention Is All You Need.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Attention Is All You Need

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.700721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.700721Z digest=sha256:6ef314a357bc6d55e4268f76376260c260257bd7dc3682a44741ef7f5b3e14f4

Observation bdab7b9c-11ef-4e1c-bffd-77bb99cddea0 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.706996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.706996Z digest=sha256:97d3b731fc8e10587a463e0627fbe7d276d969d40e1d42aa518e492826f5ddcc

Observation 76a83887-ed54-4d0b-b283-ee5ffa438ff3 · outbound

This paper cites Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem?.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.711824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.711824Z digest=sha256:c6ac5ab38826f4c24f3217da51015b4c3a69313186ebdbe1bc2ba627d5ade5d0

Observation c9c746f1-31d2-490e-8228-fa4366114b7c · outbound

This paper cites Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.714494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.714494Z digest=sha256:7b884740bcce1cb60c083b8b59cfb783b6cc4c3bc00c9ed11f978197d92f68c2

Observation ee264910-fc3c-4278-844e-046696587501 · outbound

This paper cites PPT: Token Pruning and Pooling for Efficient Vision Transformers.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models PPT: Token Pruning and Pooling for Efficient Vision Transformers

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.717242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.717242Z digest=sha256:00a70bd2527b6499762260e7c5fe28c37f07e609b695342a950ba172d1f3caa8

Observation c932c310-bd25-4d2e-a585-1509734538bf · outbound

This paper cites PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.720397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.720397Z digest=sha256:37ac0821bd19f48f721b14ace6c901c0f9307016f79e2bd8419a10fa49ee04af

Observation de262bfa-471e-4cac-a700-7eaeb0dd31f7 · outbound

This paper cites LLAVADI: What Matters For Multimodal Large Language Models Distillation.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models LLAVADI: What Matters For Multimodal Large Language Models Distillation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.723737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.723737Z digest=sha256:c830af9333128b77f121b43da9ca1ca79a3da1ac6eda54cfb9fbb9984906e282

Observation c02edbf5-cee7-49ad-8321-415708e887bf · outbound

This paper cites an unresolved cited work.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.726470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.726470Z digest=sha256:58c6bcf06d408b42b5ba3a851a93fb0b7aa4f5f9d6eab8634e4538b6a0009bdc

Observation 1805c7fa-a03b-4a40-92f1-0c7d95361a3f · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.728937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.728937Z digest=sha256:701bf728300477a884e3e7078679bb901ff7903e4e645a03c7c593248d75b61f

Observation a1a373ab-ce2e-4b1e-8eac-7867ae7a5305 · outbound

This paper cites Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.731513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.731513Z digest=sha256:a0fa7bd57902e92ecfa7cca91e7a7ec374bee6d12e683f3821cd462eb302da34

Observation 03c0985e-2934-4751-9bbb-32103dfcca6e · outbound

This paper cites Cross-modal Information Flow in Multimodal Large Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Cross-modal Information Flow in Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.735239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.735239Z digest=sha256:7b6b2bc5f22f79eeae59cbc8a141319355d1a7fa331dffd7202b7bab2ed65861

Observation 3f8c6000-e94e-4211-804f-d21e7746ec13 · outbound

This paper cites Accelerating Multimodal Large Language Models by Searching Optimal Vision Token Reduction.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models Accelerating Multimodal Large Language Models by Searching Optimal Vision Token Reduction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.738046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.738046Z digest=sha256:12071a66963d2a7a09d8477f94fdc0f08aef32e715112bcce480417061b94070

Observation 907bf8a8-c945-432f-95bc-dd27e53298f0 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.740963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.740963Z digest=sha256:74a9ea3e26d1157ab5bd32a34a4156a4866023004350911552904e0a76e66bf5

Observation 9296db2a-e545-4242-800e-a71b07ffbaba · outbound

This paper cites , " * write output.state after.block = add.period write newline.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models , " * write output.state after.block = add.period write newline

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.744129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.744129Z digest=sha256:9203aeb37416103fa2dd60ed6d74f01edc8ce6e71887d269f8df5bd76eed7a7d

Observation 0c7ae028-da2a-4028-bf68-ba1266c43ef7 · outbound

This paper cites write newline.

GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models write newline

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:53.748135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:53.748135Z digest=sha256:3579d487c2202962f684b3e171b17475abd216d3167f4dd4faa36691b04e56c5

Pith citing papers

Observation c4f1005c-db49-4c65-b769-425c85878e97 · inbound

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs cites this paper.

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs GreedyPrune: Retenting Critical Visual Token Set for Large Vision Language Models

Reference 257

Resolution
unresolved
no resolver link, observed 2026-08-01T08:36:15.783708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:36:15.783708Z digest=sha256:81735e7fc896f8cc8bc315402f2cb32350212f51616b54cebe2bff61d42c5ab8