Pith. sign in

Paper Citation Record · LEDGER

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study

As of 15 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2507.20749.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20749 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:22:48.676225Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved52
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cde1480c-a661-4101-87e6-a7b4054270e9 · outbound

This paper cites SliceGPT: Compress Large Language Models by Deleting Rows and Columns.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study SliceGPT: Compress Large Language Models by Deleting Rows and Columns

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.042601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.042601Z digest=sha256:dbcbdf8f040cf014344c0c7233981dcd58731bf576e4b03b59dee238bcb5a50d

Observation 71696497-427a-4cb6-b9f3-d267c42b3438 · outbound

This paper cites BinaryBERT: Pushing the Limit of BERT Quantization.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study BinaryBERT: Pushing the Limit of BERT Quantization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.098152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.098152Z digest=sha256:73512c8d4c1a32c09db72e0e9e92eb12a667031c9c74e553b3db5ab1a0bc84c9

Observation b1f1ec04-f4bf-424a-afd5-c1f0e236a85a · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:22:52.182974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:44.158846Z digest=sha256:795cc7525864117f8a60cd325fc9d85d285717544c27f20e547b2c255044486f

Observation 9a3a4765-f16a-48f8-9a9c-b7f8a7db46db · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.247960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.247960Z digest=sha256:1a4b2668e44a18fe42613b758a031b894fb37da9326947176538f5d88e8a0810

Observation 89b2e58c-ab52-4e06-b7ff-523d1e445c2c · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.288220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.288220Z digest=sha256:b91d889fc268e7e14825a729f67a469e79b47d71c0918989b0b56daa7708c59c

Observation e5f32ebe-0e91-4a1b-ad90-061213ed3def · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:22:52.009137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:44.382454Z digest=sha256:3bfa75498f80dcc4ceb79d4b65d27f8c5c4419704800ac8b0ee4124dc2a5c72e

Observation 8544ec58-e517-4bfc-97f5-7b8cb41c1d5c · outbound

This paper cites MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.438590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.438590Z digest=sha256:3677a02ecadd075c70e17cdf4f09800918317f5484856b3cc3e5e579a1189bcc

Observation 504bbf55-2121-414f-a9e0-ee8cb76bbafd · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.503133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.503133Z digest=sha256:e17be4bd248fafe352df94608e9d9202a804690d6d72eb2043e93bdf9b589329

Observation d5878f7d-8e4d-4ae7-8ef7-abbf0b611e3d · outbound

This paper cites int8 (): 8-bit matrix multiplicationfortransformersatscale.AdvancesinNeuralInformationProcessing Systems 35, 30318–30332 (2022).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study int8 (): 8-bit matrix multiplicationfortransformersatscale.AdvancesinNeuralInformationProcessing Systems 35, 30318–30332 (2022)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:51.820101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:44.570859Z digest=sha256:76649aa154cbc6ab5baa1fb249e75988dd9fd9602ae36088b2eaad8dc93642de

Observation acf926ff-1937-477c-8b0c-130bef12dffc · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:22:51.695570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:44.621053Z digest=sha256:0530f65bbc20c1335086a2fc249036d30e4409aadefb12feca65e53dda0a252a

Observation 4290d994-172a-4b45-b8f7-189b17ba2e12 · outbound

This paper cites Learning to Prune Deep Neural Networks via Layer-wise Optimal Brain Surgeon.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Learning to Prune Deep Neural Networks via Layer-wise Optimal Brain Surgeon

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.661615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.661615Z digest=sha256:b0790d930718978b7e60da2b8e5a7c04b23b0c213b3e0f9a3244d410b3e471f6

Observation 3776073e-b025-4559-a80c-f2d61610c801 · outbound

This paper cites Reducing Transformer Depth on Demand with Structured Dropout.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Reducing Transformer Depth on Demand with Structured Dropout

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.726762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.726762Z digest=sha256:b546736795e555c58459ecb337de3957eb030b63318a6dfc2e4540645a717c1f

Observation b314f67b-cb0a-4bf8-856c-e1dc8b7a240f · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.801416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.801416Z digest=sha256:fa619d704962e86b6992a11d804cd6a63fd1225adf4fa19781b6e588e591f884

Observation 16654f6e-d277-4a49-b6ac-3de1ccce50d8 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:51.495830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:44.881891Z digest=sha256:36ada3cc2107dcb76df4ff8c81c3c9fb6cdd7452058a5da2123d3a5476ba7167

Observation 1cc44439-724d-4e5c-843c-84e2f4d2c5cf · outbound

This paper cites The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.947753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.947753Z digest=sha256:b126a4f972b701e7586130080b5b6101123a45efb9c328705a15b5a373aef35d

Observation d07fa3d9-2024-433d-9135-a580bfdec58b · outbound

This paper cites International Journal of Computer Vision 129(6), 1789–1819 (Mar Pruning and Recovery Techniques for Compressing MLLMs 15 2021).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study International Journal of Computer Vision 129(6), 1789–1819 (Mar Pruning and Recovery Techniques for Compressing MLLMs 15 2021)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:44.982668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:44.982668Z digest=sha256:2332e895f3225a7c32c621a28519d92fb9b578090f21dd8ff14d5268d1cffe89

Observation 0e0e02db-332f-41d8-8a0e-2cc44286cb70 · outbound

This paper cites MiniLLM: On-Policy Distillation of Large Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study MiniLLM: On-Policy Distillation of Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.128888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.128888Z digest=sha256:38b7c57e61753640b9f60f68b9d555dcacc28b9ed2c9562ea6f8fb4a8e4a0398

Observation e6e16ed1-ea5d-45d0-a58b-b8b29de3acce · outbound

This paper cites Efficient Multimodal Learning from Data-centric Perspective.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Efficient Multimodal Learning from Data-centric Perspective

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.165347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.165347Z digest=sha256:20f4a34f4329248635518b5d814e0f95c663f442f74fc16defc8745904521bc3

Observation ef1e9ca0-ff73-46ea-91bc-b8c8248b5a39 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Distilling the Knowledge in a Neural Network

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.229918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.229918Z digest=sha256:fbe7df70896c137201f84904564cc0582f0c37fe6fb925b7ff9618f2c8ec3b35

Observation 7d23b2a2-47e7-413f-bf93-54a9a47512fe · outbound

This paper cites The Curious Case of Neural Text Degeneration.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study The Curious Case of Neural Text Degeneration

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.286552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.286552Z digest=sha256:bd5bfcb57079b138149fd1cabbae4c023abdd21802ce89318a0735e5a19ed8e9

Observation a63f0888-6f7e-4bd3-9843-6e4532d24289 · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:22:51.294215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:45.323392Z digest=sha256:6bfdb300f720147d52c5c7d61b565918c8bff54844a177b8c560d889cad3b821

Observation aa152e63-2d7a-48d5-94bd-6ee8740a0077 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study LoRA: Low-Rank Adaptation of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.474761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.474761Z digest=sha256:d357ef080c4af1164f1167e5540cb7c96bc69a221328c94d1e0e762408e681ba

Observation 4b02ecb8-5706-4f78-9280-1de82ff192ef · outbound

This paper cites In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.511186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.511186Z digest=sha256:978e307b12b565f25937efde63d9611c89dbed25f22805bb04982fb2c1326bd9

Observation 7a41c6bc-6b1f-46ea-abd1-4542610f4d0a · outbound

This paper cites Microsoft Research Blog (2023).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Microsoft Research Blog (2023)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:51.106896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:45.560083Z digest=sha256:4eacfbcda7cdcd2c04ec75c840b982c793ead563364c841ff4e0b48f725d0775

Observation 7945b797-6fe1-48bc-a8cf-aaa36466f759 · outbound

This paper cites Master’s thesis, University of Washington (2024).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Master’s thesis, University of Washington (2024)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:50.902477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:45.621599Z digest=sha256:3bcde5880cee6ba869ec2e8ede50fb5834057eb4f027a4f0f408d65589cd38cb

Observation 1830bc6a-bdb2-4108-8f5a-0332b13c47ad · outbound

This paper cites TinyBERT: Distilling BERT for Natural Language Understanding.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study TinyBERT: Distilling BERT for Natural Language Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.689017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.689017Z digest=sha256:e2c984d657bffd2dabb07f27eeb1c336cadd79487430bed7c7244e9048020876

Observation 97d2059f-e96f-406e-9738-42bc79a9939b · outbound

This paper cites Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.746187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.746187Z digest=sha256:3073a9b80a9c50e8d8727f3347541cc6b5476b94a3b954ecdf1a8e937be7a351

Observation 9062c940-8f27-491d-b23c-0817595dff84 · outbound

This paper cites ALBERT: A Lite BERT for Self-supervised Learning of Language Representations.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study ALBERT: A Lite BERT for Self-supervised Learning of Language Representations

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.812030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.812030Z digest=sha256:91651aa3818ee8757c23746821cec4ea3e775a66eced8f8230abc26cef8b8127

Observation adc4c753-1de6-4343-a4ba-4d565c7aadda · outbound

This paper cites A Signal Propagation Perspective for Pruning Neural Networks at Initialization.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study A Signal Propagation Perspective for Pruning Neural Networks at Initialization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.868130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.868130Z digest=sha256:282c0d6ce675bbf25f295f00d6bc0b4b1c1c9ddbdb8c9b910ec5580885863559

Observation 7b9e497b-40bc-4f95-bab5-1fcd9783b7ea · outbound

This paper cites Pruning Filters for Efficient ConvNets.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Pruning Filters for Efficient ConvNets

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:45.937930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:45.937930Z digest=sha256:19cfe23c931be6baa7210664902c22bd85cf3192149fbe52acbc49b94730fc38

Observation bc4072fa-b7b5-49dd-ad2e-37d41680e44c · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Evaluating Object Hallucination in Large Vision-Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.010798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.010798Z digest=sha256:5833dac30994a18492856a8d89a01061bd71086f88a190f48b6bed96f07494a2

Observation 813551f2-7d42-4ace-995a-2b6eefcd90d3 · outbound

This paper cites MixKD: Towards Efficient Distillation of Large-scale Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study MixKD: Towards Efficient Distillation of Large-scale Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.069815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.069815Z digest=sha256:2bf11afd17cd9574844f7ef23a4c28a825a90894df16f027901c653fbe962879

Observation b7b0fe10-9d47-40ce-a3c2-e0339c882a9c · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:50.741243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:46.130382Z digest=sha256:e37f697262aa4784b8ec4ff2ca3f625c1577d53929d30378f5653ccf99a6110f

Observation aff4ce76-5fdc-43ab-ab30-14bb3b3b6d08 · outbound

This paper cites Visual Instruction Tuning.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Visual Instruction Tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.228942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.228942Z digest=sha256:79544b45b49c1d5463672fc3b84484d6cf282ab4a56017d6536b97130425cad5

Observation e404fb81-d970-468a-afcb-4cac69586276 · outbound

This paper cites Group Fisher Pruning for Practical Network Compression.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Group Fisher Pruning for Practical Network Compression

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.284667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.284667Z digest=sha256:e84bfe8df442c950a272c9d8f13d8078e24eea401ccae0f3995105874b121986

Observation 41f59a31-6279-4d4c-a070-0af503067cef · outbound

This paper cites Advances in Neural Information Processing Systems 35, 2507–2521 (2022).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Advances in Neural Information Processing Systems 35, 2507–2521 (2022)

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.381619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.381619Z digest=sha256:22932f0827ecf2228dfa62b5eaaba31046b02f6f3cfad1d0877d1ac6f2e29363

Observation 0ae59497-55cb-4f04-b6fe-ccc62c01cbca · outbound

This paper cites Advances in neural information processing systems36, 21702–21720 (2023).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Advances in neural information processing systems36, 21702–21720 (2023)

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.487160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.487160Z digest=sha256:4a39a80285b9af3f2e13f1c3dd7c427d8f916ab39e8145c6e94021b88a10ee02

Observation 49e3e980-9831-4937-9fd4-aa83b45533c7 · outbound

This paper cites Structured Pruning of a BERT-based Question Answering Model.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Structured Pruning of a BERT-based Question Answering Model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.547975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.547975Z digest=sha256:212b386f10eb5f21fe4ee8ae3462c31a2e72bf2a3831aeeb0b108162c12dff94

Observation dca33805-b5df-47e6-8ef6-65a740e9008f · outbound

This paper cites ShortGPT: Layers in Large Language Models are More Redundant Than You Expect.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study ShortGPT: Layers in Large Language Models are More Redundant Than You Expect

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.639736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.639736Z digest=sha256:cef9001b1758c1a5cae97d9993dbf6af350d38eda2b5f08a73aa3d4b5dbd13a6

Observation 2a2b67f8-9578-49d2-ba24-4212364c4287 · outbound

This paper cites an unresolved cited work.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:22:50.561751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:46.725744Z digest=sha256:185be0e1e080c044963fdcf0bee85b6a7c3dc2a728d90227d4e01fd50714af8c

Observation b162db12-7f10-4949-8a80-8ccd808f1a45 · outbound

This paper cites Lookahead: A Far-Sighted Alternative of Magnitude-based Pruning.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Lookahead: A Far-Sighted Alternative of Magnitude-based Pruning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.791956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.791956Z digest=sha256:e1cb7001b37b05105628ce6efa63c1de3a1040c5c4a51758e61477300befac48

Observation 80e2e3bf-68dc-4bd1-90e9-bdb2235a8f95 · outbound

This paper cites Zero-Shot Distillation for Image Encoders: How to Make Effective Use of Synthetic Data.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Zero-Shot Distillation for Image Encoders: How to Make Effective Use of Synthetic Data

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.867030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.867030Z digest=sha256:a48817c220bb6fb42e4bbd45203cbaf04573e7e3489134c4dd233dfc41b3be4c

Observation 7062b5e9-312c-44b8-b478-76755a4a6203 · outbound

This paper cites In: International conference on machine learning.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: International conference on machine learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.889339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.889339Z digest=sha256:e7086fa2a59fb367d61070b8422fbdf8c3d2e0957b6bcdcb77bd0d7d9672da51

Observation dd142509-ef06-474c-913c-6061df413d8d · outbound

This paper cites Computer Speech & Language77, 101429 (2023).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Computer Speech & Language77, 101429 (2023)

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:46.953068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:46.953068Z digest=sha256:d5e12c55091046cd9d162cae7de295c1f2a451f49028cf82733bf6520fd8ad40

Observation 6d7b8da9-6408-42dd-9ddb-2559ec0adbc0 · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.025758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.025758Z digest=sha256:91ef7de7836c9fc650d695b023709200ce2efae292ec34b5077ff607ea68cb31

Observation 15bff839-5809-4602-a976-2975f142d3a1 · outbound

This paper cites Movement Pruning: Adaptive Sparsity by Fine-Tuning.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Movement Pruning: Adaptive Sparsity by Fine-Tuning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.097104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.097104Z digest=sha256:24c13e9dc6bcc3c5638a0c1bdc53b63e3b19d22fe3448e5d2d2a466cf62057bf

Observation 427cc078-087e-4d6b-a1d0-186875127e40 · outbound

This paper cites Patient Knowledge Distillation for BERT Model Compression.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Patient Knowledge Distillation for BERT Model Compression

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.217308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.217308Z digest=sha256:29d843b6389d8b6ee5bc51c270f8c05287b46415556afa424226f3d202b5f268

Observation 8888c43e-686e-49a3-bf51-2325d029d1f9 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study LLaMA: Open and Efficient Foundation Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.333939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.333939Z digest=sha256:cf4a89cf766d74e359cab8c48e2b71a05a3800f75b42f500418a66e158b74eeb

Observation a3f4a32f-4f2d-4e7a-afd8-1b776e1b54f7 · outbound

This paper cites Analyzing Multi-Head Self-Attention: Specialized Heads Do the Heavy Lifting, the Rest Can Be Pruned.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Analyzing Multi-Head Self-Attention: Specialized Heads Do the Heavy Lifting, the Rest Can Be Pruned

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.423140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.423140Z digest=sha256:ab526593b421a3fc6e5d7655239bf60edbd5e47de8acf89343b8a2ab8a4de718

Observation 899eaffc-bd30-49da-9ac8-13f86b13e5e4 · outbound

This paper cites MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.515023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.515023Z digest=sha256:3308acefd432f3c66e3ec4c279291dd639948235a8f20468e70f4866a9a1c97d

Observation 0d6e61e8-845d-43aa-9d8f-5c0e98c4af06 · outbound

This paper cites Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.582094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.582094Z digest=sha256:f397e27950c41949048287f2cc939502d776bc820404a5390a028d1ce973f66f

Observation e3447fd5-6547-4b55-a2c3-f0b007303cf5 · outbound

This paper cites A Survey on Knowledge Distillation of Large Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study A Survey on Knowledge Distillation of Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.656391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.656391Z digest=sha256:4dcb5c8481e623cded1e06e778455d12c532a41a367895c79fb94f60bdc73e70

Observation b44a2ab7-f7f3-46ea-bb8c-342c9ae1d263 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:50.309347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:47.717825Z digest=sha256:db4feb2b7403237de5eeae047720ca0c96bdf9a7a170ac12396682ba677c9b3f

Observation 934d023b-00d6-45e5-a64e-f1d9d38e76e3 · outbound

This paper cites ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.848505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.848505Z digest=sha256:0cad2fc17ab0d3c339e1d96d5ce34cc30557c1daa28211ea085e1b2f485c9e25

Observation 3d57b322-8fa6-4449-aee6-d7525aa7edd1 · outbound

This paper cites A Survey on Multimodal Large Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study A Survey on Multimodal Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:47.959885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:47.959885Z digest=sha256:3c40579e1dddc23d1484e83a00920436a5ae3d7ea8afca80fdd9dc6d51f09ee7

Observation 4fcd6331-8720-4d8d-9447-6aee4685d197 · outbound

This paper cites Gate Decorator: Global Filter Pruning Method for Accelerating Deep Convolutional Neural Networks.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Gate Decorator: Global Filter Pruning Method for Accelerating Deep Convolutional Neural Networks

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:48.083495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:48.083495Z digest=sha256:8783bfc7a034e015881a02d35fcd92aa2ee95c4617b03176521a1e338d0753b0

Observation 39b8d7ee-30b3-49b8-b505-406f390d2741 · outbound

This paper cites In: Proceedings of CVPR (2024).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of CVPR (2024)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:50.120328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:48.198110Z digest=sha256:19716e74df05a4097b6253308c8b375d01003571e7c0635a226ccaa453f256c9

Observation 72696593-63a9-4078-9549-9df1553f1be8 · outbound

This paper cites In: 2019 Fifth Workshop on Energy Efficient Machine Learn- ing and Cognitive Computing - NeurIPS Edition (EMC2-NIPS).

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: 2019 Fifth Workshop on Energy Efficient Machine Learn- ing and Cognitive Computing - NeurIPS Edition (EMC2-NIPS)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:48.316422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:48.316422Z digest=sha256:bee6f5753b866f69cdd404321c5fcf4c835d2eccb8e507aca7e954d1f624062a

Observation 1c8c4468-5aff-456f-869d-c32b64457647 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:22:49.965938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:48.404334Z digest=sha256:faccb0304d434278b48ad8834581624389755d6a97b553974bb3264e910de980

Observation 80e71a35-495e-4fb4-979e-48df0d34a593 · outbound

This paper cites Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T13:22:48.525847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:22:48.525847Z digest=sha256:5920a0d9716d6f7d9ffcd6a91d2fafd6361e453979e0cb3ffa8124cdabbc1413

Observation cefbfc6f-6959-4fda-a4a6-fbf15b377446 · outbound

This paper cites In: Proceedings of the 1st International Workshop on Efficient Multimedia Computing under Limited.

Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study In: Proceedings of the 1st International Workshop on Efficient Multimedia Computing under Limited

Reference 63

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T13:22:49.766910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T13:22:48.676225Z digest=sha256:1d5c1a96f7c24a0f8c99545eb8d7d8f6c6d1d11d3ebbae6f48283f056453d202

Pith citing papers

No inbound Pith citation observations are available.