Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:21:52.377583Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 4 inbound Pith citation observations for arXiv:2412.19509.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:21:52.377583Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T19:12:56.131001Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:18:37.366607Z
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2ccd005e-c2e7-4511-bd86-d3ef88c200f8 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Flamingo: a visual language model for few-shot learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8f2e767-f845-4a23-854d-5852311f48d1 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Vqa: Visual question answering
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af04e1b6-f081-448d-9815-fc87e0416989 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78b32ce7-ef0b-4309-9ad8-d0918eff8080 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a62efcc-4947-44b1-9d88-508cdefd994f · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models PaliGemma: A versatile 3B VLM for transfer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18e799cc-3e96-4d80-bad7-3be4a4b751a9 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Madtp: Multi- modal alignment-guided dynamic token pruning for accel- erating vision-language transformer
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2f70b025-08aa-45a6-86e5-f71421afb842 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Quip: 2-bit quantization of large language models with guarantees
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4649ac03-06bf-4d4e-a620-172b4beb1571 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ff507a7-7fdc-46ce-9903-dc28985ad3b3 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ab70f971-27a1-43e3-847c-74cbf10be103 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fa41131-0df4-46ab-997e-05de94b2030e · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 87bd6ebe-a169-4823-a15a-41161d57b92b · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0527c9fd-9b03-416e-ba9c-4a9ca20110e1 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a81c7215-b073-4b70-a25e-9b81be08aff8 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models SpQR: A Sparse-Quantized Representation for Near-Lossless LLM Weight Compression
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c08c9cf-1851-4048-a4b3-a95242fd1e98 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f9b97e6-56b7-4c7a-9ed8-26436026e99c · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Moa: Mixture of sparse attention for automatic large language model compression.arXiv preprint arXiv:2406.14909, 2024
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2f72d89-ac3b-4c8f-8360-2ca3d5b06f43 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc66d3c8-698b-4521-a5a1-ea89041de2cb · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Vizwiz grand challenge: Answering visual questions from blind people
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80130e5f-f41b-416d-aeaa-6b63ddd9dc19 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 298ad498-c48f-42f5-9415-68f218437624 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Llava-next: Stronger llms supercharge multimodal capa- bilities in the wild, 2024
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 00ee3d80-6cce-4d23-a9b5-4a1f09edcb9d · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models LLaVA-OneVision: Easy Visual Task Transfer
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a71ab9fc-5a5e-4ef9-862b-b3fea50e9344 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5492805a-4c72-4c41-aa82-6a7a0658dcf1 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Llm-mq: Mixed-precision quantiza- tion for efficient llm deployment
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7e8fcf70-d007-408d-8ef1-4d849fa730ed · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Evaluating quantized large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0e393eaf-cb02-4872-bcfd-40bf8cff4b9d · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457cc0aa-c337-4925-99da-fc2f404f3bf1 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb250ba1-e5e7-4766-9cd1-548bb4100176 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Awq: Activation-aware weight quantization for on-device llm compression and acceleration
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e5e05113-4ff1-48e2-8401-565607cab953 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Vila: On pre-training for vi- sual language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1248ba68-d2f8-48eb-9cc1-147a40f1b0af · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e31ec82e-f815-4370-bf47-b7610183a5ff · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6f2fe14-8d04-457a-83c6-f411447120fc · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Visual instruction tuning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f6438b1d-8b42-40fa-bbe1-d64fc72be943 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Ocrbench: On the hidden mystery of ocr in large multimodal models, 2024
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b7552357-7818-412b-ae44-394082910cfa · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models SpinQuant: LLM quantization with learned rotations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e2d7be-32a6-4958-a20e-3651839a077c · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc5e0330-9cfa-46ee-9090-0804226388ec · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 999bfc33-81c3-4292-91f9-62db3437da67 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Skeleton-of-thought: Prompting llms for efficient parallel generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 87dfa7f9-5307-44c3-b52d-374d49bd534e · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models VL-Mamba: Exploring State Space Models for Multimodal Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f595520b-1b07-41a3-a7cf-a5541f737394 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Llava-prumerge: Adaptive token reduction for efficient large multimodal models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7a20f75-d9d6-46fa-9a54-149ee760f6ab · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec444c70-24b4-43c2-a44e-a7bb796c0678 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Towards vqa models that can read
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 84f6cc41-98bc-4567-8db8-f5aa81e0f34e · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63b7a57d-0ecd-4844-8583-3ab40b3369d3 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Attention is all you need
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0999704a-26e5-418d-9632-184d2537b9a9 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Smoothquant: Accurate and effi- cient post-training quantization for large language models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bb5e6cbd-ce04-42d2-8e04-52df6428f9a5 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Efficient Streaming Language Models with Attention Sinks
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca19e749-1942-4d8e-87fb-4d3f901b9cce · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e13bac50-265a-49da-8a76-8133c812febf · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models ZeroQuant-V2: Exploring Post-training Quantization in LLMs from Comprehensive Study to Low Rank Compensation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1f7cbd6-4511-47d1-a7eb-990d43074858 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Mmmu: A massive multi-discipline mul- timodal understanding and reasoning benchmark for expert agi
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 92f982ef-741a-470e-9b61-2fe9c039edb3 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Sigmoid loss for language image pre-training,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95bd41c-7765-48b4-953f-85419d44d3b8 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Lmms- eval: Reality check on the evaluation of large multimodal models, 2024
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3af66ba1-ac2d-419c-b818-cf3e37abc2ac · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Debi- asing multimodal large language models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 13e58d4b-495e-40be-ab0f-72f490fca7d7 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models H2o: Heavy-hitter ora- cle for efficient generative inference of large language mod- els
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7b56f215-0911-45c3-8102-17d9289bdd3a · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa11f565-5bc6-4306-853e-cbbd9ca7db2e · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Atom: Low-bit quantization for efficient and accurate llm serving
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 968423d1-a085-480d-974d-62db0ea1534b · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models A Survey on Efficient Inference for Large Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a67f440b-5e8c-4cb4-89e8-23d5b4647d28 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Llava-phi: Efficient multi-modal as- sistant with small language model, 2024
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1b8a54a9-3e88-4268-9074-7d3abfb44a92 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models The Inference Process of VLMs The inference process of VLMs is shown in Fig
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d82b6987-bed0-405f-b22f-4d2aff834e60 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models LLM Quantization Post-Training Quantization (PTQ) techniques are widely used in LLMs to accelerate the inference process
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2b75c01d-38d6-41ae-84f3-e19e94dd071d · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models W4A16 and W8A8 Results on Large VLMs As shown in Tab
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 87b1608f-7d4c-4052-be24-b5c0131895ec · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 50a5ff3d-43dd-42e9-bfb1-2d8e8d236efd · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 30f51fb5-687d-40c2-b914-3f0bbb37c015 · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 94d6c4e6-3ad1-4783-a0cb-4fb8dcde253c · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 21771865-8bd4-4057-b2af-4dd4a109a19f · outbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models No Output
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ed410a1c-caa0-4802-9fcd-db2d6b9c6bf6 · inbound
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization MBQ: Modality-Balanced Quantization for Large Vision-Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a21d352c-9f33-497b-9c6f-b426317f3ed6 · inbound
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation MBQ: Modality-Balanced Quantization for Large Vision-Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 96b75b99-8981-44aa-b58e-d14b31774a73 · inbound
LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models MBQ: Modality-Balanced Quantization for Large Vision-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 95409af6-cc6c-4d04-8135-01f72d2376a3 · inbound
SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models MBQ: Modality-Balanced Quantization for Large Vision-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.