Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:23:18.738520Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 10 inbound Pith citation observations for arXiv:2412.05819.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:23:18.738520Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:57:32.530971Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T19:48:11.386460Z
73 of 73 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation af9cca93-691e-416b-aece-26acfe1cacb9 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Flamingo: a visual language model for few-shot learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cc3c0b1-7af7-40d9-9745-df66af626c35 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Multimodal large language models in health care: Applications, challenges, and future outlook
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8b1831cd-5dcf-436b-9ffe-2064b2d15e2b · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 977e4197-4e10-4757-80a4-3048724a4db1 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 755f1968-0778-46cb-a56c-2023e2213c83 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Introducing our multimodal models, 2023
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation aaca4f51-1d4d-4002-8de8-9af90ae780e4 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Token Merging: Your ViT But Faster
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2279f39-cdca-4968-bd6a-6f88994610fa · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd24611f-5d71-47f2-b737-b7ba2b7a7ac1 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Show, observe and tell: Attribute-driven attention model for image captioning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d50f9d2c-8fc1-4c9b-bad4-d1a90f425e0c · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Imram: Iterative matching with recur- rent attention memory for cross-modal image-text retrieval
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a1f51f80-4cd2-4477-b5aa-eaf8a52d8ad6 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db7cf982-5136-4bdd-8a21-7b8c849cdcf1 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MLLM Is a Strong Reranker: Advancing Multimodal Retrieval-augmented Generation via Knowledge-enhanced Reranking and Noise-injected Training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d066c14d-e3fa-479b-9329-c87cb9a403b4 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 181bc5c6-cede-4180-a16c-967a7f8e7cf5 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f97ebf50-acbf-4769-baee-01799f66bb91 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c115939-9158-4eea-8c30-386cea8a6f76 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Pearson correlation coefficient
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 628f96cd-d9e8-4ba3-8396-e15707789759 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs A survey on multimodal large lan- guage models for autonomous driving
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ee1115bc-874c-401e-891a-d828a6b126d5 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082a0047-1d8e-4ef8-babc-301cd1ce799f · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs HiLM-D: Enhancing MLLMs with Multi-Scale High-Resolution Details for Autonomous Driving
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0234ad18-c79d-43ed-8c49-05159676214b · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Explor- ing structured semantic prior for multi label recognition with incomplete labels
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4685981f-1e81-4143-861e-703334cb700e · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd472030-80f4-47e9-9dda-2415f5a30f8b · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs PaLM-E: An Embodied Multimodal Language Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac042213-1efa-4ed4-b1d5-dcae4239769a · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs The Llama 3 Herd of Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3aa61ae-4231-4d93-a69f-8d6a900c5b4d · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a70487e-312a-47c5-bb4c-e18ee6319381 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Anomalygpt: Detecting in- dustrial anomalies using large vision-language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5ca43d82-4d3f-4ee4-a754-c7936e043c96 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Efficient Multimodal Learning from Data-centric Perspective
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 589894f9-6786-4385-a1c4-3d6dfe505f86 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c4ac31-27be-42db-a67b-6f948b24ce9a · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1061a524-6356-4223-ac99-80a8a9316454 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Introducing idefics: An open reproduction of state-of-the-art visual language model
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 46b589e5-bec0-4a3d-bf5f-75d3dd71c1b6 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Phi-2: The surprising power of small language models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e259e76a-c326-4baf-8d8c-cce36f5f320c · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Seed-bench: Bench- marking multimodal large language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 970a86f2-3855-4ca7-af90-422022a341a3 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Llava-med: Training a large language- and-vision assistant for biomedicine in one day
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 31aff3c2-f15a-489a-b608-26a52167a965 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffa2c15f-b5a3-4d35-9ecb-a1318814a07b · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Evaluating Object Hallucination in Large Vision-Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e5d65b-674f-469c-b890-4ebdf1513676 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 506f0483-8b85-402b-8cbf-aa427a9dc19e · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5916b2a0-6577-4a96-bbd8-10eca8e57e71 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 866ee1f5-fce6-4aad-8900-51efb24b6527 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Microsoft coco: Common objects in context
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106e35d1-3e28-49cb-a613-9b7eeb4fb0bc · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Improved baselines with visual instruction tuning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 63c86595-7beb-41fc-8714-8c5862dd76cc · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Visual instruction tuning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 46d60ad2-f810-4a4f-8652-873b21524d2e · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MMBench: Is Your Multi-modal Model an All-around Player?
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e792b842-d7cc-47c0-8aee-2433f36c1769 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Least squares quantization in pcm
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b8e8b71c-995f-4f41-82e0-0258167937c2 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Robollm: Robotic vision tasks grounded on multimodal large language models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b697e2cc-3337-41b9-9c0f-d012ec4cbd5d · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b70180-7039-44b3-8fb8-d7bcb64d7eec · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Some methods for classification and anal- ysis of multivariate observations
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b123146d-63d4-4d92-8c87-a48a13b683dc · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdec5db0-8672-4016-9cd2-ebdc0d60e3f5 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs The impact of multimodal large language models on health care’s future
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0637749e-4422-4456-9a3f-32347f4afa38 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 535ad5b2-2954-46c8-93b2-702b2ad801d0 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Learning transferable visual models from natural language supervi- sion
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 28a5b9f5-0173-4799-aa3e-cfad16624250 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Designing network design spaces
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51d6f2cc-a7f1-4e0b-87cc-1a0dcf127825 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Spearman’s rank correlation coefficient
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c150cfcc-ba0c-428a-bbdb-603e35be8f39 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Llava-prumerge: Adaptive token reduction for efficient large multimodal models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38d302b1-57fa-4005-9473-e754cfc6471d · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Towards vqa models that can read
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2aca5a95-6775-4eda-af5d-71e564924332 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Gemini: A Family of Highly Capable Multimodal Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd20365-a70f-4ce4-8f81-1ccd2ec40144 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs LLaMA: Open and Efficient Foundation Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f963ca89-69a0-452d-a544-64d57abc17ad · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add6ff33-db20-4c88-8b2e-dc0fdba1a4c0 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Teaching matters: Investigating the role of supervision in vision transformers
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 975e5536-81a4-4778-a559-4a3bfa2652df · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Hierarchi- cal prompt learning using clip for multi-label classification with single positive labels
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3cf076f9-39f4-465f-8f30-de9f6feb38f3 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Cait: Triple-win compression towards high accuracy, fast inference, and favorable transferability for vits
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9f9d1d74-e30a-4b10-a53c-3f52d4d93f73 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Repvit: Revisiting mobile cnn from vit perspective
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9174ab73-291a-4aae-9be2-b27dc785481c · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs YOLOv10: Real-Time End-to-End Object Detection
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c27f763-85ac-4383-aa07-0b32be338146 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Large Language Models for Robotics: Opportunities, Challenges, and Perspectives
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17676b0-df9e-4568-9076-ca336662f59a · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs CogVLM: Visual Expert for Pretrained Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab0c678a-7cf2-491e-b9b1-8dc5e47dff05 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Drivemlm: Aligning multi-modal large language models with behavioral planning states for au- tonomous driving
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30b716bc-0e36-48b4-9771-2b3c199b3e7b · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs A Comprehensive Survey of Large Language Models and Multimodal Large Language Models in Medicine
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39613f6b-db26-414f-8f4a-bf6c37f746f7 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Drivegpt4: Interpretable end-to-end autonomous driving via large language model
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 984286a4-b366-405e-97c6-53abeeab1c45 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bcebf37-bfbd-47d8-be5d-00105a08efd5 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94685f31-4229-4530-9ccc-b34e367696c1 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97b20cd8-00d5-4aaf-92ac-9380db2105b9 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 306ad005-f887-40c6-9df7-590f6111247a · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Object recognition as next token prediction
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b245ff0a-2cbe-4362-9aa4-30e236ab5991 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b9e68a6-f9f5-4d1c-b5c0-98c06c15dbdb · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93971e6d-2a07-44de-a0b5-a99af9012878 · outbound
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs Llava-phi: Efficient multi-modal as- sistant with small language model, 2024
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cad0e01f-f301-4383-a6e4-48966203a249 · inbound
Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 927d1f5a-eb39-4d25-a0d6-5dd9d7e7c996 · inbound
B-VLLM: A Vision Large Language Model with Balanced Spatio-Temporal Tokens [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0684c49d-fb94-46a0-b763-5b078f39c65b · inbound
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34f56709-6ad3-47d4-80cb-9719785e5851 · inbound
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf08dcb9-4566-433e-b52b-f72d9f749cab · inbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0453dc7a-29bb-4b04-bbfe-a7ce4230f910 · inbound
Training-free Token Reduction for Vision Mamba [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9cfad62-f673-4087-bdbf-bcec45303e3c · inbound
METEOR: Multi-Encoder Collaborative Token Pruning for Efficient Vision Language Models [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e058a539-f6ac-4009-9e94-a32b58190d49 · inbound
Efficient3D: A Unified Framework for Adaptive and Debiased Token Reduction in 3D MLLMs [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d5a48be8-1ce3-429a-aea9-64a04bc5560f · inbound
Geometry-Guided 3D Visual Token Pruning for Video-Language Models [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e0112050-4ed6-437b-a8bc-307ebd8fa627 · inbound
Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.