Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T12:40:56.954792Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2607.26596.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T12:40:56.954792Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1f88dbb6-2687-4aa1-9f9a-997a3ba183fc · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d9c196b-f4a7-4bbd-9d2f-052b3c22c11b · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a61cdb0-9ed5-417f-95a6-a18341fcdc43 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution VILA: On Pre-training for Visual Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 830ae66c-781d-41e5-9326-2c2e2ff03646 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Decoupled Weight Decay Regularization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3100d93d-404d-4521-a8ec-ab1fecd3994b · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 790da517-9b0a-423b-a228-e98ba5390d4b · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25be7dd9-6dfe-479e-a1ac-7ed684938b88 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00355b0e-bfab-4df1-8d82-d3624c9629dc · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution FLGo: A Fully Customizable Federated Learning Platform
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abbab7a4-eef6-4abe-8d74-a2efe25f6549 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2d46f7d-6730-4e90-80dc-490bfbf64241 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6facfb72-b996-4426-bca3-9d5f5b3d3a2b · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution LoRA: Low-Rank Adaptation of Large Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e6b62ce-890d-4902-a46d-f722d15d2379 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f387456-358b-4518-b435-c8a1f63056a1 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c30a087-2d53-4d67-ad03-b22958e7e088 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1902258-08e3-4dc7-bbb3-d3a19e07e521 · outbound
Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.