Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T20:20:32.109683Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2509.08715.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T20:20:32.109683Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fbb2a09b-aa9d-4105-9a95-bce59f67ac5f · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8054e82c-2154-437a-b521-68e7650f06f3 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Lawrence Zitnick, and Devi Parikh
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d45ce047-a741-4fed-93d2-9641d82650a7 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14489560-6e77-4f19-9ec3-1b4ad1aa35e7 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Yu, and Qingsong Wen
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e740c634-1b0a-4eda-9da7-179d44b7fd16 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54da9b1e-d160-4c7f-acde-6cf7137bd3a8 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2a88261-d317-43b1-a9ad-f213266a70fc · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1b100a7-4b74-4ea2-9e53-6e88cee99c79 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion PaLM-E: An Embodied Multimodal Language Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ccf931e-efb8-4764-98e1-615a06a004e8 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d0006f5-74bd-45e7-9cf2-e0e7399119a1 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b91d14b-0dd1-42a1-b67f-ee7a8caa2952 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6212fe34-b008-491a-b149-b50f7dd94cd8 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion The Llama 3 Herd of Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e95421e9-1028-40d3-b9a3-2088cf0e4329 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Stangl, Anhong Guo, Chi Lin, Kristen Grauman, Jiebo Luo, and Jeffrey P
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c52993bc-6e3d-48aa-8c6d-39a00471e33a · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 580ce551-b7eb-42be-9eed-2eb2f9535327 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Hudson and Christopher D
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c998cc76-bf49-4db6-b8b9-28943e1fb2d1 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aa8c892-eb5e-4c99-8ac5-c79ce5a5f5f5 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6c1c8fc-d7b3-4a46-ad37-91c48d44c0ac · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion FloodLense: A Framework for ChatGPT-based Real-time Flood Detection
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c418fee-50a4-4bd3-b9f1-e7b13d5756c0 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion OBELICS: An Open Web-Scale Filtered Dataset of Interleaved Image-Text Documents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69b4835f-f2ef-4edb-ab63-7bcf7618857a · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion LLaVA-OneVision: Easy Visual Task Transfer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a91a41bd-446a-48fe-abb4-286cd22dfedf · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b53a083a-b41d-4047-a0f2-c1c37004469a · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6419b4aa-93fe-4f01-a2d1-66768114bc2c · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Selvaraju, Akhilesh Deepak Gotmare, Shafiq Joty, Caiming Xiong, and Steven Hoi
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb3d657b-2e0a-4e38-abd9-018bed994755 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion VILA: On Pre-training for Visual Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ead14f3-015d-46ef-9ae3-76c2a2b19fc4 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1ba2790-d1ae-45be-8d0e-9065fbef0ab2 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dadd9655-ac2d-43c1-ad80-3fe0ffc9c14b · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e5b9c70-ada8-4e25-9c40-a95c80c9b626 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion InfoNCE Loss Provably Learns Cluster-Preserving Representations
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba5e1a2-04e3-449b-8f0b-7b3cf0b0c433 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Qwen2.5 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 355c30c5-e9c5-4a4c-89ce-12b1ece284fe · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Learning Transferable Visual Models From Natural Language Supervision
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a0fdbf8-e3ce-415e-b1b4-8ccc2a3ae36d · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17c89c5-b1fa-4083-8e4b-7dd3ed0451db · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f049bd5f-62cd-44bb-8670-d6f785b50f1a · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Gemma 3 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 051c0ae2-8b8e-42db-b539-5eeb95a61fbc · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6085218d-969c-41d2-b5cc-c0b53ed990a9 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6922f30-9ecc-4b35-a002-01d5d5489067 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc9fbfb4-0ee6-4750-b893-1bcc4667eb55 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7acde425-109a-4556-ba41-63ba2d732315 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion online" 'onlinestring :=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d0cd1a6-8fce-4da4-8edd-c2268fd10c56 · outbound
BcQLM: Efficient Vision-Language Understanding with Distilled Q-Gated Cross-Modal Fusion write newline
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.