Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2205.12005.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:28.217225Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:29:56.632867Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation fe4fe666-abea-4baf-a398-efc8c3d7cea8 · inbound
GIT: A Generative Image-to-text Transformer for Vision and Language mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cbc7e6ea-4b69-49b7-9690-b57795862f74 · inbound
OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0547117b-2545-43d3-8918-a74655c7f44c · inbound
ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a22d27b4-0d92-4ae2-a3e9-4c5822de9a63 · inbound
The ART of Composition: Attention-Regularized Training for Compositional Visual Grounding mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23b3c1fa-12d1-4479-8187-630a5c314b33 · inbound
HyperCap: Hyperspectral Land Cover Captioning Dataset for Vision Language Models mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f4722c6c-9f1a-44ad-9e7b-9df5dcc26f57 · inbound
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts? mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f23c11-628c-4107-8673-7a996920c4ab · inbound
GETReason: Enhancing Image Context Extraction through Hierarchical Multi-Agent Reasoning mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a0b8c8a-8ab5-424b-8abc-6111086bdfc5 · inbound
ReFrame: Rectification Framework for Image Explaining Architectures mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 650fe551-ea2f-416f-8532-9aa34c0f1079 · inbound
INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c2d67bd-b0b0-4603-9b6d-329e2e9c975a · inbound
Multimodal Feature Fusion Network with Text Difference Enhancement for Remote Sensing Change Detection mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 757067d8-2202-4a58-84c5-823ce07363e1 · inbound
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7014bc0-822a-4cee-8cad-1eca19200620 · inbound
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eeb09fb3-51ba-4736-a08f-288dded6a40d · inbound
AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 55733916-8a99-4c15-871f-0e32c9a9a304 · inbound
VisChronos: Revolutionizing Image Captioning Through Real-Life Events mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.