Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 24 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2003.13198.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:02:31.632210Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:43.933033Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a642598c-7b87-44b7-8bec-c36abd4048e0 · inbound
Natural Language Understanding and Inference with MLLM in Visual Question Answering: A Survey InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 263
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 939adbb6-a456-46c7-992a-95cdb71130bd · inbound
Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model Enhancement InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8adc19f-c2b5-4ba0-8cf0-85c9d03affb5 · inbound
Semantic-enhanced Modality-asymmetric Retrieval for Online E-commerce Search InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e69e4afb-502b-4b3c-88bc-c5efe01eba00 · inbound
MULTIBENCH++: A Unified and Comprehensive Multimodal Fusion Benchmarking Across Specialized Domains InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cea11255-1ed5-44a1-a8a8-507e810d7400 · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 131
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0e5a0528-1149-40b1-853c-d2b2a835cc56 · inbound
Qwen-Audio-VAE Technical Report InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
Reference 146
Source-reported events for the cited work
Unavailable: canonical work link unavailable.