Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:01:13.283168Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:1908.06066.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:01:13.283168Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T13:30:37.351291Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T18:56:48.399498Z
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 57f21e8d-bbdc-40d9-b965-da369a6c3c88 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 629aeb46-409d-49bc-9c0a-18157c1721af · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training UNITER: UNiversal Image-TExt Representation Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d571d1d-7950-465c-a25c-7c2f6d927fe2 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Unicoder: A Universal Language Encoder by Pre-training with Multiple Cross-lingual Tasks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 55160a59-8588-412b-9488-78404417ba9c · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Cross-lingual Language Model Pretraining
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c52cbe4-2736-4864-993a-4a72b1742700 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training VisualBERT: A Simple and Performant Baseline for Vision and Language
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 646cbc41-217c-4cf3-a704-543d6adf4a31 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee6b0e0-0cef-45be-b6ad-9436b121f399 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74967b05-acba-452a-955a-957c1b8c84e2 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e77d5504-8c88-424c-8260-53f0907d3b74 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training VideoBERT: A Joint Model for Video and Language Representation Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34afb34b-0a1b-4e82-8ee2-daa0cc7cba2d · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5f54d87-ddce-49b9-a752-40a05e710b58 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training XLNet: Generalized Autoregressive Pretraining for Language Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b5d464f-2550-46e4-95e0-c764e5ecf590 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training In 2009 IEEE conference on computer vision and pattern recognition, 248–255
Reference 2009
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dc70a464-2a05-470d-aef4-48481e2d1098 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Very Deep Convolutional Networks for Large-Scale Image Recognition
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d270b1b0-f24e-48f2-bff2-0345e1f3ea6b · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training A large annotated corpus for learning natural language inference
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a986e2f-d17c-4e60-aede-58f4c888cc73 · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training SQuAD: 100,000+ Questions for Machine Comprehension of Text
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d69f56a-31b7-4b92-9e98-ec399527844f · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training VSE++: Improving Visual-Semantic Embeddings with Hard Negatives
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd129c4a-1d34-4cde-94a7-c8db5588cd4d · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58888512-c5e6-46c3-8956-85fd6c948f8a · outbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training Fusion of Detected Objects in Text for Visual Question Answering
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfbe82c1-dc26-4f58-bde9-8af72792eba4 · inbound
Fusion of Detected Objects in Text for Visual Question Answering Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95b9913-23a5-4f1d-8966-17d993acb34c · inbound
RPN 2: On Interdependence Function Learning Towards Unifying and Advancing CNN, RNN, GNN, and Transformer Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.