Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:19:28.185701Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2504.13123.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:19:28.185701Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:39:02.739280Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T15:39:09.371232Z
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation df80247f-9dbe-4d6c-8759-a1859b2f72d9 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35ba0703-4018-4ea1-819b-dbc4735ea5ff · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1b793377-d092-43fd-832c-35b96981fdca · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training PaLM: Scaling Language Modeling with Pathways
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53b1445a-8dd7-40c6-9a1b-e65051dd1f96 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Write and Paint: Generative Vision-Language Models are Unified Modal Learners
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1c7148b0-ee69-4ad2-b3a7-cdcdbb6b9c9e · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training An image is worth 16x16 words: Transformers for image recognition at scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 982c87cf-2d6a-4270-aa18-a6d232f17a26 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training The Llama 3 Herd of Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c1fc6d0-4b07-4ffc-8f2d-25f315d1a5b0 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dfdb7e5-e9fa-47a9-9f5e-be16e34f524f · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 883350f9-ef5a-48ba-a37c-fd4de2882fbc · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Visual Instruction Tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20ab5fa4-0c8a-4954-84a6-01cdac79a04f · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bd060094-ba59-480a-aef5-77476307b95d · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df753a51-f896-4d75-a577-43a4884a0720 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training LAION-5B: An open large-scale dataset for training next generation image-text models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400eec56-5304-47d1-a992-312df3bd0b83 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 284a5f42-329e-4ffb-bbd7-4d565552ee9d · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training FLAVA: A Foundational Language And Vision Alignment Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f2388d4-4ad5-4f52-a172-f5362daad767 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training LLaMA: Open and Efficient Foundation Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac68f509-cb66-40d8-9000-f7d2aa8c660e · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1813411b-6ceb-48eb-b25e-a1be8ad0c85f · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2809b66c-517e-4068-94bf-c1d847f00f4d · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2427627d-fc08-415c-b78b-22deac0bb1d7 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f33a485-a52a-49a4-a327-a3fcc13483bb · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Qwen2.5 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28a89139-89db-467c-8bdd-cea58f085994 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Multi-Grained Vision Language Pre-Training: Aligning Texts with Visual Concepts
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f592739-6d90-468a-857f-003e5b5c2459 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a84a82e-fcc5-491b-8251-9570a7698409 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2cc7478-edd2-461d-8a31-25c568c76904 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8888dba3-c773-4b41-916d-3092aacc2eb7 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Unresolved cited work
Reference 2011
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3080e3a3-abdd-42a7-bdea-6fe2c804baa1 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Will we run out of data? Limits of LLM scaling based on human-generated data
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b5fc0dc-625d-4310-acdd-c5a1d37cbf51 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Minwoo Byeon, Beomhee Park, Haecheon Kim, Sungjun Lee, Woonhyuk Baek, and Saehoon Kim
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 31d2a297-0def-479b-939f-871f9e0c4932 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c67cd18-fafd-4628-869a-1946f968ed7b · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Hallucination of Multimodal Large Language Models: A Survey
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b1744e7-b6dd-42ed-bfec-f381e83df06a · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training Flamingo: a Visual Language Model for Few-Shot Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42d773f4-123a-4294-9cb3-a3e46099ffce · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training LLaVA-OneVision: Easy Visual Task Transfer
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a61a0b-f893-490b-9aff-a2210c217a17 · outbound
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training GPT-4o System Card
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c188f1b7-3ace-4697-8ca8-53ccadd7f3ac · inbound
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3a734a36-931b-4707-b5bf-109b8efadd8c · inbound
CAPEval: A Decoupled Caption Evaluation across Understanding and Generation Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.