Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:06:54.293800Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 2 inbound Pith citation observations for arXiv:2505.20053.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:06:54.293800Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T09:52:01.773382Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T07:16:44.587845Z
69 of 69 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c06619fa-9623-4a5f-ae7b-52d9b53670fb · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion High- resolution image synthesis with latent diffusion models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5670b8-5329-4dd2-873e-efea9da8a204 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c39549-f37a-4bdc-b515-ceb6be0a7050 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82fbe835-84d6-42cd-b205-46448a2e69de · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Denoising diffusion probabilistic models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca4bca66-6924-4fec-9c8a-c7ab9994c99a · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Glide: Towards pho- torealistic image generation and editing with text-guided diffusion models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 516980c7-ff1f-4bc7-bdeb-0e440de27559 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Photorealistic text-to-image diffusion models with deep language understanding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fae598c4-1250-4417-9dca-ad311a928cd5 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Diffusion Model Alignment Using Direct Preference Optimization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d8ce54f-ea13-4e07-b26e-8cb25d398790 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Reflected diffusion models for text-to-image generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10ad2305-3411-41e1-992c-3b249212fc2f · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Zigzag diffusion sampling: Diffusion models can self-improve via self-reflection
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4b9019f-129b-4818-8f7d-a5c4ea30eb39 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f69bfa58-c945-4c6f-98e1-15914aa25905 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd0a1b04-1430-4840-a719-c88703dc6216 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Improved Baselines with Visual Instruction Tuning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea24c6c6-0162-4ff9-96f8-00de82811965 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c08b10a6-45de-4872-9748-4f015951ff6b · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Flamingo: a visual language model for few-shot learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e219843-5cf5-4465-a564-1712dca7ca49 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e50bf329-a9fe-459d-a036-13a62c91e7f9 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Gpt-4v(ision) system card, 2023
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea40dd3e-078d-4a32-806f-ee9f5d2867fe · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab32b750-5c39-4104-bad3-bc2629e92b08 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion The claude 3 model family: Opus, sonnet, haiku., 2024
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 301aff66-ce55-45c8-91de-c03f5c645ec0 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 473e7b59-2b52-4a8b-a714-8d517271f1a8 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Visionllm: Large language model is also an open-ended decoder for vision-centric tasks
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca8069a4-8084-40fd-b260-cd8a3a3f04f4 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion CogVLM: Visual Expert for Pretrained Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc00e381-7b25-419d-911a-94e666446bf5 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c37bb812-1d23-4f05-a0ba-81645fefeb15 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfddfa8-bcb5-48b3-81ca-e67109e115cf · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Visual instruction tuning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d778432-d4a7-4989-967c-683703e71c7c · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Llava-next: Improved reasoning, ocr, and world knowledge, 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abae598a-bd1a-4bf1-b652-9477b1db0b36 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion video-SALMONN: Speech-Enhanced Audio-Visual Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f486cbc-df12-485f-8dc2-8c4e544ab8ae · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9832be2-b9f5-4fdd-aeac-e089fbae9e01 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion NExT-GPT: Any-to-Any Multimodal LLM
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 158b09e0-b90c-4054-9de1-0ec7c148dae2 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 609b4299-45d5-421d-9257-8c03625df872 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Gem: Empowering mllm for grounded ecg understanding with time series and images
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 924717a1-24cd-45db-99c1-8602fcd9fe8b · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee156c4d-6fdf-467e-be13-d763710906db · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Instructblip: Towards general-purpose vision-language models with instruction tuning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe97c230-cacd-4740-b9b9-67ccfdc43bc5 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Deep unsuper- vised learning using nonequilibrium thermodynamics
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation effea327-ff42-4398-8134-c0a3d82fabb8 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Generative modeling by estimating gradients of the data distribution
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642ddab9-115c-4701-b043-540e0ae28fcb · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Score-Based Generative Modeling through Stochastic Differential Equations
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2df6ad0c-f992-4801-a5b2-60d4b5b6a25c · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Improved denoising diffusion probabilistic models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60078f2f-56e9-4a43-849f-07f80ff1ba27 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Variational Diffusion Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76e04cf0-65aa-4a61-b6ea-3ba52601ff6a · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Diffusion models beat gans on image synthesis
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d43efcaa-0729-4ffe-a8fb-5c40c94dee3a · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Denoising Diffusion Implicit Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0e00451-32fd-4ecc-a9df-06ff70552686 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion On fast sampling of diffusion probabilistic models, 2021
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ca835f-bac8-4b5f-b7ae-61564d957ac3 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Progressive distillation for fast sampling of diffusion models, 2022
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fc98851-47c4-44a9-a0f7-12ba55f5e3b4 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a5553f-4d25-4df0-91bb-6ba334d54caa · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion A morphology focused diffusion probabilistic model for synthesis of histopathology images
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 42c2468a-1653-4443-8c81-46d292bea758 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion ViT-DAE: Transformer-driven Diffusion Autoencoder for Histopathology Image Analysis
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fa73ae7-5566-407a-ac17-46718afa9a6b · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion A multimodal comparison of latent denoising diffusion probabilistic models and generative adversarial networks for medical image synthesis
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7bc475f4-c260-4ee1-8ef0-fb19c053fbaf · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion PathLDM: Text conditioned Latent Diffusion Model for Histopathology
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9ffe029-eaad-4e40-ae29-8cc014e4afa2 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Generate Your Own Scotland: Satellite Image Generation Conditioned on Maps
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 835f4c14-42e8-4a81-b790-b0dfe214c2cf · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion RSDiff: Remote Sensing Image Generation from Text Using Diffusion Model
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38a8099b-876f-4083-9b15-fd5f310778fb · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Diffinfinite: Large mask-image synthesis via parallel random patch diffusion in histopathology
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6588839-6739-4c54-a6b4-2b9d52746045 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Adding conditional control to text-to-image diffusion models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b5878c5-cde4-430d-8d15-44266ed48e53 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Zero-Shot Text-to-Image Generation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea94fad3-b85e-411e-9b82-40c19f678f2c · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Classifier-Free Diffusion Guidance
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6c5ff1e-165e-4788-a6c9-b6fe757af986 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac61415e-7d8c-4bbf-82ff-03c186bd7891 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e115430-fe59-4210-8c0d-54f64c306505 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Proceedings Combined 30th International Workshop on Expressiveness in Concurrency and 20th Workshop on Structural Operational Semantics
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3b398bc-8821-45dd-bc29-fee3c6833b06 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fd60d17-982f-4d9c-8cba-f4daaa680f76 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Pixart- σ: Weak-to-strong training of diffusion transformer for 4k text-to-image generation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00d7166c-da29-472b-910f-19f7d6158fc3 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion High-resolution image synthesis with latent diffusion models, 2021
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e3bcc20-d7d9-46e8-9144-f7e9d13468b4 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Qwen2.5-VL Technical Report
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5849a7b6-c064-4c6b-b828-6a0e448de70e · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9ea0a37-f9b8-4b22-b3d2-a4fcb87725f0 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d8ea50b-2b65-464c-bc35-36390c751286 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion "" Generate an refined prompt to better match the original intent
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3aed37fd-81df-490e-8b7b-e62f8595a9e2 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5dc84d7c-c437-4cf6-a6bf-81ff2e2ebcd1 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a612da21-ff5b-440a-8550-9912a649a6af · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation edb1aa06-00e2-4db7-819c-5e0df1c42cfc · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe3c1149-e799-43e8-931e-5db7bb1af3d6 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion "" Generate a omission h ig hl igh t to e li mi na te unwanted elements
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 332eaefd-0085-4fa2-86c3-ec6e302ecb70 · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b50e467b-8dcf-4084-85a6-d29d84eb4f9c · outbound
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion OLLL”, “ZALL
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9baf39b9-98f5-4453-b4fe-7d4d6105034a · inbound
Evaluating Reasoning Fidelity in Visual Text Generation Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fb06972-df6a-42e2-ba65-76813b6d0c0a · inbound
EventOD: Event-Aware OD Flow Generation via LLM-Guided Semantic Modulation Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.