Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:24:41.613546Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 4 inbound Pith citation observations for arXiv:2502.10458.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:24:41.613546Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:41.848847Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T07:16:44.528778Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 22f885db-7021-4c7d-b4d2-f2e02da7de95 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ed10bf5-00e1-4e10-b85f-b2740dd74a1c · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4894015-2a6d-478c-a88c-cf596569162e · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Emu: Generative Pretraining in Multimodality
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3c019a-4926-4b07-9d9f-7d2714d09cc2 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models MetaMorph: Multimodal Understanding and Generation via Instruction Tuning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 083c57dd-e15a-4e01-b9b9-032256d952e3 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58444dd7-0d96-4e22-af4f-c55c151bc7b2 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models OmniGen: Unified Image Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89300a6a-ce86-42d5-9296-ee9b39e789eb · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33892029-21d1-4f12-b64c-54f9cd3d0a5e · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models X-VILA: Cross-Modality Alignment for Large Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66aca6c8-bbd4-4ea9-99fd-bb9ff985fd77 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb751775-0dc6-4f23-9fda-ee7c7109f809 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Limitation Despite ThinkDiff’s strong performance in reasoning generation tasks, several limitations remain for future work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f35c36b-0d8d-4518-b496-4576e66c666d · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models These images are preprocessed using Qwen2-VL, which generates detailed descriptions based on randomly selected text prompts from a predefined set
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2d81f8f4-bbf6-47d5-8c78-f11dd0ff8683 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Kosmos-G: Generating Images in Context with Multimodal Large Language Models
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82064ae3-1537-41f5-954b-ed80812f159e · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 774753ec-f91e-4c9c-a525-2acc165eaf77 · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ddc80c-8930-4c48-b240-0512dee2b3da · outbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b9f5c38-17bd-4f58-bd65-fb5155828610 · inbound
Fake it till You Make it: Reward Modeling as Discriminative Prediction I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdf387b8-c76b-48f7-bc82-8ba6fb216934 · inbound
Canvas3D: Empowering Precise Spatial Control for Image Generation with Constraints from a 3D Virtual Canvas I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78e3f980-ffb4-4779-8a20-f93842855331 · inbound
MentisOculi: Revealing the Limits of Reasoning with Mental Imagery I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7caac5a9-9141-425b-8b35-17869fbb82d9 · inbound
Evaluating Reasoning Fidelity in Visual Text Generation I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.