Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:02:07.964380Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 11 inbound Pith citation observations for arXiv:2501.00289.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:02:07.964380Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:28:32.362554Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T13:02:18.322222Z
92 of 92 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 12614b94-8db3-4552-9103-3fcda37994f4 · outbound
Dual Diffusion for Unified Image Generation and Understanding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f830245-6bd7-45d8-b696-4ff51a1c28c3 · outbound
Dual Diffusion for Unified Image Generation and Understanding Flamingo: a visual language model for few-shot learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad0f78d8-c3f1-4973-973d-25b316d7e2ff · outbound
Dual Diffusion for Unified Image Generation and Understanding Reverse-time diffusion equation mod- els
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca9ba559-4f16-4a77-b559-01469055e3e7 · outbound
Dual Diffusion for Unified Image Generation and Understanding Structured denoising dif- fusion models in discrete state-spaces
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56943249-3212-46e4-9730-4315a4b6d84b · outbound
Dual Diffusion for Unified Image Generation and Understanding Openflamingo: An open-source frame- work for training large autoregressive vision-language mod- els, 2023
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1965964-ac7f-4f26-9c84-5c1b953a397b · outbound
Dual Diffusion for Unified Image Generation and Understanding Qwen-vl: A versatile vision-language model for un- derstanding, localization, text reading, and beyond, 2023
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd605e9f-67f2-4df8-bbc1-2ab7cac49b69 · outbound
Dual Diffusion for Unified Image Generation and Understanding One transformer fits all distributions in multi-modal diffu- sion at scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09d5f859-714c-4b32-a2ea-a9a90571c2ac · outbound
Dual Diffusion for Unified Image Generation and Understanding Vizwiz: nearly real-time answers to visual questions
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd54cb74-3fc5-49fb-ad6a-79fb2f58ec0a · outbound
Dual Diffusion for Unified Image Generation and Understanding Language Models are Few-Shot Learners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aee87cd-613b-4e54-9fcc-4bd4d9b978bd · outbound
Dual Diffusion for Unified Image Generation and Understanding Pixart-α: Fast training of dif- fusion transformer for photorealistic text-to-image synthesis,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59817b7a-a16f-4ccd-9710-f1310c4e2075 · outbound
Dual Diffusion for Unified Image Generation and Understanding ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7203fae7-e095-4923-9c80-a057ab80e1cd · outbound
Dual Diffusion for Unified Image Generation and Understanding Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2faa5857-6ba0-4b48-ad13-9d3c01f7b4fc · outbound
Dual Diffusion for Unified Image Generation and Understanding Ex- panding performance boundaries of open-source multimodal models with model, data, and test-time scaling, 2025
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de9daf0c-0f99-488d-b378-75d101d30b85 · outbound
Dual Diffusion for Unified Image Generation and Understanding Instructblip: Towards general- purpose vision-language models with instruction tuning,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 090fbbe5-ba13-48da-a809-22bfe8708d5c · outbound
Dual Diffusion for Unified Image Generation and Understanding Diffusion models beat gans on image synthesis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26b5011a-6190-45d1-95a2-7f294344c698 · outbound
Dual Diffusion for Unified Image Generation and Understanding Continuous diffusion for categorical data
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0df6326-2066-45a5-ab1e-fcf6cf96f285 · outbound
Dual Diffusion for Unified Image Generation and Understanding DreamLLM: Synergistic multimodal com- prehension and creation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd51f741-1ab5-49f0-946c-b8753c1bee38 · outbound
Dual Diffusion for Unified Image Generation and Understanding The Llama 3 Herd of Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebce373a-e02e-4d81-b674-7c46aac7f1d7 · outbound
Dual Diffusion for Unified Image Generation and Understanding Taming transformers for high-resolution image synthesis
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 930f4408-39f7-43c5-82a4-c79502c14177 · outbound
Dual Diffusion for Unified Image Generation and Understanding Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a4d15cd5-6240-4594-ae61-5553ae042622 · outbound
Dual Diffusion for Unified Image Generation and Understanding MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78465120-2039-486d-8f58-0bce3a27cb57 · outbound
Dual Diffusion for Unified Image Generation and Understanding Datacomp: In search of the next generation of multimodal datasets, 2023
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 70ca07cd-d460-41aa-aafc-a56c2210b7c7 · outbound
Dual Diffusion for Unified Image Generation and Understanding Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f792e5-a7e0-47b7-a4e7-37dc30453433 · outbound
Dual Diffusion for Unified Image Generation and Understanding Discrete Flow Matching
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cf65f1b-7790-4aa7-8ee8-aa3128b26611 · outbound
Dual Diffusion for Unified Image Generation and Understanding Planting a SEED of Vision in Large Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19b6c957-0c07-4a83-a829-506788e29008 · outbound
Dual Diffusion for Unified Image Generation and Understanding Geneval: An object-focused framework for evaluating text- to-image alignment, 2023
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de7c55c-9195-4211-8e07-c22878fc2d6e · outbound
Dual Diffusion for Unified Image Generation and Understanding Making the V in VQA matter: Ele- vating the role of image understanding in Visual Question Answering
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4cb9e3ad-bcfc-4183-be5f-ffe778fe1486 · outbound
Dual Diffusion for Unified Image Generation and Understanding Likelihood- based diffusion language models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9e65b7ce-101b-4ed8-8b00-b0d12639ba9f · outbound
Dual Diffusion for Unified Image Generation and Understanding Classifier-Free Diffusion Guidance
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 849ff9ac-7ad1-4e6c-9deb-4d0c1c8f1937 · outbound
Dual Diffusion for Unified Image Generation and Understanding Denoising diffu- sion probabilistic models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3c038f1c-7be4-44df-8f1d-10a7c2728172 · outbound
Dual Diffusion for Unified Image Generation and Understanding T2i-compbench: A comprehensive bench- mark for open-world compositional text-to-image genera- tion
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 339486c2-7de3-4ebc-89db-faf2eb3fcdf9 · outbound
Dual Diffusion for Unified Image Generation and Understanding Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17cfd9b0-443e-4cd4-bbd1-3db1ba19e0df · outbound
Dual Diffusion for Unified Image Generation and Understanding The Platonic Representation Hypothesis
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff5a9fe-a7dc-4ae5-94ea-5f27130a97b6 · outbound
Dual Diffusion for Unified Image Generation and Understanding Elucidating the Design Space of Diffusion-Based Generative Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 532711f8-ed2a-4204-bd54-9568dc147d8f · outbound
Dual Diffusion for Unified Image Generation and Understanding Variational diffusion models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3922031b-bf94-46f9-8381-46ca209261f4 · outbound
Dual Diffusion for Unified Image Generation and Understanding Auto-Encoding Variational Bayes
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dfc06ce-04d2-4071-852d-491caa67c15e · outbound
Dual Diffusion for Unified Image Generation and Understanding The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bea8a73-09bd-4243-86d2-3e5ee57e83d3 · outbound
Dual Diffusion for Unified Image Generation and Understanding Rush, Douwe Kiela, Matthieu Cord, and Victor Sanh
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15968370-18d9-4465-bb75-559a4d0ddd94 · outbound
Dual Diffusion for Unified Image Generation and Understanding Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d208d2e-32fd-4b93-b2a4-6b727ecee7b4 · outbound
Dual Diffusion for Unified Image Generation and Understanding Likelihood train- ing of cascaded diffusion models via hierarchical volume- preserving maps
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20515266-7ab0-4758-b556-20ccb6555b9b · outbound
Dual Diffusion for Unified Image Generation and Understanding Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97be30ac-5b86-452a-bebb-329aa00c7287 · outbound
Dual Diffusion for Unified Image Generation and Understanding Autoregressive Image Generation without Vector Quantization
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f64cbffb-6e9c-4e02-b47d-a5808e2e5de0 · outbound
Dual Diffusion for Unified Image Generation and Understanding Diffusion-lm improves control- lable text generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 104eaeb3-cee6-4a01-ba8e-7d73a36759c7 · outbound
Dual Diffusion for Unified Image Generation and Understanding Diffusion-lm improves control- lable text generation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c19cb48d-4b04-4567-aa31-da0a91370c34 · outbound
Dual Diffusion for Unified Image Generation and Understanding What If We Recaption Billions of Web Images with LLaMA-3?
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c961acc-18f1-44db-8fd7-ee2acfd6ec59 · outbound
Dual Diffusion for Unified Image Generation and Understanding Evaluating object hallucination in large vision-language models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cf28e946-5e85-473b-a009-df6110367d9d · outbound
Dual Diffusion for Unified Image Generation and Understanding Microsoft coco: Common objects in context
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97384505-e0a4-4802-aa41-59ab308a0144 · outbound
Dual Diffusion for Unified Image Generation and Understanding Flow Matching for Generative Modeling
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd9f8e1a-d532-4776-a7bd-1c0c12cf4239 · outbound
Dual Diffusion for Unified Image Generation and Understanding Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c76f7b-f8aa-4da4-9d17-5fa2ca516084 · outbound
Dual Diffusion for Unified Image Generation and Understanding Visual instruction tuning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 526a1bdb-27f3-4409-a0c3-67a857194a3b · outbound
Dual Diffusion for Unified Image Generation and Understanding World model on million-length video and language with blockwise ringattention, 2024
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bae71810-f060-4f91-ab9b-8b4ce1c5aa4b · outbound
Dual Diffusion for Unified Image Generation and Understanding World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8aac922-88d2-4efc-a815-2823e36b8326 · outbound
Dual Diffusion for Unified Image Generation and Understanding Flow straight and fast: Learning to generate and transfer data with rectified flow
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c71bd467-dcdd-46b6-94df-917dc4bb56ac · outbound
Dual Diffusion for Unified Image Generation and Understanding Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b4b32d0-1100-457d-8898-d59f75a540e3 · outbound
Dual Diffusion for Unified Image Generation and Understanding Latent diffusion for language generation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2ebabc2a-7cc5-46b4-8fc2-fa2d594ed6cf · outbound
Dual Diffusion for Unified Image Generation and Understanding Ok-vqa: A visual question answering benchmark requiring external knowledge
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e08a9967-99b0-421c-86ce-53545f525c9c · outbound
Dual Diffusion for Unified Image Generation and Understanding Concrete score matching: Generalized score matching for discrete data
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2d0eafc0-37c1-4a2a-8b4a-d4fd2131db73 · outbound
Dual Diffusion for Unified Image Generation and Understanding Improved denoising diffusion probabilistic models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3428c539-73a8-4fe9-a21c-a568fc103fb3 · outbound
Dual Diffusion for Unified Image Generation and Understanding Scalable diffusion models with transformers
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3344b70-868b-40d7-8a1b-7eba5905ce38 · outbound
Dual Diffusion for Unified Image Generation and Understanding SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecad3bcf-0eff-44c1-9b2d-f9f8166f9037 · outbound
Dual Diffusion for Unified Image Generation and Understanding Improving language understanding by generative pre-training
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8fc8d875-4d0e-4d33-852a-cd441fbb3441 · outbound
Dual Diffusion for Unified Image Generation and Understanding Language models are unsu- pervised multitask learners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4782ba2c-048e-4375-a8ee-a82646ce6121 · outbound
Dual Diffusion for Unified Image Generation and Understanding Learning transferable visual models from natural language supervi- sion
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6318971a-872a-4f1e-aac3-58850c1bc3d5 · outbound
Dual Diffusion for Unified Image Generation and Understanding Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e59f7b03-a240-4b35-bbca-59ea593e955a · outbound
Dual Diffusion for Unified Image Generation and Understanding Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2535679-94ca-49e0-9e45-5118b9863964 · outbound
Dual Diffusion for Unified Image Generation and Understanding High-resolution image synthesis with latent diffusion models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ca8d73-2488-44ef-8329-5b8461d41994 · outbound
Dual Diffusion for Unified Image Generation and Understanding Photorealistic text-to-image diffusion models with deep language understanding
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f69cd402-92ef-4d14-9658-0d9c7d1c8b49 · outbound
Dual Diffusion for Unified Image Generation and Understanding Simple and Effective Masked Diffusion Language Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b150719-b58e-4344-a8e8-2ea6f480e9f6 · outbound
Dual Diffusion for Unified Image Generation and Understanding Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9d1ca3b6-2f14-4707-a057-10c6f6454802 · outbound
Dual Diffusion for Unified Image Generation and Understanding Deep unsupervised learning using nonequilibrium thermodynamics
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f9163c04-2049-4aac-8b53-22507e1425d3 · outbound
Dual Diffusion for Unified Image Generation and Understanding Score-Based Generative Modeling through Stochastic Differential Equations
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6652dbdd-6f6f-4f01-a86b-98a2ff2e03c0 · outbound
Dual Diffusion for Unified Image Generation and Understanding Maximum likelihood training of score-based diffusion mod- els
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ccf16b99-5962-467b-8070-713973b7da2f · outbound
Dual Diffusion for Unified Image Generation and Understanding Emu: Generative Pretraining in Multimodality
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fc79b0a-0f60-4fd0-8c0d-667845e8388b · outbound
Dual Diffusion for Unified Image Generation and Understanding Any-to-any generation via composable diffu- sion
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1a3a1287-3d8e-446a-af6a-7de4476f5680 · outbound
Dual Diffusion for Unified Image Generation and Understanding Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5554944d-8fef-4dbc-a107-a310e4eb2d63 · outbound
Dual Diffusion for Unified Image Generation and Understanding Gemini: A Family of Highly Capable Multimodal Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a85e97e-667e-4a98-aeff-3fc37818342f · outbound
Dual Diffusion for Unified Image Generation and Understanding Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce767ee-4f28-479e-9a5b-7d095aa1768b · outbound
Dual Diffusion for Unified Image Generation and Understanding Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49080754-2798-46a1-b8bc-35d5a6a1faca · outbound
Dual Diffusion for Unified Image Generation and Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d1253d4-a196-485f-987d-2ec1352ebd18 · outbound
Dual Diffusion for Unified Image Generation and Understanding A connection between score matching and denoising autoencoders
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6633b1ed-d3a6-4578-add5-08de184f0d66 · outbound
Dual Diffusion for Unified Image Generation and Understanding Emu3: Next-Token Prediction is All You Need
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8af2dd09-1b19-40ad-90c6-8cb3addcb339 · outbound
Dual Diffusion for Unified Image Generation and Understanding VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4bcfe59-c79a-4ba7-afc4-d7bead0b7692 · outbound
Dual Diffusion for Unified Image Generation and Understanding Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7edc67d6-4819-48cd-a2de-de9b643e65fe · outbound
Dual Diffusion for Unified Image Generation and Understanding Versatile diffusion: Text, images and variations all in one diffusion model
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 135d5202-5355-4b3f-b18e-10559816eb6e · outbound
Dual Diffusion for Unified Image Generation and Understanding Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cba4a55-bae1-4657-9598-9a2761189983 · outbound
Dual Diffusion for Unified Image Generation and Understanding Scaling autoregressive multi-modal mod- els: Pretraining and instruction tuning, 2023
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 646f88c1-ec84-4785-a9a0-ca8a56d6aacd · outbound
Dual Diffusion for Unified Image Generation and Understanding Sigmoid loss for language image pre-training,
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a96a108-076f-4901-a72c-4f033c1eab6a · outbound
Dual Diffusion for Unified Image Generation and Understanding Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed18e898-53b6-4a40-80b4-02a19d08ff69 · outbound
Dual Diffusion for Unified Image Generation and Understanding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 893d3773-da98-4aff-ad5c-b2ad65017bd2 · outbound
Dual Diffusion for Unified Image Generation and Understanding Dual pretrainContinued pretrainInstruct
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 981ccc90-16f0-4585-9ab8-e18da6861801 · outbound
Dual Diffusion for Unified Image Generation and Understanding (B) FID ↓ SD-XL [60] Diff
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 84498b3d-5063-4915-ba59-cbe31d579ee6 · outbound
Dual Diffusion for Unified Image Generation and Understanding On T2I CompBench, we find that after dual diffusion fine tuning the model performs worse in texture
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eb9db4b8-1c2d-402f-870d-5a2b59dde72d · inbound
Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling Dual Diffusion for Unified Image Generation and Understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 49bb7643-8ddb-44c6-a10d-9aacbc2e1f4c · inbound
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation Dual Diffusion for Unified Image Generation and Understanding
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0563063a-900d-40b3-9637-a4bb099a7ec9 · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Dual Diffusion for Unified Image Generation and Understanding
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b40dcffc-9bae-4483-bd3b-485d13ea5e07 · inbound
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning Dual Diffusion for Unified Image Generation and Understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ed0f780d-9212-47a9-96ab-f8b587539ce6 · inbound
OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks Dual Diffusion for Unified Image Generation and Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba11db8f-ae42-4811-9e68-11470bdeffb9 · inbound
Jodi: Unification of Visual Generation and Understanding via Joint Modeling Dual Diffusion for Unified Image Generation and Understanding
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45345e19-f584-4399-ae16-ff1d85fe254c · inbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dual Diffusion for Unified Image Generation and Understanding
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e001e7-be36-48f5-adc0-f3ca80abc84f · inbound
Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model Dual Diffusion for Unified Image Generation and Understanding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e2932b53-4de1-4e27-9e66-765eb64805cf · inbound
ReasonGen-R1: CoT for Autoregressive Image generation models through SFT and RL Dual Diffusion for Unified Image Generation and Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b74913d7-a687-49da-bcfd-170f68407a83 · inbound
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Dual Diffusion for Unified Image Generation and Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ac33dff-0e29-4430-b42d-934d3e4af888 · inbound
Show-o2: Improved Native Unified Multimodal Models Dual Diffusion for Unified Image Generation and Understanding
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.