Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:23:39.194911Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 6 inbound Pith citation observations for arXiv:2411.18301.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:23:39.194911Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T10:46:55.020093Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T22:02:50.398472Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9b7ac967-ad2d-469f-b8f1-0833d46bd194 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation A-star: Test-time attention segregation and retention for text-to-image synthesis
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6b36b74-55fa-402f-a2f7-786da14c8a48 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Separate-and-enhance: Composi- tional finetuning for text-to-image diffusion models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa70afa9-7a6c-4d04-81f7-b2c69bd1fbbb · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Make It Count: Text-to-Image Generation with an Accurate Number of Objects
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e1ba90-e595-4ba4-9e11-527794ca7a5d · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Attend-and-excite: Attention-based se- mantic guidance for text-to-image diffusion models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d6dea554-32b1-4bff-afb4-d1a4bf206666 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Training-free layout control with cross-attention guidance
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83f0ef26-6420-4f2f-8cac-dd06f889bf94 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Reproducible scal- ing laws for contrastive language-image learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cf3d7aa-6c68-418a-8afe-ab3ca7c4052b · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Be yourself: Bounded attention for multi-subject text-to-image generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1b5e9962-00ee-4125-b24f-dd661b0b1715 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1daa7a5c-3cf4-46b9-9a32-f40123d61b80 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Training- free structured diffusion guidance for compositional text-to- image synthesis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bdccbab0-6763-4c00-b01c-b515802fe92d · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Initno: Boosting text-to-image diffu- sion models via initial noise optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d9f2fa3-cece-4ba7-aef5-a1ea67106a9a · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Optimizing prompts for text-to-image generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3f75a776-ea8c-411a-a646-47dcc18c425a · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 49d280ce-a727-4b10-8bd7-f62a4de6b2e1 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Classifier-Free Diffusion Guidance
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdd953a8-8959-458d-a57c-51bd1f9222d3 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Denoising dif- fusion probabilistic models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a7572d9-e64a-4f5d-afbc-e07231b7c4a4 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Scal- ing up gans for text-to-image synthesis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579048d3-32b6-456e-9780-b08236d76551 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Counting Guidance for High Fidelity Text-to-Image Synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a996a307-abfb-422b-ba7c-1e7d880a0cf2 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Compositional visual generation with composable diffusion models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e2196b6c-4a59-4244-8517-5ccbd57801d6 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Grounding dino: Mar- rying dino with grounded pre-training for open-set object detection
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation edeb82ba-38fb-4ce5-a7a5-c462eabb3979 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Design guidelines for prompt engineering text-to-image generative models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7118c30c-170f-499b-9d9c-6a3c05c1c738 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Flow straight and fast: Learning to generate and transfer data with rectified flow
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 990cc766-901a-481e-8f4e-b0dac7a31e8f · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Conform: Contrast is all you need for high- fidelity text-to-image diffusion models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1e700a80-f3a4-4760-b714-24a165f5a6cd · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 228509c5-4103-470a-9d7f-963f73be38c6 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Gpt-4o mini: Advancing cost-efficient intelligence
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d882c672-e587-4f36-bbd5-b90109b6f892 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 71b2bb50-a706-4842-8904-281f3d02f709 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Scalable diffusion models with transformers
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 017b59e2-16c9-4f77-9d47-34dc4eba6bc1 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ca5c12-b16d-42e1-93e7-ad7b07fa1124 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Learning transferable visual models from natural language supervi- sion
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33675c2e-1086-4084-9b9f-8670baea0a5d · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4d9d0ce6-8b5b-413d-a542-f780ee614bf0 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Hierarchical text-conditional image gener- ation with clip latents
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bbd40059-bcac-475d-9ee8-36fcaa52e920 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Linguistic bind- ing in diffusion models: Enhancing attribute correspondence through attention map alignment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3dcf6604-d04d-4d1a-a5cb-732fd86de452 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation High-resolution image synthesis with latent diffusion models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b67d03b8-8170-4235-af1c-9b8d201f5b6d · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation U- net: Convolutional networks for biomedical image segmen- tation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 833800c9-b2ad-4aa3-8cf2-26678c44e1f4 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Photorealistic text-to-image diffusion models with deep language understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 23863b69-e10b-4b33-8ced-ae041f090424 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Denois- ing diffusion implicit models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d781a978-54d4-47d8-bc54-39682746c2ee · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Predicated diffusion: Predicate logic-based attention guidance for text-to-image diffusion models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0935ab0c-1cbc-422a-a4a6-4218777fd227 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Diffu- siondb: A large-scale prompt gallery dataset for text-to- image generative models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eaeb710e-71b3-4f0b-ba2c-9a265ee8c7ff · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Hairclipv2: Unifying hair editing via proxy feature blending
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 386e4eb0-ce3a-4e38-9ec9-68c1cb288040 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Investigating Prompt Engineering in Diffusion Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3b47955-6427-45c6-b4fa-3806bf67e7cf · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Attngan: Fine- grained text to image generation with attentional generative adversarial networks
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3937ba5c-5a9e-4c13-9691-b1a363dc67b4 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Scaling autoregressive models for content-rich text-to-image generation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4a5d5aa7-326c-437b-b9d7-3813dbe6e9f3 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Cross-modal contrastive learning for text-to- image generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34a71818-393a-4627-8d68-66ed0062a84a · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Enhancing semantic fidelity in text- to-image synthesis: Attention regulation in diffusion mod- els
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9d8882ee-abb8-4a05-8a66-9349763eba9c · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Object- conditioned energy-based attention map alignment in text-to- image diffusion models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f5e1e28f-bba3-4172-917f-852acd81b526 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6a8b929b-d4ec-465c-8218-775c44242288 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation a chicken and a duck and a goose
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 265dfc8c-1306-4a5d-bcba-b7c41d48760d · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e4b1081-bce4-4cd1-9985-1ab3b9884c33 · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation Bad seed, rejected sampling!
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d0ead241-0e83-4425-a886-40783f5d320b · outbound
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation 7, 8, 12
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 64cd234c-1b99-421e-83c6-3df2b3ec6c14 · inbound
Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 856ae90a-91db-4aef-bce8-f6a7f4ba09ff · inbound
JEDI: The Force of Jensen-Shannon Divergence in Disentangling Diffusion Models Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de35d14-f731-4746-ba70-481e00658dbf · inbound
Scale Your Instructions: Enhance the Instruction-Following Fidelity of Unified Image Generation Model by Self-Adaptive Attention Scaling Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bfe246e-1ad4-4f52-8f8e-d93c200fa060 · inbound
Detail++: Training-Free Detail Enhancer for T2I Diffusion Models Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50c1ff24-d2f5-4c3c-a041-996742e60fc0 · inbound
LaRender: Training-Free Occlusion Control in Image Generation via Latent Rendering Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d5ea9f5f-5257-4c32-a219-cd3e9235ba5b · inbound
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.