Pith. sign in

Paper Citation Record · LEDGER

Test-time Prompt Refinement for Text-to-Image Models

As of 14 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 3 inbound Pith citation observations for arXiv:2507.22076.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22076 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:01:52.609484Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:13:10.574530Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:29.278949Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43569dd6-df28-472c-b09b-7dd3ef991d24 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Test-time Prompt Refinement for Text-to-Image Models Flamingo: a visual language model for few-shot learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.176538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.176538Z digest=sha256:2558ee7efb86e09cfef29d5b9828ff980b0f16305717720bfe526de68650b17f

Observation d086ba3e-d8bd-4d9a-b8bc-3c593c135bf0 · outbound

This paper cites Blended latent diffusion.

Test-time Prompt Refinement for Text-to-Image Models Blended latent diffusion

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.721694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.182180Z digest=sha256:51d57eedda54a5031db991017a7b8a5852b562aed649204af625d1b18fd013c7

Observation 2e92cf0d-7eeb-4d4c-9d31-50ebba10271f · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Test-time Prompt Refinement for Text-to-Image Models eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.186385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.186385Z digest=sha256:44e5f8df561bfccc5a2eec0ceaee06a2884bf6a1b49b761a728e9ae78a521cd4

Observation 7f4b6d50-f2df-4c29-b53a-be447b2216c3 · outbound

This paper cites Multidiffusion: Fusing diffusion paths for controlled image generation.

Test-time Prompt Refinement for Text-to-Image Models Multidiffusion: Fusing diffusion paths for controlled image generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.700213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.191212Z digest=sha256:9b9919e088a125290f1f5c0056ec17de6cfae7f220610473c0b2ab2a05f3a7c3

Observation 9c19095c-5274-4f7e-8dbd-ea1d0be6be35 · outbound

This paper cites Improving image generation with better captions.

Test-time Prompt Refinement for Text-to-Image Models Improving image generation with better captions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.679210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.195527Z digest=sha256:8c69840568853d1f0cd38052297833844e4529ab3b70c2895d6c63d9eb5641ad

Observation f1065118-1e2e-437d-8dbc-4419629261b4 · outbound

This paper cites Training-free layout control with cross-attention guidance.

Test-time Prompt Refinement for Text-to-Image Models Training-free layout control with cross-attention guidance

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.663088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.199854Z digest=sha256:06b23f5513ce2eae183edcabf2de5b490fb945515d6c3ed0da1ed0216bf3bcdd

Observation e5531e3a-38e1-423a-8b42-ad44ed35c59a · outbound

This paper cites Masked-attention mask transformer for universal image segmentation.

Test-time Prompt Refinement for Text-to-Image Models Masked-attention mask transformer for universal image segmentation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.645208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.204947Z digest=sha256:b7026e65fc99eb662f941bc806e5c6276abe06c51b43e67dde4f6e46717d81aa

Observation 7a840b82-e008-4775-abc9-d7e62f1a7d72 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

Test-time Prompt Refinement for Text-to-Image Models Cogview: Mastering text-to-image generation via transformers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.627074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.210094Z digest=sha256:4c5bc801e84f6cf008cfeef93245d7a865f218fb0071f9cf09ddb28c09c9da62

Observation 8cad3c74-3d26-436f-ac96-f036b47e3f99 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Taming transformers for high-resolution image synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.215333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.215333Z digest=sha256:767691946f93d0e8d3f5cc84f7434c27ddbdff98123f17376cc674446a9f58a9

Observation 41cb1387-aa9d-4c76-ae4e-f2046d7fd1be · outbound

This paper cites DPOK: Reinforcement learning for fine-tuning text-to-image diffu- sion models.

Test-time Prompt Refinement for Text-to-Image Models DPOK: Reinforcement learning for fine-tuning text-to-image diffu- sion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.598235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.220132Z digest=sha256:5a7e0a7107fbd531cd62f8d203ec4290ea3091d31c9e10962ccf281fde33aadd

Observation 2ec0318e-7b3d-4610-9ee5-19cc19ca7bb2 · outbound

This paper cites Layoutgpt: Compositional visual plan- ning and generation with large language models.

Test-time Prompt Refinement for Text-to-Image Models Layoutgpt: Compositional visual plan- ning and generation with large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.224431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.224431Z digest=sha256:5954678de738da9874e2091f0c15281ca6785d4f27a346f4ff10c3a892de2e33

Observation b668eda1-6ab9-4385-b393-4e9b5541dd2f · outbound

This paper cites Layoutgpt: Compositional visual plan- ning and generation with large language models.

Test-time Prompt Refinement for Text-to-Image Models Layoutgpt: Compositional visual plan- ning and generation with large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.570232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.228582Z digest=sha256:8d5f1ba89b6ef6104a4c9f8ddaf7ce1e82396df6cac6655033b05d58d36159b7

Observation d2feb80e-38fb-46c3-bf8f-41a51c693c73 · outbound

This paper cites LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts.

Test-time Prompt Refinement for Text-to-Image Models LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.232968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.232968Z digest=sha256:ae9f2a852c5e489e47b9f6019676399738a73b83928a941567f38773d050e84b

Observation 7fb07d2c-bc98-4ba5-a989-8c9a404d06ba · outbound

This paper cites Geneval: An object-focused framework for evaluating text- to-image alignment.

Test-time Prompt Refinement for Text-to-Image Models Geneval: An object-focused framework for evaluating text- to-image alignment

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.551291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.237769Z digest=sha256:9628deb9b811f3ad130a75d9b0fa512e3d861282ef792d4397c893bcb2da31d0

Observation 607496b9-0335-4ba5-8b1e-70b831dbe442 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

Test-time Prompt Refinement for Text-to-Image Models Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.241657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.241657Z digest=sha256:8f699d27d1cba2ec01049d24f15d2bf68417054759dfa912e76e9a2545643f1c

Observation 5d66487a-27c2-485b-a521-03b2ffe11767 · outbound

This paper cites GPT-4o System Card.

Test-time Prompt Refinement for Text-to-Image Models GPT-4o System Card

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.246022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.246022Z digest=sha256:4cf7fe6f11a763882b11b0ddf541098caa0d8500fb34a968af35a1b222d117fe

Observation 7b7040ec-03fc-4003-bbd9-b6762caf11d4 · outbound

This paper cites Few-shot classification and anatomical localization of tissues in spect imaging.

Test-time Prompt Refinement for Text-to-Image Models Few-shot classification and anatomical localization of tissues in spect imaging

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.529471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.250084Z digest=sha256:cb0ad1eb187b53a51a24fc94ad90324387fa0de994231b2326b1ebb0719e903c

Observation f334e179-6c0b-4836-805a-a3a84e487e3a · outbound

This paper cites Clas- sification of microstructure images of metals using transfer learning.

Test-time Prompt Refinement for Text-to-Image Models Clas- sification of microstructure images of metals using transfer learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.510571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.254159Z digest=sha256:a451d814a1da0b6db0d8814d4bbf239ddfa473a5eb2885fe6e9433cc304fc4e5

Observation ac1bfb0e-b880-4f05-a352-0d20f21c9884 · outbound

This paper cites Alina: Advanced line identification and notation algorithm.

Test-time Prompt Refinement for Text-to-Image Models Alina: Advanced line identification and notation algorithm

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.492598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.258108Z digest=sha256:db9a20cedd91b57795b01a7265c3b388ebafbab8dd7accc6ddebac225242c1ea

Observation 4fbd46a4-5c4d-46b5-96cd-ec8e50de5cb6 · outbound

This paper cites Gen- erating images with multimodal language models.

Test-time Prompt Refinement for Text-to-Image Models Gen- erating images with multimodal language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.476091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.262537Z digest=sha256:f23536c716200df811b4ac6b0316d7e225a73f190ee3b12abbe35df54c16c77f

Observation 1e8729aa-64a5-400f-8ff6-9fa8c1c86155 · outbound

This paper cites Zero-shot Text-guided Infinite Image Synthesis with LLM guidance.

Test-time Prompt Refinement for Text-to-Image Models Zero-shot Text-guided Infinite Image Synthesis with LLM guidance

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:01:52.927761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.267531Z digest=sha256:8754fd138c0573d58dad83a734ede94c7aad86f7e03c5d39454a4c154cf29430

Observation 0b04754e-ca52-4c10-965f-b0fc759b9aa5 · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.458307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.273106Z digest=sha256:18d7796cf5cd952bcbed74b4368567b527232c8eab9aabe9449b43097b809a3c

Observation c458aae0-ddd2-42cd-8521-ce422abd5c84 · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

Test-time Prompt Refinement for Text-to-Image Models Aligning Text-to-Image Models using Human Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.277213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.277213Z digest=sha256:d5aae0ab6e3ad386c21b5051b0eaec87eb0bc3f8632b529e6445657d290c3cda

Observation ed207dd5-73fc-477d-b770-76d4b4edfa32 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Test-time Prompt Refinement for Text-to-Image Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.281822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.281822Z digest=sha256:1a5945a8bc2ff3f6f306e3b59226c2c71c68238fb1f35ebdfb398ba019d3ea66

Observation b870a995-e2de-4ba2-9621-f8a92b33a46a · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Gligen: Open-set grounded text-to-image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.429341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.285748Z digest=sha256:0878035740252c345868e177d0d4e8adcf4bc9cd472f8a98af4a8e55a4e9a03f

Observation 0f586462-6019-4bfd-a6d6-84efb249685e · outbound

This paper cites LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models.

Test-time Prompt Refinement for Text-to-Image Models LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.290435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.290435Z digest=sha256:91a5d916d0d2f7e81b7e7d687a4e89a82d602301da228b1b0218083561a04c42

Observation 666964f2-506d-4ca3-860c-6dda68be8baa · outbound

This paper cites VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning.

Test-time Prompt Refinement for Text-to-Image Models VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.294824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.294824Z digest=sha256:e0b8c4c9cec383877e28d34c7f1ace68ebf7d8708ab60feb4890538e42c2cc8f

Observation d2b90bad-7946-40f5-a4b1-644a8afd93a2 · outbound

This paper cites Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps.

Test-time Prompt Refinement for Text-to-Image Models Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.299646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.299646Z digest=sha256:4cbbe9a0e217f285dd40586328be1fd4ca3c0642510e34c7e966922d1a828e93

Observation c69a07c1-9aec-482b-b49f-5660feabae24 · outbound

This paper cites PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models.

Test-time Prompt Refinement for Text-to-Image Models PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.303932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.303932Z digest=sha256:3e980b32edbb21b5d13ef632e31fefecc95bafe03af213057f2a660f5e2a111e

Observation 0827f241-f3f2-4252-abbf-8283292521de · outbound

This paper cites Simple open-vocabulary object detection.

Test-time Prompt Refinement for Text-to-Image Models Simple open-vocabulary object detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.409155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.308732Z digest=sha256:a73f4187a491670d3714d49449235c876dcf47a9de135648fe2998109eb39819

Observation 7e5399fb-d84e-42d3-b10c-bd914fde2f4b · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Test-time Prompt Refinement for Text-to-Image Models GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.313124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.313124Z digest=sha256:e141226fccb0457fe3b84930915fa4857d16225f94d1c537d8100baf4ac6f733

Observation 3edf1190-2cbd-4ab6-9b5a-52d72c946208 · outbound

This paper cites Localizing object-level shape variations with text-to-image diffusion models.

Test-time Prompt Refinement for Text-to-Image Models Localizing object-level shape variations with text-to-image diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.390860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.318058Z digest=sha256:6272e5623ded82fc6b7e522bd4a69e174e65a6497198f6d94adbade630fe704a

Observation 7e66b17b-116f-467a-a772-e431decfacdc · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Test-time Prompt Refinement for Text-to-Image Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.322258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.322258Z digest=sha256:79c76f6af65f27caca5b762b80c892cfcb68ba6ce74a78645e77ecd65af00911

Observation 8fc2749c-911d-4d41-a400-bcca2f8dbe4f · outbound

This paper cites Diffusiongpt: Llm-driven text-to-image generation system.

Test-time Prompt Refinement for Text-to-Image Models Diffusiongpt: Llm-driven text-to-image generation system

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.327594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.327594Z digest=sha256:401cf24c28a86a6065b9243814e76980d89c5590e113952523407f717acc8d03

Observation 4be411d5-7a06-439d-8ffb-7f5fd95a4a65 · outbound

This paper cites Layoutllm-t2i: Eliciting layout guidance from llm for text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Layoutllm-t2i: Eliciting layout guidance from llm for text-to-image generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.373242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.331685Z digest=sha256:769c82b2aff8899e2abf8b0ecbf3733aa5ef14b0e3fd299549f9889dc5ec5cf6

Observation 95ae111d-5157-41a2-bd76-2f672cfde516 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Test-time Prompt Refinement for Text-to-Image Models Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.356271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.335654Z digest=sha256:927268fd12204e5a6939e300d7d34ee2c127046bfc1bd2f4394520ad2fb00b5b

Observation dbe6e7c9-470f-4e2d-94b5-667c65a5c52f · outbound

This paper cites Zero-shot text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Zero-shot text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.335735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.339583Z digest=sha256:acaea5088b1418cfb17087bc6b8060d2b22da0ad32c8cc8256cca42eb59c09ba

Observation f24677d5-041d-4729-8ad2-7a39540dd740 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Test-time Prompt Refinement for Text-to-Image Models Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.343453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.343453Z digest=sha256:a54834263255a35558548da45eead8120360e20f21f0e3d7e4c0e3039e1553d7

Observation 5b6bbfd6-dd56-4e85-bb72-145d4820cb02 · outbound

This paper cites Generative ad- versarial text to image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Generative ad- versarial text to image synthesis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.347957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.347957Z digest=sha256:c10a78f7f095d19f7b8afb47644ff11725bec795c4475ea0bc06b66ece8efda5

Observation 054a60bb-6615-44f5-bcb2-ca5f51cd2a55 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Test-time Prompt Refinement for Text-to-Image Models High-resolution image synthesis with latent diffusion models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.307768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.352213Z digest=sha256:9e93875800279032bf6863f652988f6c6e8f3978d36342e29b4ce86443314021

Observation 34456057-989b-4c33-95f7-f18177b0478c · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

Test-time Prompt Refinement for Text-to-Image Models Photorealistic text-to-image diffusion models with deep language understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.357931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.357931Z digest=sha256:6ea11be822829b2c4202fc3319ac426cf6503c587f82b324b1dcbb948deac96c

Observation a3b624fe-b3e8-4325-9b8a-77b0efa9309a · outbound

This paper cites Benchmarking awesome diffusion mod- els.

Test-time Prompt Refinement for Text-to-Image Models Benchmarking awesome diffusion mod- els

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.281385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.362885Z digest=sha256:656c1160cdf31ad2ec8ef665ecd0411b1812935cc0134e43ba02c4bd94fc8985

Observation 6f5a8638-aeef-488c-9eb8-5dd93c618c76 · outbound

This paper cites Df-gan: A simple and effec- tive baseline for text-to-image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Df-gan: A simple and effec- tive baseline for text-to-image synthesis

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.249741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.371093Z digest=sha256:8b1a0485888ed156f7f944e889ee0dbe44ea76102ed5d3be04eaf53703be226b

Observation 52b951c6-ecf2-45a9-ab21-7763cce1a45b · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Test-time Prompt Refinement for Text-to-Image Models Qwen2.5: A party of foundation models, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.233186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.375107Z digest=sha256:0c2de535a08a76695c31a3580dbb686cd03ac85c8512041afa673e7587b622db

Observation 8757f061-212b-41f5-9a4d-2d7b2673e23a · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

Test-time Prompt Refinement for Text-to-Image Models Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.379789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.379789Z digest=sha256:4c6b6057c9789a94b61546c832f3f6efb1a3d06e6e21a47aafe03069be71ca4e

Observation 306d0874-42d8-4f2c-8762-a23c7dabc5b3 · outbound

This paper cites Self-correcting llm-controlled diffu- sion models.

Test-time Prompt Refinement for Text-to-Image Models Self-correcting llm-controlled diffu- sion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.218936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.384252Z digest=sha256:b95c7248fe0e8c096b0b1a14d40067bcc92c7f2eb8dac96cac2946fd9830cf00

Observation c40d7134-6a2b-4b31-99d2-0c92c2e7e2f6 · outbound

This paper cites Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion.

Test-time Prompt Refinement for Text-to-Image Models Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.201848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.388650Z digest=sha256:ddf9f25f543bb84c658c9ec08f2d9e430ce28915fb52cf79011693fd530a8e29

Observation 585525c1-dfa8-4590-8882-11ae613bb283 · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.187417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.393591Z digest=sha256:a3cb26e66b50c0a044969a863a3162b60f4531850a6e01140131ee5f337c55ae

Observation e1077ab1-cc15-4f07-adb7-24b287ce6577 · outbound

This paper cites Qwen2 Technical Report.

Test-time Prompt Refinement for Text-to-Image Models Qwen2 Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.540099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.540099Z digest=sha256:52e8957939d978e3ed1a9fdf641dce77377500027ed31c8b4f5f195581050b7a

Observation de47594c-35f6-4206-8575-02af20cd8924 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion mod- els.

Test-time Prompt Refinement for Text-to-Image Models Paint by example: Exemplar-based image editing with diffusion mod- els

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.172630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.547403Z digest=sha256:60e729687b25c530e7d6d5fa3a257e49e8da5f4f34bcaca34a60b8d118f2bee8

Observation f64aab34-d6b6-48cc-8bc1-1d97b6cceb17 · outbound

This paper cites Reco: Region-controlled text-to-image genera- tion.

Test-time Prompt Refinement for Text-to-Image Models Reco: Region-controlled text-to-image genera- tion

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.554275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.554275Z digest=sha256:7a7e15a5f079b72e25ed39ca1e0765020cbb1585318e4e6459fb214150b090a0

Observation 958ad8bd-cfe6-412a-8dfe-7f255a1b844e · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Test-time Prompt Refinement for Text-to-Image Models Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.148697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.562603Z digest=sha256:21b01f7d5beb87746b7e4346419519070f6203a64bc2a51c1b4adef1d1796295

Observation 85a3587b-fc2c-4b79-928e-7d51e7409aa8 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Test-time Prompt Refinement for Text-to-Image Models Adding conditional control to text-to-image diffusion models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.569855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.569855Z digest=sha256:7ea11566de6008f21bbcf4f195a3f9f9c5364f9cf9593f3b1cee6eaa7b431ffe

Observation 2513d08e-c38d-43c6-8d7f-5e9f15a9e864 · outbound

This paper cites Controllable text-to-image generation with gpt-.

Test-time Prompt Refinement for Text-to-Image Models Controllable text-to-image generation with gpt-

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.123489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.576873Z digest=sha256:3a8d57f7428df21d2bcfedab47d826bb9198588c5f77cb0764b06f499762ebe2

Observation e7e3f439-9ce0-49a1-891c-c8ddd560bf47 · outbound

This paper cites Dm-gan: Dynamic memory generative adversarial networks for text- to-image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Dm-gan: Dynamic memory generative adversarial networks for text- to-image synthesis

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.108208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.593493Z digest=sha256:e37f998d8096b795ba636130b4025b74579a807b1b798a826a6be6d7a9bc529b

Observation 6c7405c6-5644-493e-8975-3742ad478835 · outbound

This paper cites Controllable Text-to-Image Generation with GPT-4.

Test-time Prompt Refinement for Text-to-Image Models Controllable Text-to-Image Generation with GPT-4

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.585344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.585344Z digest=sha256:0bff0b82619ec6eb6fcc9beae3184070cd10f23b88772fa38586053e1f6bca8b

Observation b152ec5c-59c1-4354-b5b8-6dd4a6c28df3 · outbound

This paper cites Benchmark Datasets We use three benchmark datasets to assess compositional fidelity, prompt comprehension, and generalization: 9.1.1.

Test-time Prompt Refinement for Text-to-Image Models Benchmark Datasets We use three benchmark datasets to assess compositional fidelity, prompt comprehension, and generalization: 9.1.1

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.093043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.601255Z digest=sha256:6c4d80389225806462b2e585545e80dbaabde5ba4f5c2b51d770af61430f5a76

Observation 0c053d58-a99d-4125-a188-9f923a76db6d · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.078443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.609484Z digest=sha256:cd4f4d3c9a5c4b28b77fe8c64959d34c6b56552397ecf3d9de207d1611355b1e

Observation 7bb962e7-2f81-48b0-a76f-fcc33561f370 · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.266861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T15:01:52.367049Z digest=sha256:5bc080fc09fec9085b9dfef114fcd655d4090ac897fc15b23375cbccb8db51aa

Pith citing papers

Observation 5af3e4d7-3f17-45fc-acab-134cacffe0e0 · inbound

Evolutionary Token-Level Prompt Optimization for Diffusion Models cites this paper.

Evolutionary Token-Level Prompt Optimization for Diffusion Models Test-time Prompt Refinement for Text-to-Image Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:40:57.856590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T17:05:00.723170Z digest=sha256:ed7a35107f166d21c1c190e7f799575ff79c0c013b1f796fff60c12d2ebd5115

Observation 5a7d0d4a-ab9d-4063-99a1-367dc1aa8057 · inbound

DuET: Dual Expert Trajectories for Diffusion Image Editing cites this paper.

DuET: Dual Expert Trajectories for Diffusion Image Editing Test-time Prompt Refinement for Text-to-Image Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:29.280305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T06:58:39.254427Z digest=sha256:899c4e08d445b6b9f54f2a882e3fb43051517f93613078bc053484e8fcc24539

Observation ee1caf41-9d5b-4243-9317-9ad86911a2f3 · inbound

DuET: Dual Expert Trajectories for Diffusion Image Editing cites this paper.

DuET: Dual Expert Trajectories for Diffusion Image Editing Test-time Prompt Refinement for Text-to-Image Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T02:13:10.574530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:13:10.574530Z digest=sha256:03b6e0662bfa3a76f11d679b2fd1053c9168d24b909099081ee8eafb1005b72b