Pith. sign in

Paper Citation Record · LEDGER

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps

As of 15 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2501.14046.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14046 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:28:56.209954Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6167753-5f32-4bdb-9b97-dece69809d71 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.748611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.052571Z digest=sha256:11010dd13c515c95d3ca55e5098c59ad469ac312527e26f338beed8a29eeb04f

Observation 77067a6d-39aa-4e86-b442-a160c7d6ab37 · outbound

This paper cites Break-a-scene: Extracting multiple concepts from a single image.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Break-a-scene: Extracting multiple concepts from a single image

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.734847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.057822Z digest=sha256:ca1bf41666ea807c43ebb8d954d383a1cb70822d532c3fd43066a2272dea7cef

Observation be8f0e30-fece-4698-bcdb-3937ce529e8a · outbound

This paper cites Blended latent diffusion.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Blended latent diffusion

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.721423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.061956Z digest=sha256:5919b3ca6f180c1651a8cb33f7ae2ad201448f1e972ef1e34429f4569dfb4e78

Observation 728c1fff-a6f1-4764-9035-40f8d3e28bb7 · outbound

This paper cites Universal guidance for diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Universal guidance for diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.707737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.065961Z digest=sha256:cbf5deaaac9490e62476bf410e77e86841a1639088c806615e22b29f946501ea

Observation 3cfb0b04-3d65-4569-a84d-5aa0f7cf9c14 · outbound

This paper cites Improving image generation with better captions, 2023.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Improving image generation with better captions, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.694474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.070002Z digest=sha256:c9d65f45593622591e16ed1a6f441135e46a5e927359c4c3bfb3f1298e75a47b

Observation b0a9e1b1-23cd-4663-aa66-663c290a1583 · outbound

This paper cites an unresolved cited work.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:28:56.681188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.073967Z digest=sha256:3a3069a398e92cadb4904a76c084939cd6f098e2a666f387ec72797590954386

Observation 2574eac1-a817-405e-851f-b46c423e705b · outbound

This paper cites Language models are few-shot learners.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Language models are few-shot learners

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.668831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.078457Z digest=sha256:b177311ef9bd6862a5acca1fef8609ace7f9d77cc49e0dd7989fe95c475a4e02

Observation 1b26c6b8-dc92-493e-a2ab-213c6b82eddd · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.654311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.082428Z digest=sha256:e63f5df1732476022575b5bb4bac50bc681ab3a454137becaf37ae41b5c17079

Observation fefe8c52-f2a5-4e08-93f9-250a8882c3f5 · outbound

This paper cites Diffusion models beat GANs on image syn- thesis.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Diffusion models beat GANs on image syn- thesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.639725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.086298Z digest=sha256:7c2451c3ac38f6d38fa5ef8f36eebf9451e4dab8c948302244c5cdbdfe994a73

Observation f7a5825d-9a3b-4877-a7dd-c447e3178876 · outbound

This paper cites Diffu- sion Self-Guidance for controllable image generation.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Diffu- sion Self-Guidance for controllable image generation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.625908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.090184Z digest=sha256:52d3c3c7ab764384f7b444fa9e4e299156d10b2bb9efcb7d4ab16eb078c8cdb8

Observation d8827997-e7d0-45d6-8238-7a97cae0c884 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.094553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.094553Z digest=sha256:2ce34941db2829c8b78acad9a6e6dcd1f0cfab3203620adfd63d7869fffd7f3c

Observation 780e8d1a-449f-48d8-92ff-6f32a01adc9a · outbound

This paper cites An image is worth one word: Personalizing text-to-image genera- tion using textual inversion.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps An image is worth one word: Personalizing text-to-image genera- tion using textual inversion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.612301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.099142Z digest=sha256:cfcdbb083ec843510fda8c2d2734620af5e686f8aa5a19f68fb32287a9a372bd

Observation cbea1d66-5243-49e8-8855-017be9d9d569 · outbound

This paper cites Diffusion models as plug-and-play priors.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Diffusion models as plug-and-play priors

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.598469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.103250Z digest=sha256:d6e3e219bcee01e0b9302a986e720a560c1f523ca28f64e10793e88b46bda81f

Observation 15a42542-c963-4cc9-b32d-756ee4eb125c · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Prompt-to-prompt image editing with cross-attention control

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.585383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.107687Z digest=sha256:53deb2b17ed248d72c22c0b557f8daffb6521e420d5b86ef6ef264a529703310

Observation ef35354b-de24-40af-9a0a-1b2134293da7 · outbound

This paper cites Classifier-free diffusion guidance.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Classifier-free diffusion guidance

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.111938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.111938Z digest=sha256:fe27518bfdc80f52337bee98ef2dea999844490afd3ebd1ad867fb6dabe38f5e

Observation 914b68cc-25f8-48b9-83d5-ce8129a356fc · outbound

This paper cites Denoising diffusion probabilistic models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Denoising diffusion probabilistic models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.116102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.116102Z digest=sha256:288b86ef3a11607d2f8844a5d69b4473f34f482ddd789b56e7e9a4e97ef2db6b

Observation 11de8e1a-e3a6-404a-98fb-b0d200e71379 · outbound

This paper cites Mistral 7B.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Mistral 7B

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.120194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.120194Z digest=sha256:8b9ac18b0f00db6e9dda0f5a204ea20989d54edec0291922bf31bd3d238a2932

Observation 888dfb6d-b480-41b3-94cc-e76376368e73 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Imagic: Text-based real image editing with diffusion models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.553992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.124872Z digest=sha256:edf8a06897a2a05185bc3932f9d46d43e0c8a96d72db675604cbc4d97df7e3a4

Observation 61a8b553-bcb5-4023-99f1-eecbef662f6f · outbound

This paper cites Variational diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Variational diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.540130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.129137Z digest=sha256:269aeeefbaac9bfb603dee9733926ccedbe5aaf0d0ac094fbb8fcd96f5b89abf

Observation 6c3b6794-85c5-4013-a5d0-4a24e36c4564 · outbound

This paper cites Multi-concept customization of text-to-image diffusion.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Multi-concept customization of text-to-image diffusion

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.133237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.133237Z digest=sha256:19011702ce762ee200a22716b07d96104b58771872696a16c9d9556b4f0a4216

Observation 3d237b25-88e6-47e2-ae72-8f8ceeee0e56 · outbound

This paper cites LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.136764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.136764Z digest=sha256:82acfc377ebbb45bfc868eda292d10aad0b6d5ea96f9b1f55f362c47bebc8a5a

Observation ab75c67c-b9a5-4f83-a478-f3a9aaf09f37 · outbound

This paper cites Compo- sitional visual generation with composable diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Compo- sitional visual generation with composable diffusion models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.517184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.141061Z digest=sha256:8153d896ae0229fd18bf08d936bd56df46aabe3307aec7a9fb11c0bc9ba0e9f6

Observation 33576f37-d36a-450c-bd66-883292bba580 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Gemma: Open Models Based on Gemini Research and Technology

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.144895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.144895Z digest=sha256:0a27fdfaf5a1724631ba719215ff985d103a9fb47efb2d2669b845ca5f88ee35

Observation 0c4e6579-f3e0-43b1-ae5f-af391c46a444 · outbound

This paper cites Scaling open-vocabulary object detection.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Scaling open-vocabulary object detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.503279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.148676Z digest=sha256:25699aa50d710308f60283d0d6438cd5178fe74788e5d767b47608f435ef2fd6

Observation 596db516-6881-480b-9b51-7194dfb41a36 · outbound

This paper cites DragonDiffu- sion: Enabling drag-style manipulation on diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps DragonDiffu- sion: Enabling drag-style manipulation on diffusion models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.487822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.152276Z digest=sha256:ed5cbfd87e7fd99b5170979b445fe197e75d551b389c4e7c2d6df1846672054b

Observation 77caa4de-f3b4-4a37-ab3b-cd691590c931 · outbound

This paper cites SDXL: Improving latent diffusion models for high-resolution image synthesis.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps SDXL: Improving latent diffusion models for high-resolution image synthesis

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.473292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.155662Z digest=sha256:7e900bcad1c0eb68944b2ddf32b8c5813048b805b749c7fea4fc9d3a6e0ae252

Observation ab422940-b4c1-44f7-9aa4-4240efdd7c19 · outbound

This paper cites Learning transferable visual models from natural language super- vision.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Learning transferable visual models from natural language super- vision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.460327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.159283Z digest=sha256:9b72851bfc8796193898cdd9b252b9f37e000fe01d32f8b3cceeb0ab05b44172

Observation bdc06ff8-cf0e-4e95-baaf-c000b90a9e6b · outbound

This paper cites Zero-shot text-to-image generation.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Zero-shot text-to-image generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.446881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.163552Z digest=sha256:5c196137030658e80728558b6fcb2ab22b7b1dba5faf638a94dc55016bef1778

Observation 94ecbbad-4356-458d-a05a-01280c522a47 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps High-resolution image synthesis with latent diffusion models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.433497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.167716Z digest=sha256:22417df1fca662d1501d579139c82cd90aefddca79388b463bfb895aa945b4e7

Observation 22e2dc97-2617-45ff-806a-3ad8d1e82cb6 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.172328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.172328Z digest=sha256:a69826396de661bb36f30fa743d05c260a23bc06cc77316bb6f777136c9c0da4

Observation d635cb3d-7c17-437d-b6e7-c34c29ba6a71 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep lan- guage understanding.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Photorealistic text-to-image diffusion models with deep lan- guage understanding

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.408594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.176633Z digest=sha256:6ef419222207e8eb22b8350a838eeac73622e7ab4182061f9e04dbbd12ffa22d

Observation 13ce9922-8653-4a78-b5e6-c0d3cda17966 · outbound

This paper cites Denoising diffusion implicit mod- els.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Denoising diffusion implicit mod- els

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.180780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.180780Z digest=sha256:13d0314032febd5fefc97a8222d04f668c357ace0713ea71d902a79e935a749b

Observation c301d897-a153-4a2f-909e-223da9dc4a4b · outbound

This paper cites Score-based generative modeling through stochastic differential equations.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Score-based generative modeling through stochastic differential equations

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.184943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.184943Z digest=sha256:22d1b7cfc5a73a7a0b240ba481cd4aa2bf1d533a05581048264741a8f4d5dd74

Observation e61b4fd5-6962-46b9-a686-2ea88c2c5d62 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.189112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.189112Z digest=sha256:1829b60b9b5c3349669747fc8b6b3e66fe66b48b407a6af69e8825072e2c0894

Observation 39a5271a-a6dd-4b1d-8d46-e6f79eb07550 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Plug-and-play diffusion features for text-driven image-to-image translation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.374856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.193676Z digest=sha256:73a27414b60ad711938500670bb75ccf41ce85bd5b7958c6b1634d56a8e97440

Observation 52aad501-4818-4260-9355-0b57c81a164a · outbound

This paper cites Dy- namic prompt learning: Addressing cross-attention leakage for text-based image edit- ing.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Dy- namic prompt learning: Addressing cross-attention leakage for text-based image edit- ing

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.360036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.197814Z digest=sha256:4db95047aa6007761728793ac16914d2de0019f76809c4243b919520d0d44393

Observation fd16baf5-8fc1-430c-a23e-4259795f900f · outbound

This paper cites Self-correcting LLM-controlled Diffusion Models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Self-correcting LLM-controlled Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:28:56.202130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:28:56.202130Z digest=sha256:b786ae836e3ebf2befeb1762dbc4330ab22533f0025b02872eb6421243fc0ac4

Observation 018fbb71-7047-4334-90a9-c711368fbe90 · outbound

This paper cites Adding conditional control to text- to-image diffusion models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Adding conditional control to text- to-image diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.345042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.206392Z digest=sha256:5d64b6f8dd57574a3697b67053cacfc64a2042f885ecf04bbe98f12df0449122

Observation 24088f72-890c-414a-815a-facae9994ccc · outbound

This paper cites Sur- adapter: Enhancing text-to-image pre-trained diffusion models with large language models.

LLM-guided Instance-level Image Manipulation with Diffusion U-Net Cross-Attention Maps Sur- adapter: Enhancing text-to-image pre-trained diffusion models with large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:28:56.331088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T15:28:56.209954Z digest=sha256:b943c7a6bf9748959d7c292ed2ac227f03ae7faf55f7760b09ac16bede80e780

Pith citing papers

No inbound Pith citation observations are available.