Pith. sign in

Paper Citation Record · LEDGER

InsightEdit: Towards Better Instruction Following for Image Editing

As of 15 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2411.17323.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17323 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:19:22.592578Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-16T16:07:53.054355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T16:07:53.165809Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 760af4c4-abb5-40fd-949b-68587d0a0794 · outbound

This paper cites Blended latent diffusion.

InsightEdit: Towards Better Instruction Following for Image Editing Blended latent diffusion

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.336623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.358642Z digest=sha256:a17fc1a572bfe1f2f80741e2a08ab208316a5be1ffafe8746ca6bb0c03e07b9f

Observation 5667b81a-71fb-4832-9207-d985f4552312 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

InsightEdit: Towards Better Instruction Following for Image Editing In- structpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.317771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.364636Z digest=sha256:3a0d729daaa8022bef1bcc79a97779fcff2ebf4178556771d3e4fb26a9c7470d

Observation 1200f067-85c1-4936-b5c0-b6728a3f9478 · outbound

This paper cites Coco- stuff: Thing and stuff classes in context.

InsightEdit: Towards Better Instruction Following for Image Editing Coco- stuff: Thing and stuff classes in context

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.370164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.370164Z digest=sha256:b247e725107469ed19d95c23718e5478bd7ac81471db62ee503ca68ca1ad3e96

Observation 03de9598-47b0-4693-b29c-66309b4f3ff9 · outbound

This paper cites Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing.

InsightEdit: Towards Better Instruction Following for Image Editing Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.376187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.376187Z digest=sha256:3e12789aea2450ed278cfeb4fb32f46be2e8bf3b352059040c2b0bb945b54e24

Observation 2c38f988-f88e-420d-910a-9b79d30b4127 · outbound

This paper cites End-to- end object detection with transformers.

InsightEdit: Towards Better Instruction Following for Image Editing End-to- end object detection with transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.383137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.383137Z digest=sha256:b239cbc2dda673d5b7dca7ca89a80c1838d9b67d351b91cf44af4469d7b86833

Observation 19b28b7c-5657-45e1-a1cd-e182da48246c · outbound

This paper cites Learning to Follow Object-Centric Image Editing Instructions Faithfully.

InsightEdit: Towards Better Instruction Following for Image Editing Learning to Follow Object-Centric Image Editing Instructions Faithfully

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.391172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.391172Z digest=sha256:a98a860b1b2350ab2394e8746b3282296a24d5904666e7d26a215c6a9b2a5c2e

Observation 882868f8-e7e3-4a11-9b74-55ee7ab7ecd5 · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

InsightEdit: Towards Better Instruction Following for Image Editing Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.397476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.397476Z digest=sha256:d0fa9b9dc2b13e59808b758c8e4d7373a7214245ae4677ff2bf0bc61f4e304f4

Observation b6fd0b5c-dc44-48d3-9e4e-2c0151863f78 · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.402699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.402699Z digest=sha256:9799138aef69fda9d46dc4d72777f09b7f0ad5f05d48d4dc574584ae0a8478a3

Observation bb97db91-9e4c-4659-a3ff-27323f5ddce0 · outbound

This paper cites Instructdiffusion: A generalist modeling inter- face for vision tasks.

InsightEdit: Towards Better Instruction Following for Image Editing Instructdiffusion: A generalist modeling inter- face for vision tasks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.253477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.411977Z digest=sha256:4b216b3fa1e147f82a060b2e48d1469174e0d6812d8d8e7929d6f9e048ba006c

Observation d7007744-1940-4c6d-ace5-9b3b6649b8a3 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

InsightEdit: Towards Better Instruction Following for Image Editing Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.417571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.417571Z digest=sha256:2b71720919be293aa73f1e6c100f19fa92e89865f2f10c9a6cf7fccdd575d1b8

Observation a66cee22-bf40-44ed-b424-ba9352f5317f · outbound

This paper cites Image quality metrics: Psnr vs.

InsightEdit: Towards Better Instruction Following for Image Editing Image quality metrics: Psnr vs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.422518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.422518Z digest=sha256:335e7e0e8f6dfb7612ca38272859225d481b6565d7f5a4a340fbe449e8329162

Observation de62810b-1005-42e4-a081-900ba0658e5e · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.427569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.427569Z digest=sha256:757d32220f8cdf94a2abd9edc38ccd111a169dab8a22efcbe5f17653bf2c189b

Observation 6f051b66-ec4c-4448-bd43-30e5e500137a · outbound

This paper cites Diffusion Model-Based Image Editing: A Survey.

InsightEdit: Towards Better Instruction Following for Image Editing Diffusion Model-Based Image Editing: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.433319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.433319Z digest=sha256:461402b01046824918bbda6b0e66d5b64934d03f65fa1cc76e6dfb647af32133

Observation de0a4650-16b1-486c-a064-2d003056840a · outbound

This paper cites Smartedit: Exploring complex instruction-based image editing with multimodal large lan- guage models.

InsightEdit: Towards Better Instruction Following for Image Editing Smartedit: Exploring complex instruction-based image editing with multimodal large lan- guage models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.219965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.438521Z digest=sha256:9d53b53cae9227c7534e660300e381f3f8167b3bc4356477cb08e64508af4923

Observation fe37e4f9-7969-402c-b9a2-e67348baf9a0 · outbound

This paper cites HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing.

InsightEdit: Towards Better Instruction Following for Image Editing HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.444615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.444615Z digest=sha256:e233346224f7d705ec984311c0dc0b6c9ed9c229455edd46a7d2334660e1b07c

Observation 61fb3233-c4ad-4751-9c00-c3c0b97c250d · outbound

This paper cites BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion.

InsightEdit: Towards Better Instruction Following for Image Editing BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.448901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.448901Z digest=sha256:3e417551a52bdefe522beaba169404bcea3f62471878b5fa0844f4fdd65e2f86

Observation ca56dd4b-0d50-43bd-945d-05172560eff7 · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

InsightEdit: Towards Better Instruction Following for Image Editing Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.453180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.453180Z digest=sha256:c66c3942f6bd69a931f54c1f90aa46b72dd60241bfe570486f2de1e72b337087

Observation 1f1a3217-d4b8-461e-8caa-767ac1591bf3 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

InsightEdit: Towards Better Instruction Following for Image Editing Adam: A Method for Stochastic Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.457748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.457748Z digest=sha256:abd81d27fb446cc523ee6d833b3a6a4401fc934a5fe1c76917793e1415994f45

Observation cb616302-0503-4f2d-86d8-46387ef6a077 · outbound

This paper cites Gen- erating images with multimodal language models.

InsightEdit: Towards Better Instruction Following for Image Editing Gen- erating images with multimodal language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.193354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.462634Z digest=sha256:0f1c5a8ca1026f1b4a9acb565cf09a6fb034cb7a151ea14827228008c7c712a9

Observation 8da70949-c9f3-4256-9c3e-a6a79f2ac986 · outbound

This paper cites VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation.

InsightEdit: Towards Better Instruction Following for Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.467898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.467898Z digest=sha256:8125cf1da2b3f3fcb361f0129f863173bb7b3c02baf275f93afbe1f7ed44236e

Observation 2f37db29-59a6-42d9-bc99-caed9a1506e3 · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

InsightEdit: Towards Better Instruction Following for Image Editing LISA: Reasoning Segmentation via Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.472727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.472727Z digest=sha256:4339bb243b99d473d336fb52e37eb14dd549f54024d911a20fb37c932454175c

Observation 0f4ef9b0-c56e-4ca3-80ab-7b70d7955d68 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

InsightEdit: Towards Better Instruction Following for Image Editing Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.478124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.478124Z digest=sha256:be184ccdfa8e6646d49c08bac15cdd79c2fd1decdea5ada69a34c8d0b73efe87

Observation c1e1e345-5fde-4f4a-8013-4cd084f15ac4 · outbound

This paper cites Microsoft coco: Common objects in context.

InsightEdit: Towards Better Instruction Following for Image Editing Microsoft coco: Common objects in context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.482309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.482309Z digest=sha256:0f840087617d607aa91510ceb33cf98889391b34f5f0a8e1b2537e81b49cc070

Observation 7733db1f-c23f-44dc-97a9-151a93ffcf1c · outbound

This paper cites Visual instruction tuning.

InsightEdit: Towards Better Instruction Following for Image Editing Visual instruction tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.487643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.487643Z digest=sha256:a9b0f449af2d65d0c2f9ea4d0d6163a010e4ce165f42366895dc882eb79b241c

Observation cfcc8c28-9204-46af-a22e-70b99b42fe9d · outbound

This paper cites Language Models are Few-Shot Learners.

InsightEdit: Towards Better Instruction Following for Image Editing Language Models are Few-Shot Learners

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.492943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.492943Z digest=sha256:7482d99bb2fc69e9f6b3cb1661ac72c1f9e029a9dbbadfd2f47279b226fd63ee

Observation fdf02c19-15fd-49ec-acfe-6e52678a7234 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

InsightEdit: Towards Better Instruction Following for Image Editing Learning transferable visual models from natural language supervi- sion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.497473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.497473Z digest=sha256:370f6672f147a91e233cdfb91d4adf1c431b0ec8e70399475f2e0a67081f7210

Observation d6c01bc4-8e7d-469c-ae3a-cfdc09c86c80 · outbound

This paper cites Grounded sam: Assembling open-world models for diverse visual tasks,.

InsightEdit: Towards Better Instruction Following for Image Editing Grounded sam: Assembling open-world models for diverse visual tasks,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.501781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.501781Z digest=sha256:2e29629a7b13ce444e66f21b675b9f99230d11c40e48886610df2727184fc8b5

Observation d628a74b-b5be-4599-840f-ade024b74f16 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

InsightEdit: Towards Better Instruction Following for Image Editing High-resolution image synthesis with latent diffusion models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.512037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.512037Z digest=sha256:8427fa6dd1ba2793b69e054689ea28d10b5755c3b881bc82bd0b3e6d2ad9d371

Observation 5a35c3bd-5b08-45d0-ab1c-4a4b750fa8ea · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

InsightEdit: Towards Better Instruction Following for Image Editing U- net: Convolutional networks for biomedical image segmen- tation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.516863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.516863Z digest=sha256:3abcc8ae3d0977e6eee448ee6454acc7eef172f55e8e291043643dd88c408ae6

Observation cf115904-528c-4411-8cb5-77acd804dc56 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

InsightEdit: Towards Better Instruction Following for Image Editing LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.521176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.521176Z digest=sha256:b71c1e5d09257b03e58e51582597738605cfc432954f86a605ea30e45a7c3c28

Observation 32efd130-9350-476c-a03e-facd9d1d5c7a · outbound

This paper cites Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting.

InsightEdit: Towards Better Instruction Following for Image Editing Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.097887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.525213Z digest=sha256:8851f9472043d8c1750700e08f7f1e399f88a73adea1fa711c8b47a7739acf46

Observation 3d182635-52bc-4e1e-9513-7d3cad92a339 · outbound

This paper cites Smartbrush: Text and shape guided object inpainting with diffusion model.

InsightEdit: Towards Better Instruction Following for Image Editing Smartbrush: Text and shape guided object inpainting with diffusion model

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.069036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.529202Z digest=sha256:5399af241fd523469d3d873450594dadd9f119b30f2fbccfdb75af1875760258

Observation f447d16f-197e-454e-8e7a-a1533ad4324d · outbound

This paper cites DreamInpainter: Text-Guided Subject-Driven Image Inpainting with Diffusion Models.

InsightEdit: Towards Better Instruction Following for Image Editing DreamInpainter: Text-Guided Subject-Driven Image Inpainting with Diffusion Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.533203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.533203Z digest=sha256:362275213f047c33f9e27d84fe83aa78bc6e07e748e4b39d15f6b810ad29c1f0

Observation 86b2cd0f-58d6-44a1-8edb-27a4cda6c50f · outbound

This paper cites Qwen2 Technical Report.

InsightEdit: Towards Better Instruction Following for Image Editing Qwen2 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.538501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.538501Z digest=sha256:44683b6713da3a29b1b20cbd761cd98a44f9589364c67ccdcd9bb976db221991

Observation 34aaf49b-0adb-4be0-87b3-f493d10e4670 · outbound

This paper cites EditWorld: Simulating World Dynamics for Instruction-Following Image Editing.

InsightEdit: Towards Better Instruction Following for Image Editing EditWorld: Simulating World Dynamics for Instruction-Following Image Editing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.543846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.543846Z digest=sha256:5f1bd20d1cd362f76e117282062d815ab8651999d7356c6847b5fac3991ded8c

Observation d00bebc9-12d3-4492-97cd-ffaa4de39556 · outbound

This paper cites LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model.

InsightEdit: Towards Better Instruction Following for Image Editing LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.549062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.549062Z digest=sha256:2502d0c3b5b9fba9b1f57791ba0661f6bfe5674a7722f1820f41fb82242e5664

Observation 3ccb8b6a-4558-49b4-9845-1b0de044f3f3 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

InsightEdit: Towards Better Instruction Following for Image Editing IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.555889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.555889Z digest=sha256:0b07ea96815ef4a734335293e24060673697b59217efae2d9005fdb16bc02fd6

Observation 224856ad-fe9b-461a-a225-04d0c5c05e68 · outbound

This paper cites Modeling context in referring expres- sions.

InsightEdit: Towards Better Instruction Following for Image Editing Modeling context in referring expres- sions

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.560892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.560892Z digest=sha256:7cecbe42dc3ca4ca1dfb9ace6851c90634c5099e1723f32baa02fa9ffc80057d

Observation 1c677321-181c-41d1-9f1e-0cfacc1aa4e2 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

InsightEdit: Towards Better Instruction Following for Image Editing Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.044724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.565590Z digest=sha256:cb7dcfc1285286037e74306285963ecb1e83867c7d11aebe958a5007972cf64c

Observation a4dde6b5-45de-46ff-98b7-1eea7f7d18fc · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

InsightEdit: Towards Better Instruction Following for Image Editing Adding conditional control to text-to-image diffusion models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.570581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.570581Z digest=sha256:5ed18fc8be16fa0f6367cdb116b945177bcf1a9d05bbeee85dff02887d48d067

Observation da9bea9e-ba98-4c84-87c3-b640a70532c4 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

InsightEdit: Towards Better Instruction Following for Image Editing The unreasonable effectiveness of deep features as a perceptual metric

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.574739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.574739Z digest=sha256:c69dc19912e580bae46ba5677c32aa3d8f4c63e9b103865b2ce81abd5805a945

Observation a95793dc-a990-4046-84a9-7c30c9ca6501 · outbound

This paper cites Hive: Harnessing human feedback for instructional visual editing.

InsightEdit: Towards Better Instruction Following for Image Editing Hive: Harnessing human feedback for instructional visual editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:19:23.006566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T12:19:22.580148Z digest=sha256:86e22aa7dc48d762161db1d1fcf5393988cfea410ff125066487f2b135e0be88

Observation 4ed2c6d5-5dfa-4d7e-865f-9236dc97d98c · outbound

This paper cites UltraEdit: Instruction-based Fine-Grained Image Editing at Scale.

InsightEdit: Towards Better Instruction Following for Image Editing UltraEdit: Instruction-based Fine-Grained Image Editing at Scale

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.587711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.587711Z digest=sha256:d1aa3397d93c771550a82810dfee4208163d687a3598cc658fd33f5757e5cbd5

Observation c5f2084e-7948-4f04-a07f-57d1120e7dcc · outbound

This paper cites A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting.

InsightEdit: Towards Better Instruction Following for Image Editing A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.592578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.592578Z digest=sha256:de79e1d82f5df69e4b853eaf6a1edfd9dd126edae518046c4bc9eb9987deddc1

Pith citing papers

Observation 17882952-f7a9-410d-b29f-8dd72fff0933 · inbound

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer cites this paper.

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer InsightEdit: Towards Better Instruction Following for Image Editing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:07:53.167573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T16:07:53.054355Z digest=sha256:bbb2f0752e1281d47947d2cdc1a5f5118716836218be3061883aa72de5a7a233