Pith. sign in

Paper Citation Record · LEDGER

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation

As of 13 August 2026, this Paper Citation Record lists 88 of 88 outbound references and 0 inbound Pith citation observations for arXiv:2507.07317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07317 v2

Coverage vector

measured 88 of 88 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:48:46.673046Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

88 of 88 outbound references displayed

  • verified exact2
  • verified fuzzy36
  • unresolved49
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 73803670-10a8-400a-90f2-086b1f14f79f · outbound

This paper cites GPT-4 Technical Report.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:36.859682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:36.859682Z digest=sha256:97fcd78fbf100b12674071683e01c3b47c02e1ef0d75e5559efdc77315826fef

Observation b3b8855b-32cd-4c55-9187-c9a8286ff7c7 · outbound

This paper cites Cos stable diffusion xl 1.0 and cos stable dif- fusion xl 1.0 edit, 2024.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Cos stable diffusion xl 1.0 and cos stable dif- fusion xl 1.0 edit, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:36.939717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:36.939717Z digest=sha256:15e05a92aad860df65051249963197e4a3618af29b743412f241db3c44a2ba31

Observation 0443baac-97af-4ca7-b4a8-32a38837091b · outbound

This paper cites Claude 3.5 sonnet model card addendum.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Claude 3.5 sonnet model card addendum

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.088224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.088224Z digest=sha256:32a80f9b63d09056229d8fe37b2ea9e742fbea1b2892cbe7956eba4af2873a8c

Observation 7760d3d8-8267-4825-8be7-572708b72096 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.185937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.185937Z digest=sha256:774f6a2dfdbd6a4d1ec89847df1658e3b8182f8dd1f0048f6c877582177ea114

Observation 1fd3bd14-795b-438d-9757-48d709760987 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.334872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.334872Z digest=sha256:5a921db6dfc19605c5e14a8807f3504e1d7c0a71046ef850deec8802dad5a564

Observation 1a00ba9d-c536-4cc8-8ecc-4ce29ebee588 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Text2live: Text-driven layered image and video editing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.444474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.444474Z digest=sha256:5f3cf1015aa56de1cf20716e825795da444a2cd2defa1fdaa223d86c4a41587b

Observation c1d90854-d996-4922-ac60-437f3c9f41bb · outbound

This paper cites Introducing our multimodal models, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Introducing our multimodal models, 2023

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.535105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.535105Z digest=sha256:4e186140be1a041ed475fca35ba04646ff17f0279d1165001a6bab192717759d

Observation 6062fa7b-067f-4ecf-a83c-1ba8d949ddd2 · outbound

This paper cites Is clip the main roadblock for fine-grained open-world perception? In 2024 International Conference on Content-Based Multimedia Indexing (CBMI), pages 1–8.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Is clip the main roadblock for fine-grained open-world perception? In 2024 International Conference on Content-Based Multimedia Indexing (CBMI), pages 1–8

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.672421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.672421Z digest=sha256:a7981db18b40f1e4d2738c2bc6feb68c158815f213530b33fce090422ec75ce3

Observation 55afb521-809f-451e-8042-e60ee9701ddc · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation In- structpix2pix: Learning to follow image editing instructions

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.770834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.770834Z digest=sha256:f09cd86f65d41e0e36eb7e4630bf915e1be8ff772af0333b9c81fdfc4ab2c21c

Observation 030b1cbf-c569-4ae2-b528-c72b49bda671 · outbound

This paper cites Lan- guage models are few-shot learners.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Lan- guage models are few-shot learners

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.883358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.883358Z digest=sha256:e6db512213159359a460084694e120b7eefa84d0e983c6cf1410a7409debee31

Observation 31cdb809-1e0f-4bba-8d2f-ac017d213150 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Emerg- ing properties in self-supervised vision transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:37.978585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:37.978585Z digest=sha256:8126569e13bcb170a408ffba08f67af1a783ef2ff5f1b0882871fe3340ed4058

Observation 1c8db78e-3087-448a-b247-82dc018016e8 · outbound

This paper cites MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.097897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.097897Z digest=sha256:f1a96756a62a5051d5c74846d3500f76c07b2c338a0f3717126ec05b263874a5

Observation 318e331a-086b-45b9-a626-cd80b3d6c8e9 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.227993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.227993Z digest=sha256:441d525e2512d6f70177ed5b33b620c48f6a6183ac5ba390ceda3698bb18b32f

Observation 841ff0b5-0a4b-4a14-81b7-9e66f2ce5a7b · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.322722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.322722Z digest=sha256:598579d523b007f970089ea519118aec68a8cf6ded6aa74b6d7e0b93ae4dd5e8

Observation 00cdd155-fda2-4c85-922b-7ba276b3128e · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.482789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.482789Z digest=sha256:40c687b3140a6071e44b3bb4f1c5fdc27c49d7882f52d9ed82d7919882db5938

Observation 5e89a477-5d13-41d2-a3c4-40be01ee4a4d · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.623306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.623306Z digest=sha256:520945eca38845635be4ef1243b0e361a2f7268c3c4a4afa3239991544de2e3c

Observation fcd48602-d109-4347-a20c-69e931a1a218 · outbound

This paper cites The epic-kitchens dataset: Collection, chal- lenges and baselines.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The epic-kitchens dataset: Collection, chal- lenges and baselines

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:57.287543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:38.694374Z digest=sha256:1d107854c9ade7c9e32534356914ea6c49e1575794dff5459f564f29532ddc20

Observation 3df92554-ed0c-47eb-9ff2-99fdf29e2844 · outbound

This paper cites Diffusion models beat GANs on image synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Diffusion models beat GANs on image synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:57.107238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:38.805037Z digest=sha256:2a48d61d1179dd61afc2c2185a429e8a5fdc44f524fd7b60ae6a4a1c61d6b425

Observation aa04f5c5-a8c4-414d-b9ae-cc4f7c45248b · outbound

This paper cites Dreamlike photoreal 2.5, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Dreamlike photoreal 2.5, 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.865429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:38.903563Z digest=sha256:754f5014ea4efbe113ae4973456fbbcbf48b33e3f24ccbe18a2e9008bca56e96

Observation 315354ec-94f7-4a86-b2a4-6c612b76d4d3 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:38.987435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:38.987435Z digest=sha256:eded7702ca8688131b4f23fc93091c1c3acb6bd01bce09773cad0a0d2abf85eb

Observation bf0f9e38-56d8-4007-b7e4-53875d7f8dfd · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.067547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.067547Z digest=sha256:0f44f7bf30b0edea79602f9b0647740e5117f33676122e1d1ae03396e47b676f

Observation 44dd0e9c-a8d4-494b-8314-94d6cda05bdd · outbound

This paper cites SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.186905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.186905Z digest=sha256:210797ef0a4b4353228cf1b6b5a228d47d51b931b5e50790aee966c746a1b752

Observation 812faeec-4c2f-4a18-a3a6-57115f07b1ec · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.256568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.256568Z digest=sha256:149db7404ac40150045134c4ff903a33ad1f23250370c5fc4ee8a11a90b380d3

Observation 939bf505-b7a9-4d98-b509-d71f4ba618f1 · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The” something something” video database for learning and evaluating visual common sense

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.712513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:39.416699Z digest=sha256:4b9c9a338904718679da6a7ae078f6e0fc2baae0a0d645efadf40b67cb7c6417

Observation 7e212353-a222-4a26-a374-efc1aa6b8390 · outbound

This paper cites an unresolved cited work.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:48:56.489778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:39.526651Z digest=sha256:b4b76ab4467b432f2e6eb3ad01a592ffbd3727521f3971d5476b1ac01a457db4

Observation 45219a64-2c0e-4161-becb-94c209e23cef · outbound

This paper cites Multi-Reward as Condition for Instruction-based Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multi-Reward as Condition for Instruction-based Image Editing

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:48:47.999415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:39.649531Z digest=sha256:9889326c33b0f38d5d286c94955589eb83d699a628dbadf50bca6f4e20a0d53a

Observation 2c2f0c50-0fd6-480d-ad51-af2ed04f2177 · outbound

This paper cites VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.761303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.761303Z digest=sha256:1c25fde677c261f5b722d4c0d5d6c33241831f8d5589c0193d9aad81986a2d68

Observation fc613fd7-3279-4765-9322-bd541e9a29f7 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.908378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.908378Z digest=sha256:ca21aabecf03a4eb9fafa81597464a9f7d423d7505e8a77d740a89c3f8e47df9

Observation cc1f3b99-1f29-40b5-a64d-2094ddebe56a · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.054243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.054243Z digest=sha256:82aab30399e101bea2e68e8ccea6b30a3c0e2f6fcd02ca0df68e1149b33bb9c3

Observation e422a841-340b-42c6-9422-4d3695b107ad · outbound

This paper cites Lora: Low-rank adaptation of large language models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Lora: Low-rank adaptation of large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.154473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.154473Z digest=sha256:df1a85f5f2168f99b2f4a699a1d5463f4f7f8b11b38fef21dbc0f465752227da

Observation 75f18027-de1c-4400-b27d-273d352bf943 · outbound

This paper cites Action genome: Actions as compositions of spatio- temporal scene graphs.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Action genome: Actions as compositions of spatio- temporal scene graphs

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:56.262111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:40.217599Z digest=sha256:e93bc8f8e9900b378118f0105c41e9835c862ee6a7accfda12b38322ff820dd8

Observation 2560f2a4-585c-465e-96cb-ecc9f473ae27 · outbound

This paper cites GenAI Arena: An Open Evaluation Platform for Generative Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation GenAI Arena: An Open Evaluation Platform for Generative Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.299542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.299542Z digest=sha256:eccfba00353748944ed731ef15f9e0c07ad34bb66885ab74f61ad0944297ed15

Observation e91b6027-b560-4ee7-960b-ed6ed8549437 · outbound

This paper cites Clevr: A diagnostic dataset for compositional language and elementary visual reasoning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Clevr: A diagnostic dataset for compositional language and elementary visual reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.981983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:40.392429Z digest=sha256:5f9aeeaf3c4fa97ff49ca36d4de949269866ac9cedc48f7b1b4af75200646eda

Observation c37c56fe-5ab7-4df3-a94d-6b5ca9c7359c · outbound

This paper cites What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.774617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:40.508645Z digest=sha256:6782a8d2ba33d45e86faac806f4a3d7ac58f1f50c935d354a921d3b20e19283e

Observation 85d71823-4eb0-4e36-af0b-940643c377bf · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image genera- tion.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Pick-a-pic: An open dataset of user preferences for text-to-image genera- tion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.524988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:40.651286Z digest=sha256:9a14acd74ded27f9a39de3e962941b2440a94fff580b8832e142a0e028140c7f

Observation 13b85a26-e7f0-47f4-996b-90bc6c1e1ba5 · outbound

This paper cites Learning Action and Reasoning-Centric Image Editing from Videos and Simulations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning Action and Reasoning-Centric Image Editing from Videos and Simulations

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.296234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:40.763102Z digest=sha256:1b0b30a17922a08f241cda36048392ddc875eca2620a9f47fe98612c249afdb5

Observation a3bd7ea0-470b-434b-9038-2b132119e827 · outbound

This paper cites VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.925409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.925409Z digest=sha256:29c9a1319b772444a4a18a16df1a70dd801d4ba4722297302278435ee362d86b

Observation e38527a1-a279-4413-bc17-0dbf56168060 · outbound

This paper cites Imagenhub: Standardizing the evaluation of conditional image generation models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Imagenhub: Standardizing the evaluation of conditional image generation models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:55.134883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:41.094239Z digest=sha256:b418ca20760824abc22e2f8f7d2c08c1d5831137160a0d14b476c4be33ae9275

Observation 273f83b8-2f55-4422-b866-b56ff5e11ddb · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation LISA: Reasoning Segmentation via Large Language Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.170936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.170936Z digest=sha256:c261fbb4708c4cbc937645b380f38cf77dd5a736bde50b1b0a167485efa2f320

Observation 5d425981-fd93-4cfc-9610-296a2aac90b4 · outbound

This paper cites Introduc- ing idefics2: A powerful 8b vision-language model for the community.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Introduc- ing idefics2: A powerful 8b vision-language model for the community

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.842646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:41.328554Z digest=sha256:76063d86e05af0b5809b0f3acd03e90cb2c146953ec013475caec4cf0742ee97

Observation c2fa4195-ca75-476c-bac3-732febbd8811 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation LLaVA-OneVision: Easy Visual Task Transfer

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.472557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.472557Z digest=sha256:d43a6935d352bded9526f0ebd7c56a9b2eda46a982c8db2c8496cb87529e78b2

Observation 70baf567-1bc0-4c63-bd42-dea7ad167215 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:41.543773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:41.543773Z digest=sha256:a158346dacf85ac1db42baa750b20f1a2b7a6dbec4975dc39720a85e9c2996a3

Observation ffb0ec51-6caf-4aa3-89b2-77cf5dd41369 · outbound

This paper cites ZONE: Zero-shot instruction-guided local editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation ZONE: Zero-shot instruction-guided local editing

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.663638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:41.666684Z digest=sha256:d95e642e0c7916717c64d4a1ab3d47a264e41e7987e4774c619457a208086724

Observation ed81f119-4d62-4a1e-b5d2-643bc5d08a86 · outbound

This paper cites Rich hu- man feedback for text-to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Rich hu- man feedback for text-to-image generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.408687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:41.778736Z digest=sha256:5d2bba07e70d47e2775c60a61015d70f9d5b909a89acc3fc3155a2bc42d945ec

Observation 35157127-5dbe-4d91-a02a-ebbc8c33f484 · outbound

This paper cites Evaluating text-to-visual generation with image-to-text gen- eration.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Evaluating text-to-visual generation with image-to-text gen- eration

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.181411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:41.932198Z digest=sha256:d566720cadf46ee9d1d2f6d331000a00c0170403138e9d8d7edb974e03c90d67

Observation fe292af6-2fba-4015-b442-285d7662e961 · outbound

This paper cites Llava-next: Im- proved reasoning, ocr, and world knowledge, January 2024.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Llava-next: Im- proved reasoning, ocr, and world knowledge, January 2024

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:54.009132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:42.044640Z digest=sha256:06858e15323b06c5d0e69688e6e27e4135c56e3b33303de51e6d1d0bcbd91ec2

Observation de4403b8-8010-4dce-935d-617b596dc5a4 · outbound

This paper cites Visual instruction tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Visual instruction tuning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.803553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:42.110125Z digest=sha256:6dc65a7f530c01f73ad68bbfd2cf996ca68583ab509f23a643176912c6559f7c

Observation efd8b911-2f75-4d3a-8a9e-9659053125a4 · outbound

This paper cites Visual instruction tuning.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Visual instruction tuning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.239498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.239498Z digest=sha256:0eaf54bf17191ebb8f955d48cadd84e99336a910e292c96de9c95860dd323140

Observation dacaa74e-4b55-4d74-a232-6d874d6d3146 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.388687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.388687Z digest=sha256:09cd73a891568941826953af3e34635fc02857c55d3d088ee601f15a1f445802

Observation 07d900e7-ac67-43a5-9e58-24830fffbc03 · outbound

This paper cites I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.511233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.511233Z digest=sha256:35cbdea5da71c764f50e38747d80d48a1059f4bdc80c14269274025fed218bb4

Observation c4b4b26c-8708-43fa-a6a8-971199d3f9ae · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.647440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.647440Z digest=sha256:c14c11943f8f42f6711e2652dc7305a392331616d9e05aae62377e78822f5279

Observation 3c18578f-3c62-4eee-9cae-6545d5f94e31 · outbound

This paper cites Watch Your Steps: Local image and scene editing by text instructions.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Watch Your Steps: Local image and scene editing by text instructions

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.622904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:42.778364Z digest=sha256:2e3ed61ddde17b27ea44b4cb2f68de0ee75848790958d1b61b936b28ff4b6549

Observation 9a4772f2-fbac-487d-803a-c6815fa42fc7 · outbound

This paper cites Null-text Inversion for editing real images using guided diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Null-text Inversion for editing real images using guided diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.365096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:42.853041Z digest=sha256:ee2ba3dc3cb01e78be041b5d82c4b4bcd967a028e10d6e5676adafe460b38db0

Observation 601f0582-615a-48d7-b0ca-9a285694f71e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation DINOv2: Learning Robust Visual Features without Supervision

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:42.928679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:42.928679Z digest=sha256:ffab7e4bfb74baf393fc1e2e3b6360e481007457a83aa676009d90683dd000ee

Observation 229a0c2d-1188-4b68-92e6-46532039745b · outbound

This paper cites Zero-shot image-to-image translation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Zero-shot image-to-image translation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:53.165587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:43.018862Z digest=sha256:56f0e718a9ca99a97821e429794f96659442d551ac9d27f40c13aeab6809fe87

Observation b3080117-a3f8-4b4a-8d7d-6d700189f94a · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.156247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.156247Z digest=sha256:cb316b330b1e3f648ac4771c42d1ebf988b10f40565706b64eeefcbc5c22e675

Observation 09854faa-21f6-4fb5-8c77-519bf6872ec0 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning transferable visual models from natural language supervi- sion

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.881947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:43.294518Z digest=sha256:3d5cf5349ddf7a482feb386ce933eb0033aef506063564f3702497906cc19459

Observation 09d01c24-5c1d-4661-992e-ce783d53ac0c · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.438852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.438852Z digest=sha256:97c569a726da86d71f221a06e6d358f811c8ca42931eb24ed99df7efec7fd1da

Observation 4e3a6838-2556-48b7-badd-5e638a794f9b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation High-resolution image synthesis with latent diffusion models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.623565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:43.638757Z digest=sha256:6e6b680eed27bce7d2a82384717af41d0ffe2e5fd8b774aa1e642eadff9ab401

Observation 35f71e6b-2bfe-459a-b7b6-b6c4089570f1 · outbound

This paper cites Emu edit: Precise image editing via recognition and gen- eration tasks.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Emu edit: Precise image editing via recognition and gen- eration tasks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.368993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:43.698204Z digest=sha256:d5314eca8f64ceff050c36af1475bb0567994cfc6823e463c3c12d43db7f8bad

Observation b32463c9-32f5-4501-ae26-d93768503d1b · outbound

This paper cites Aria: Advancing multimodal ai.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Aria: Advancing multimodal ai

Reference 61

Resolution
verified exact
raw_fallback, observed 2026-08-06T18:48:47.318442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:43.804495Z digest=sha256:2ab092629868e1f22ece2cd4dc39993e68a2c7603f2e742d1fe9f79618e53a9d

Observation 2152a224-2206-4fba-8722-4e3ed131b37d · outbound

This paper cites Denoising Diffusion Implicit Models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Denoising Diffusion Implicit Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:43.954569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:43.954569Z digest=sha256:847a62c3957d07074f3b9b5c442a3f4bed6025f5dc23ab47fedb4f188779197d

Observation 08b2c90e-c4ba-43fd-98e1-1110031dbea4 · outbound

This paper cites IE-Bench: Advancing the Measurement of Text-Driven Image Editing for Human Perception Alignment.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation IE-Bench: Advancing the Measurement of Text-Driven Image Editing for Human Perception Alignment

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.077282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.077282Z digest=sha256:b9b9d9873e310e1f2c8833e11e98953d820bb28ed615a44dbe0d75361b770d25

Observation 445f8659-fb5b-4566-8c39-0f50ed8a2670 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.180428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.180428Z digest=sha256:f28933fd4a997b1d8ce483805c6f677553af6dc8daa52c528291acee9d11ed6d

Observation 05dfb526-9c39-40de-8b2e-3edbf472e13b · outbound

This paper cites Plug-and-Play diffusion features for text-driven image-to-image translation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Plug-and-Play diffusion features for text-driven image-to-image translation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.143028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:44.295558Z digest=sha256:b99ebc09683cf81004e8a6169d9c28b655d92b2a25373f0d03b7bcf07d2fd0a6

Observation 55527c56-02bd-4408-8d9d-e4e5ecbb3a97 · outbound

This paper cites EDICT: Ex- act diffusion inversion via coupled transformations.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation EDICT: Ex- act diffusion inversion via coupled transformations

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.934383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:44.405595Z digest=sha256:30028e7b5e0f82738b7943c8d0e4dcc03c72c9eef250e6ed65968ae04c3016d1

Observation ee6d5465-21ad-4c2f-b1f7-cbfec2c7a076 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.469725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.469725Z digest=sha256:f018c5b5825c5ed1168db88500bb557266f89b1437a421c0d6f50a2b84846095

Observation a542c4b8-016b-4404-9502-a97b63c1b5ea · outbound

This paper cites Cogvlm: Visual expert for pretrained language models, 2023.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Cogvlm: Visual expert for pretrained language models, 2023

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.522758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.522758Z digest=sha256:cf22377e3fa8e95b5e34de21bcea605b42a39261a5c6090e566bcae14828b2c1

Observation 6c547196-5242-45f6-ad04-474f1adf368d · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Image quality assessment: from error visibility to structural similarity

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.661975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:44.613567Z digest=sha256:4f87fd49ef45097b350c5feac550821c62d3a6b2f744bc53c510b0eb121c36c1

Observation ec506cb9-3f42-4de6-b3f2-5ae60bd1a379 · outbound

This paper cites OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:44.706303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:44.706303Z digest=sha256:e30628d59ef6bbb4c27a8508bd6f55dc26e664b11696d872b321419f5d8f1e5c

Observation 2df7148e-531d-4d25-9337-322a52b82641 · outbound

This paper cites A latent space of stochastic diffusion models for zero-shot image editing and guidance.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation A latent space of stochastic diffusion models for zero-shot image editing and guidance

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.457038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:44.848440Z digest=sha256:05b30cbd91f4f4aae6d89e74765fc42f1c180b7dacd04aa959c82567c9c071bf

Observation a72c81cc-f3fb-433b-a8bd-549b68e2626c · outbound

This paper cites Uncovering the disentanglement capability in text- to-image diffusion models.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Uncovering the disentanglement capability in text- to-image diffusion models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.269087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:44.968134Z digest=sha256:b889094561bc64e53c57a0d8db632dc134d5ce38193af1edbf7d11b9d4faede4

Observation 98f7c310-02ba-486c-9cc2-3360905cbdde · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.124627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.124627Z digest=sha256:bad5caefbc533c7b1cfaacd469eb23acf729186b4ccdcf7f62ace1c815d660d2

Observation 5767bcff-6c33-422a-91c1-7e7ad09062b5 · outbound

This paper cites Multimodal large language models make text-to- image generative models align better.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multimodal large language models make text-to- image generative models align better

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:51.082936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:45.244020Z digest=sha256:85eaad5a3b5241bd330c2444c89d9ce89c2a36a5f9cc924a95074156466b6313

Observation f2883f65-6823-4ff0-ae40-905a7a59639b · outbound

This paper cites Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.322166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.322166Z digest=sha256:90ec9b6a707dda7031076df9211c9b389100037b98baaea0bb202ee470b9b258

Observation a236bf43-d83d-42a8-8ba2-a8ba8ae1c26f · outbound

This paper cites Human preference score: Better aligning text- to-image models with human preference.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Human preference score: Better aligning text- to-image models with human preference

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.816393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:45.426857Z digest=sha256:e211b71288a888f66b0386813814e76030a534147626156e1037b9f68f684c0e

Observation 7bc553e3-d597-4a0d-96a5-3c733c4c4341 · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.561184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:45.562550Z digest=sha256:4a2d16eb83680c9c1c7997e213782b2ed528a9e39c1598ad622051849badeebf

Observation 3e5e715d-3bf9-467b-be0d-808afe6ccbc4 · outbound

This paper cites Inversion-free image editing with natural language.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Inversion-free image editing with natural language

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.697618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.697618Z digest=sha256:0710220762656e4f90bae3ef009646f8229470ecd1acddfdb6ca996d481a9a14

Observation 9666feda-0088-42e3-b9bf-84424452a916 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.861490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.861490Z digest=sha256:7718da546b358e4424231b11d1dcf85d1ddab2ded1620697b200d46ebe69c7d9

Observation a4d73f97-fff5-47e6-9079-cf3839a4e9af · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.288946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:45.909446Z digest=sha256:d323b8701123ebf89f799cd5c8d0d200cd18f23f8ad195c059add45fedfb3730

Observation 48d636dd-8c85-446d-acd7-e7f440c730b8 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:45.958345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:45.958345Z digest=sha256:48bd0f1afd6bfae86fc857b6ac03b07e9a963b2b58442198377887a6b462755e

Observation a979365d-1191-46f0-8491-d82a85753118 · outbound

This paper cites Long-clip: Unlocking the long-text capability of clip.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Long-clip: Unlocking the long-text capability of clip

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:50.019536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.055557Z digest=sha256:f442384f3eaa7eb252907a46ea4bbb8b198f9873a787ebeeeded52a019e3cdb0

Observation 75fd2958-d867-4973-a9a6-ac2e722471aa · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.728766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.208380Z digest=sha256:79f50714b91674233208abc01c39d0aa53e1c925f45b31be3374761dc637da22

Observation 03a1bcbd-f75b-4f0d-a28f-4bbefc6f4f15 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation The unreasonable effectiveness of deep features as a perceptual metric

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.519550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.313062Z digest=sha256:04c8f7b48600b18fd9f1bd21c99270dd10d0e5dc92ebbfe710738711bfe42217

Observation e991b1d4-3cab-4cec-a433-a5ba5a8b1f55 · outbound

This paper cites Learning multi- dimensional human preference for text-to-image generation.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Learning multi- dimensional human preference for text-to-image generation

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.280332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.376211Z digest=sha256:0eb664bc4238d0bb18637b3243067e3304ffa9cd62faa1e4dd7bce4f2311a5d1

Observation aaa52f07-c2a3-4c94-b459-c7435d1050ce · outbound

This paper cites Hive: Harnessing human feedback for instructional visual editing.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Hive: Harnessing human feedback for instructional visual editing

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:49.021181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.496584Z digest=sha256:8e8388f848bd7e1ec590bab0feb46e086dbde2defeafcb840699315baebe36c4

Observation 1fef0756-b515-4779-be69-9de6e264fb0f · outbound

This paper cites UltraEdit: Instruction-based Fine-Grained Image Editing at Scale.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation UltraEdit: Instruction-based Fine-Grained Image Editing at Scale

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:46.618116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:46.618116Z digest=sha256:031da824a665206ef6a1904510e25018a9096c228f8a0f768e3fd13ca4f9cff9

Observation 1f9888cd-41c0-49be-a770-d0ce2f9a8846 · outbound

This paper cites Can you rate how successful the edit instruction [IN- STRUCTION] has been executed from the first image to the second image with a score from 0 to 10?.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation Can you rate how successful the edit instruction [IN- STRUCTION] has been executed from the first image to the second image with a score from 0 to 10?

Reference 88

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T18:48:48.700573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T18:48:46.673046Z digest=sha256:ff94f4ce0ef55971407969b7b1c2b95c52cef59a2aef663a2b99859d685f5ac2

Pith citing papers

No inbound Pith citation observations are available.