Pith. sign in

Paper Citation Record · LEDGER

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 25 inbound Pith citation observations for arXiv:2507.21033.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21033 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:05:39.722236Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:41:13.116725Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.737452Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1e13080-fddb-45bd-9fee-6173d472ece9 · outbound

This paper cites Qwen2.5-VL Technical Report.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.658388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.658388Z digest=sha256:e47dd6eb8f162d7d9329cdbc5ff2e4222cfac6b6dce53c6a3b18f11c1ffb2eac

Observation c000a00f-84f8-4056-affc-74cc710698c7 · outbound

This paper cites GPT-4o System Card.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset GPT-4o System Card

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.678552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.678552Z digest=sha256:3992efdb04a3a3fb5a7add38c0b2355d8ef41e4f41c2fcf2857a0de4b03ab5ed

Observation 32fc32f2-8c12-40a3-b483-b9d867bb167a · outbound

This paper cites UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.686427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.686427Z digest=sha256:2c210180d710884ffd3a36b3ed6c6a3fa3488fc56b40c219519a999516ce93f3

Observation 063b186a-9143-4448-89ef-493090b7bb25 · outbound

This paper cites Flow Matching for Generative Modeling.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Flow Matching for Generative Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.690728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.690728Z digest=sha256:0f893d672ca62118a774fa6b6b6093e1deee6f4fa1e549504d8a63ee5121d778

Observation fe926506-216b-4320-b9fb-1b1ec0734e3c · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Step1X-Edit: A Practical Framework for General Image Editing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.699110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.699110Z digest=sha256:69580e3f0ad009e0d11b33d1537b9efc468df819ddfa9b11d8610f1651e86be0

Observation 2240c72c-18a1-4a4a-b969-a0ada78022ec · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.702682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.702682Z digest=sha256:5af496c7dcd4c120e935c45631b0655708f194f5d4c738bb21251ac2a5d8b113

Observation 2d3e8b29-c081-4409-a199-3f34a5ce3a5b · outbound

This paper cites SeedEdit 3.0: Fast and High-Quality Generative Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit 3.0: Fast and High-Quality Generative Image Editing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.710714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.710714Z digest=sha256:7d2d40541b21c4062d32be840542a68e1d627c4f66b0281cbd064e73b894cb08

Observation 74e7d532-c840-4003-b8b1-7aa5177ed517 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.715086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.715086Z digest=sha256:5c5bdb9fa9d685bbfd7df0fc6b1b3e4a13061720d0cdb610e2b1369fc048a5ed

Observation a2ebc351-07dc-4a79-84cd-95e12a8ff3e2 · outbound

This paper cites $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.718864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.718864Z digest=sha256:5783a7155451d777fbcfaec773b36c41e5a95962dda0d7c01fe408acbdabe549

Observation 2b14b3b4-0542-41cf-880a-1293cd42c1b2 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.722236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.722236Z digest=sha256:d5b91cdc99d1edda20ca861f343bc6f654b9c3c9d612fd3906422250c54fc91d

Observation 309f22f9-9241-4f0d-b7d2-f8eccc15f32e · outbound

This paper cites ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.665891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.665891Z digest=sha256:d85aae49ec80af353dcf67c1bf2314531e53e54587ec9d423555f485df47911c

Observation d2ad8c8e-b29d-4508-a2da-1a63cf7db32c · outbound

This paper cites SeedEdit: Align Image Re-Generation to Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit: Align Image Re-Generation to Image Editing

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.706618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.706618Z digest=sha256:b1d4be3126a51f9f0660bde05d1a7892879a66fd1e40711d2a1f8f0e40812cb2

Observation 35079285-d3c6-4ba9-b891-e25033bbb6e7 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.662514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.662514Z digest=sha256:c95a175d56f786d2fb29c1899b22ca0c4d3f17dd28dc65c2261b9fefc30980d1

Observation b9ae3ed5-2307-4878-96dc-04bfe799296a · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.682539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.682539Z digest=sha256:72eb4aa61c8ec43b2037061cbf579f7a2904f9459ad9766d108952ea70b48700

Observation 90a86f81-562d-40d1-afcb-7c216e4df9a0 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.669870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.669870Z digest=sha256:c0e44e8f768ecfa70a6278e5a6b5250ca21ed428ed9cae918787ec7bd4f16655

Pith citing papers

Observation 16cd742a-e808-4f03-b47f-51c663bed9a7 · inbound

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning cites this paper.

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:52:07.965603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T05:47:19.552825Z digest=sha256:5c2fc9cac3531c7071811a701b1e39f2e93d995f347fbf72c3c6603b59a247ea

Observation 52408d33-734c-434f-824d-073db230f35b · inbound

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning cites this paper.

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:41:13.116725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:41:13.116725Z digest=sha256:0eccb6843a20d618274d90e2675a2b335e00337e9ebf23d87d4125b0542709af

Observation ccef5db5-15eb-4134-9cfd-857a8cfa91a8 · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:16.254186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:16.254186Z digest=sha256:c92dcc657541e1824e04aa251efff63bab953ca086f2176beaf5cffdbecec2d1

Observation 48524113-4174-4be8-8849-bf52cb11b5bd · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:08.215305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:08.215305Z digest=sha256:464042329672fe1211fcebdaa00c5474da8ce20f1798757f427daa19b0934605

Observation 4c23dc36-ce14-4211-a05d-f86970344e92 · inbound

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation cites this paper.

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T15:39:40.585648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:39:40.585648Z digest=sha256:245351f4610632455fa284d85af9eaca066f469202827530f2aa513da82f8f5e

Observation ac43947d-a470-4f97-b4ea-fd7ac54aa1fb · inbound

Emu3.5: Native Multimodal Models are World Learners cites this paper.

Emu3.5: Native Multimodal Models are World Learners GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:12:13.583455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T01:12:13.426640Z digest=sha256:78e0f5329da0004689851b10865fc9c4889d999e7ca0a873d72a82a63f8f7917

Observation f30cd611-26b5-44df-9794-0973834a8c51 · inbound

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models cites this paper.

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T16:21:37.664980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:21:37.664980Z digest=sha256:a052f7451ce141e1bdcd6c947cf66619d73d6c7609c11ae4b5274cc321b561f3

Observation 7bdf3801-08f8-4b70-944e-3ce1704759fb · inbound

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation cites this paper.

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:11:11.494014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T19:10:47.425041Z digest=sha256:0817315c7118811b07b30077e63aa3fafeea2cf9195bc755481d80ffc61b1bd4

Observation f74e0d7c-be25-4471-8ba8-b45fe48e28f6 · inbound

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing cites this paper.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.474554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.474554Z digest=sha256:2966bb5742ffc32def2bf7f20d2823faabbfc28329b21f0249de92aac6707bf1

Observation a13baf45-5f93-43db-a4e7-2387d978ae3b · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:76324fcb15857bc23c83e5b7d225cd4c0f83dbbcb4b32ddf5841439f55c3a3a8

Observation aefbd785-8e73-4f20-ab8a-72683fc2803f · inbound

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing cites this paper.

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:00:50.312722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:23:50.614589Z digest=sha256:4efe4084868dac9307810dd5f79ec129b725365b90a4f826c2ffaa70aee3c314

Observation cd5cec37-e215-440a-bf19-4592609a2b1e · inbound

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details cites this paper.

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:20:53.908150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:11:43.172296Z digest=sha256:2e67c4519f0d7534b7266b283d9cb99bc5dd6a69f4f77dcc2af3084afd47dd80

Observation 0bd309f4-2e51-49ea-b1e5-d849e921ec5c · inbound

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation cites this paper.

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:57.551819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:39:59.791758Z digest=sha256:1fbdc8932238dd1936cdd082f93d2378ae31816c32c695fa930e59174970b462

Observation b2185e1f-8287-4024-a27d-7a576bb61fd4 · inbound

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance cites this paper.

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:06:00.083616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:15:23.578176Z digest=sha256:eda41184f28c1d0c4ccff0285aeafe313544db4e0271f65fd3e997e218fa6c70

Observation 27b246a9-c1b2-4da7-8f1d-e3d35af300b5 · inbound

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing cites this paper.

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:21:55.104500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T07:21:41.483427Z digest=sha256:d1271e970f871364948529dd2fd0d6342ed759a761a0cb0991472a8bb14022a5

Observation ed312bff-e852-417c-8819-982019823f64 · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:55:43.713579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T17:57:08.606559Z digest=sha256:cdeea54e199b35708c019740607aa593c205c42c4790a08794ad17cbfa4f9f7f

Observation e3eddba3-60ac-422b-a8cf-c18f06e73ced · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.686633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T08:15:58.020894Z digest=sha256:ba8938030c433eb9025061dc733f11ad1d21d5ba2148e506f23980d684d34fe1

Observation a5f3e129-d018-4598-a752-d6b87a31c88e · inbound

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality cites this paper.

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:07.965833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T15:04:41.518195Z digest=sha256:63f28af5de2ee68e9bbb9a53c01242c32f2597f52bf9421a41576a9625cd7648

Observation b3805d8d-e440-43ad-8e55-c33dfb307e93 · inbound

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis cites this paper.

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.029644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T21:41:10.852265Z digest=sha256:69b146c5ab77bcfdb376b84a7047ce631da1948aea797e93ecc8cdfa350908f3

Observation dedc9623-590f-480b-8b92-1112562b4e66 · inbound

TextSculptor: Training and Benchmarking Scene Text Editing cites this paper.

TextSculptor: Training and Benchmarking Scene Text Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:19:39.343635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:16:43.756525Z digest=sha256:d0cf3d7a5cc285b92f14f8939b4dac1e6c6494b23bbbc9d452264d77c749f364

Observation 24c23ea0-e9ad-4a9e-a638-d45473cf0782 · inbound

Bernini: Latent Semantic Planning for Video Diffusion cites this paper.

Bernini: Latent Semantic Planning for Video Diffusion GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:41:10.427081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T06:39:47.124605Z digest=sha256:851799122727e3258e7274282e2ba751fce60486c810fa0f81c7d9ce4be5cd85

Observation df9272dd-e106-4133-8111-c198da1b2927 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:44.738743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T09:08:25.661515Z digest=sha256:bd936fe27c74e184663dc1a47fd89c4863a9cbfe6bc0de779aee99c28ebbd40b

Observation 1c38d0f6-d031-425e-8674-539fbdf5c845 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:19:02.587384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T23:15:09.253879Z digest=sha256:6e51f73fdad2e6613a6839be5131482c555485d333a91832c9e78fb2b4e66765

Observation 8d95ddc0-def7-4859-a94a-c6d646f64f28 · inbound

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry cites this paper.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.773307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.773307Z digest=sha256:cd3b96c2f5083ed7b6e19b68fd8dd38f417cbc1a0722a9990af67f3404bea0cc

Observation e875a0a6-b8b5-442d-8947-11f2f9821327 · inbound

Illuminating Visual Identity in Universal Multimodal Embeddings cites this paper.

Illuminating Visual Identity in Universal Multimodal Embeddings GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:43:17.318655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:43:17.318655Z digest=sha256:437567545efd34233e1ecab7563fb65526f6fd41f0f32a8ddbd2bcda7be30eef