Pith. sign in

Paper Citation Record · LEDGER

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

As of 17 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 25 inbound Pith citation observations for arXiv:2507.21033.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21033 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:05:39.722236Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:41:13.116725Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.737452Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1e13080-fddb-45bd-9fee-6173d472ece9 · outbound

This paper cites Qwen2.5-VL Technical Report.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.658388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.658388Z digest=sha256:8b5b9fdffa67ce24835f1468a471905d0b66a7478b788ead8a5901eb1622d5b3

Observation c000a00f-84f8-4056-affc-74cc710698c7 · outbound

This paper cites GPT-4o System Card.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset GPT-4o System Card

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.678552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.678552Z digest=sha256:14563a13965da07c02949a9921ee47e96e8f0519e2a125770bc93604732cb1a7

Observation 32fc32f2-8c12-40a3-b483-b9d867bb167a · outbound

This paper cites UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.686427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.686427Z digest=sha256:35e69c5352b8fadc6f071a11edce72418afe8f8e3760cd6700d2699d0042b863

Observation 063b186a-9143-4448-89ef-493090b7bb25 · outbound

This paper cites Flow Matching for Generative Modeling.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Flow Matching for Generative Modeling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.690728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.690728Z digest=sha256:6365fcde48e89c4dbd34e363bbdd7ee5707ecfe3180f74d1d915ed224035517e

Observation fe926506-216b-4320-b9fb-1b1ec0734e3c · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Step1X-Edit: A Practical Framework for General Image Editing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.699110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.699110Z digest=sha256:43c83b53ff02bb8ee9e9f45295953185637543523b876ba4ae392519c5e27483

Observation 2240c72c-18a1-4a4a-b969-a0ada78022ec · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.702682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.702682Z digest=sha256:ca7882da8fab2ab2b9f29853012f0b6c9e83fce12d45a36b3fac9669c0612162

Observation 2d3e8b29-c081-4409-a199-3f34a5ce3a5b · outbound

This paper cites SeedEdit 3.0: Fast and High-Quality Generative Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit 3.0: Fast and High-Quality Generative Image Editing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.710714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.710714Z digest=sha256:cf2f5fc61be95d10221fc5e0ddc1b0e1bd761eae896c367bdfb4f496c02873ce

Observation 74e7d532-c840-4003-b8b1-7aa5177ed517 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.715086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.715086Z digest=sha256:3d43e6e14b35fdc35bae9e8da3511ffac436dabc60e379858b96bf9ba41efb1e

Observation a2ebc351-07dc-4a79-84cd-95e12a8ff3e2 · outbound

This paper cites $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.718864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.718864Z digest=sha256:6ba6ddc0c4bbf28b871d738b35478fbf24baca54461d3e3db85140c5d3d2fab1

Observation 2b14b3b4-0542-41cf-880a-1293cd42c1b2 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.722236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.722236Z digest=sha256:9d4255c3190a4df6ed0c8e7ad364a5011e4c040d464b2354c50c7d66254d85bf

Observation 309f22f9-9241-4f0d-b7d2-f8eccc15f32e · outbound

This paper cites ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.665891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.665891Z digest=sha256:1b1c13208827dfb4e1441d383e7bf7b6a300e57cf44e0b273153711cc619f56b

Observation d2ad8c8e-b29d-4508-a2da-1a63cf7db32c · outbound

This paper cites SeedEdit: Align Image Re-Generation to Image Editing.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset SeedEdit: Align Image Re-Generation to Image Editing

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.706618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.706618Z digest=sha256:69edd3efd2fdc76f094ea7aa3fabf820d540028b72e9556dd761ca13ff01160d

Observation 35079285-d3c6-4ba9-b891-e25033bbb6e7 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.662514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.662514Z digest=sha256:cdd9c8e41cc4144623a4de6e7dd2bd3cdbd6ff78381919132e92931c458e3060

Observation b9ae3ed5-2307-4878-96dc-04bfe799296a · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.682539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.682539Z digest=sha256:f79dcb931e4df9c80e51f572db859ad38df58b2fc2576b40fb82374dcbf24a8f

Observation 90a86f81-562d-40d1-afcb-7c216e4df9a0 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T13:05:39.669870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:05:39.669870Z digest=sha256:d5c376309cd7121aef6eefb463b571bd70f502beb12b2c78d589c72b15f0e0d5

Pith citing papers

Observation 16cd742a-e808-4f03-b47f-51c663bed9a7 · inbound

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning cites this paper.

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:52:07.965603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-19T05:47:19.552825Z digest=sha256:70e02e38c67c3b22d8bdd4fcf847eb757621f29971aacf56d0309cd65210acd3

Observation 52408d33-734c-434f-824d-073db230f35b · inbound

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning cites this paper.

TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:41:13.116725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:41:13.116725Z digest=sha256:8a4fc990678a4123ab97b0ccf8bafe9b352374039df628f734f5da9917d3002c

Observation ccef5db5-15eb-4134-9cfd-857a8cfa91a8 · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:16.254186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:16.254186Z digest=sha256:03fc3469d8e7539596691b6b8557e4135cfff5506d45d8b2403edbbafc0ee204

Observation 48524113-4174-4be8-8849-bf52cb11b5bd · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:08.215305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:08.215305Z digest=sha256:5e4884bf0d13f900f02d03c7fe4fb30ab379de5f65db8c79e4a9fc3120ff7557

Observation 4c23dc36-ce14-4211-a05d-f86970344e92 · inbound

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation cites this paper.

Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T15:39:40.585648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:39:40.585648Z digest=sha256:c65d2d697594be3d2777835d089bb637d698ea99a894e420e133cb4e8b7d5d23

Observation ac43947d-a470-4f97-b4ea-fd7ac54aa1fb · inbound

Emu3.5: Native Multimodal Models are World Learners cites this paper.

Emu3.5: Native Multimodal Models are World Learners GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:12:13.583455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T01:12:13.426640Z digest=sha256:a34761bcdfb9cf1ff2acb908ea397c61196d5fdc61ae3a4d30fc7ef1d7652cf7

Observation f30cd611-26b5-44df-9794-0973834a8c51 · inbound

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models cites this paper.

Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T16:21:37.664980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:21:37.664980Z digest=sha256:d3a19e221071a87e5c95bf45553774faa0b26fcae5ce7089191d1ea131600b26

Observation 7bdf3801-08f8-4b70-944e-3ce1704759fb · inbound

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation cites this paper.

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:11:11.494014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T19:10:47.425041Z digest=sha256:34abf2938f6083ed816900e11506b4c0ea8f81b61038b10330eb9aee45253945

Observation f74e0d7c-be25-4471-8ba8-b45fe48e28f6 · inbound

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing cites this paper.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.474554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.474554Z digest=sha256:1c44df3a7f5b295bb59853f4744301e650f58db74f31d5b1d212e548872bec96

Observation a13baf45-5f93-43db-a4e7-2387d978ae3b · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:61d79d243966b5a71daea792be95f8fa4f5f44fe16ed9b4428bd7c028b60b898

Observation aefbd785-8e73-4f20-ab8a-72683fc2803f · inbound

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing cites this paper.

SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:00:50.312722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T19:23:50.614589Z digest=sha256:f6a3559568d9ad8cc266204e70639a54f9ae96f99514f1c2b80b73ea9b620195

Observation cd5cec37-e215-440a-bf19-4592609a2b1e · inbound

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details cites this paper.

RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:20:53.908150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T19:11:43.172296Z digest=sha256:964c4f5633f491765e4c3b0042ae1511ad457409233ee8f61ef4cda0a16f0e4f

Observation 0bd309f4-2e51-49ea-b1e5-d849e921ec5c · inbound

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation cites this paper.

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:57.551819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T17:39:59.791758Z digest=sha256:68d48b569909ed66def1b14d0e15e6673791f87097d423d6f36f19dc13cf0caa

Observation b2185e1f-8287-4024-a27d-7a576bb61fd4 · inbound

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance cites this paper.

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:06:00.083616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T16:15:23.578176Z digest=sha256:9a603997d38209bd10cb9cb1f92aaabb39fe5fefebc53ed52b0795806dfa9129

Observation 27b246a9-c1b2-4da7-8f1d-e3d35af300b5 · inbound

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing cites this paper.

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:21:55.104500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T07:21:41.483427Z digest=sha256:fc69d0f202e0f3212a83be505e8d011506ccf6433ff293c42842aea6cd397a94

Observation ed312bff-e852-417c-8819-982019823f64 · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:55:43.713579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T17:57:08.606559Z digest=sha256:f98572554426654609df3759ddc29b6d12c9225c5d0628895c315a3f6926fb36

Observation e3eddba3-60ac-422b-a8cf-c18f06e73ced · inbound

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation cites this paper.

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.686633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T08:15:58.020894Z digest=sha256:1f31e2cf1785a25da821271380b005dabbd32b2209b6fb85d065c162c22a6ae2

Observation a5f3e129-d018-4598-a752-d6b87a31c88e · inbound

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality cites this paper.

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:07.965833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-08T15:04:41.518195Z digest=sha256:5157e6b9e172342bd2dd1414f0ffb36476120668ae188d76506efbb189aea111

Observation b3805d8d-e440-43ad-8e55-c33dfb307e93 · inbound

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis cites this paper.

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.029644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T21:41:10.852265Z digest=sha256:a815440b87b148ef313b7e435d3287d2c394c31c215179e9cf405cbb5bdf82d9

Observation dedc9623-590f-480b-8b92-1112562b4e66 · inbound

TextSculptor: Training and Benchmarking Scene Text Editing cites this paper.

TextSculptor: Training and Benchmarking Scene Text Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:19:39.343635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T05:16:43.756525Z digest=sha256:d97a16240c00f8f96b42d1f857e02b0b5417b605f902c825701c45d0c828b230

Observation 24c23ea0-e9ad-4a9e-a638-d45473cf0782 · inbound

Bernini: Latent Semantic Planning for Video Diffusion cites this paper.

Bernini: Latent Semantic Planning for Video Diffusion GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:41:10.427081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-22T06:39:47.124605Z digest=sha256:5ee3323d8fc6d3d59f6c0dc5cb41fca874ee0a5773fccefeebe303cc92771644

Observation df9272dd-e106-4133-8111-c198da1b2927 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:44.738743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T09:08:25.661515Z digest=sha256:85d4c049c1835cc1698fb9179d6c1946ad2da06fc3e98b0af869166f2429f6fd

Observation 1c38d0f6-d031-425e-8674-539fbdf5c845 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:19:02.587384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-03T23:15:09.253879Z digest=sha256:925e7b07fb4c4ab508d7dc8bcd46ac7883dde3438ec564b21affc70499b7521c

Observation 8d95ddc0-def7-4859-a94a-c6d646f64f28 · inbound

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry cites this paper.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.773307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.773307Z digest=sha256:3e4f7b7f2338e923c63acb33b60038002a416f262193a6697daa765b1c55055c

Observation e875a0a6-b8b5-442d-8947-11f2f9821327 · inbound

Illuminating Visual Identity in Universal Multimodal Embeddings cites this paper.

Illuminating Visual Identity in Universal Multimodal Embeddings GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:43:17.318655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:43:17.318655Z digest=sha256:609d7922209d32bca99e32698b3c80bb518cb5eee7e38d1c5f7f3e0d7515d3f2