Pith. sign in

Paper Citation Record · LEDGER

An Empirical Study of GPT-4o Image Generation Capabilities

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.05979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.05979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:46:02.746041Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:49:51.431619Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 05e6ee68-f58a-4b09-ad8e-0aaced55cb8c · inbound

TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation cites this paper.

TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:09:10.636658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:09:10.636658Z digest=sha256:51e1c2a2983fb665c0cac6fb118b924c6f793495f8d5c0ac70fe6c52cedfe6a7

Observation 144adc4e-f03e-4738-885a-6ecade1076d2 · inbound

Preliminary Explorations with GPT-4o(mni) Native Image Generation cites this paper.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.746041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.746041Z digest=sha256:5382ee44a472a463bdd3bd5fccda0aacc404adc8901441b2fab7d508739917d7

Observation 45a68c89-357f-459b-8636-d0fa70b96755 · inbound

Emerging Properties in Unified Multimodal Pretraining cites this paper.

Emerging Properties in Unified Multimodal Pretraining An Empirical Study of GPT-4o Image Generation Capabilities

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:23:41.909986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T16:23:41.854132Z digest=sha256:595a0b9339a117a62d6675ae05f0be96f4a98b4045ce62c6e078b13610ecb61c

Observation dd15863d-8cf8-4044-a176-65211f606d62 · inbound

MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation cites this paper.

MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:19.368948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:19.368948Z digest=sha256:e421a30e143da060b4f10c88000f3d8f1ea5ef9c856059cac97c7e880295e6a9

Observation 957fd8ea-0870-43da-8c62-b3105c881098 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.456631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:03.456631Z digest=sha256:0ad26d9646913e8969221c712aaa157e3acf4a668a3fd96bcca7579a459e6c17

Observation eb24732f-0bf0-418c-b2d8-fe4667ea7b09 · inbound

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks cites this paper.

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks An Empirical Study of GPT-4o Image Generation Capabilities

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:57:08.065260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-19T05:55:09.188048Z digest=sha256:9c2194fba6a26c9592109ea5707e3670e2fe45871e977928f1706f333559ce76

Observation 408d4dc0-97c8-4740-929f-1fc45933843d · inbound

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models cites this paper.

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models An Empirical Study of GPT-4o Image Generation Capabilities

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:11:52.762645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T22:10:16.645510Z digest=sha256:03881a455a73ce516760eb9cad0664c89f2e1cc928b5d4c099819308a686eb0c

Observation 632f58af-6a82-4904-98ac-23309dbaabd0 · inbound

HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images cites this paper.

HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:26:22.000134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T17:26:02.023978Z digest=sha256:267ff0ba78ee671c96c64fbc6d2841f30ae9be4b2ee2d4d009311d87982f2f65

Observation 261287a6-2821-4e2a-92aa-90e776afcca7 · inbound

Structural MRI Synthesis for Alzheimer's Disease via Conditional Diffusion on Anatomical Masks cites this paper.

Structural MRI Synthesis for Alzheimer's Disease via Conditional Diffusion on Anatomical Masks An Empirical Study of GPT-4o Image Generation Capabilities

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:29:02.745471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T22:12:46.169295Z digest=sha256:360a249d6ae9d242c7388ab9f51003236c5a91e963728bee5b89b992f925376e

Observation b9179e33-676f-4962-ac36-4aa3b5babbdd · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:49:51.433341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T04:55:42.018348Z digest=sha256:4c4043abb34a8bc4f8d14c2d8f1e76c4ef493b4925b037b11bb59959be7988b4

Observation b3abc640-9e19-4156-8321-7ce8459d9e97 · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T11:44:54.717393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:44:54.717393Z digest=sha256:5286b2db067a6ed436f0f4d5bd3a19ed7233957aeb2d90bf8a3f459995e6f463

Observation 089415a7-d51b-4f27-8f3c-90aa1c8a50ef · inbound

AI-generated Images Challenge Visual Trust in High-risk Scenarios cites this paper.

AI-generated Images Challenge Visual Trust in High-risk Scenarios An Empirical Study of GPT-4o Image Generation Capabilities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T08:52:47.347766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:52:47.347766Z digest=sha256:5ef02d1cc1a79ea4a4d7008763d3dc2f14428d619987129ff797d695ba09ceca