Pith. sign in

Paper Citation Record · LEDGER

Preliminary Explorations with GPT-4o(mni) Native Image Generation

As of 16 August 2026, this Paper Citation Record lists 100 of 219 outbound references and 3 inbound Pith citation observations for arXiv:2505.05501.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05501 v1

Coverage vector

measured 100 of 219 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:46:03.070172Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T05:30:32.163273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T09:34:57.138862Z

Reference resolution

100 of 219 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2f74821-6a5a-413a-b5c7-9576da5ff7dc · outbound

This paper cites Defocus deblurring using dual-pixel data.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Defocus deblurring using dual-pixel data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.638890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.638890Z digest=sha256:b685bed3215af251e892812be590d86d602a222bdb04f3597ac1037dab3e25c8

Observation 329d17a9-d9db-4cc5-ae02-d9f49f503222 · outbound

This paper cites Learning to reduce defocus blur by realistically modeling dual-pixel data.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning to reduce defocus blur by realistically modeling dual-pixel data

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.644112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.644112Z digest=sha256:198c8e177d08d4cc805eb9074423ec692da91666905ca2a5854ef7705af5589e

Observation 00c433ad-b03b-4ad2-a622-27d6839b8735 · outbound

This paper cites Ntire 2017 challenge on single image super-resolution: Dataset and study.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ntire 2017 challenge on single image super-resolution: Dataset and study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.648456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.648456Z digest=sha256:c83c20c373854bc03519704cac0e44fea665b73a38b8c03af2d5d3d4cf0710fb

Observation 3aaddf03-7b7d-43ff-80ed-c46a7c64207d · outbound

This paper cites Dream360: Diverse and immersive outdoor virtual scene creation via transformer-based 360 image outpainting.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dream360: Diverse and immersive outdoor virtual scene creation via transformer-based 360 image outpainting

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.652848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.652848Z digest=sha256:6f0200068c2d903517384df276cc64fa6257a8c90836c27ed67deb9ad2cbf180

Observation edf06fe7-0192-412e-9f88-721dccf7e23d · outbound

This paper cites Single-image reflection removal using deep learning: a systematic review.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Single-image reflection removal using deep learning: a systematic review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.657073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.657073Z digest=sha256:a44df67f37090f1ad82426621666d5b7c65dd031709161aea4771bdb84bea0db

Observation c13648ec-b1fb-479e-bdc2-850edac22983 · outbound

This paper cites 2d human pose esti- mation: New benchmark and state of the art analysis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 2d human pose esti- mation: New benchmark and state of the art analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.661347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.661347Z digest=sha256:1d9751888f7eb04e42c9ae8d31315b8fc676b63bac23d73b4970f421f033c2f6

Observation 7ac9330b-c9ed-4d16-ac33-65b32d158f4d · outbound

This paper cites Change detection techniques for remote sensing applications: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Change detection techniques for remote sensing applications: A survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.666252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.666252Z digest=sha256:b44e20b84fbd767d241c5a736562c435599e6aac74171b7e2cd06af322c28865

Observation 3ae4da77-3dd8-433d-b42d-94571df53b03 · outbound

This paper cites Rethinking inductive biases for surface normal estimation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Rethinking inductive biases for surface normal estimation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.670557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.670557Z digest=sha256:6b4366786dbfdf59b01ebda44e5e1613f3a3348ee84aa6f13b3ceaf3b871f8d1

Observation 030a4bb0-bb52-4a85-a8a4-eb7fd7198e7b · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Preliminary Explorations with GPT-4o(mni) Native Image Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.674974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.674974Z digest=sha256:482a0e193cb1bc00d4ab31a08da5473eab03cba9eddaca19b4a90ca80d8de868

Observation d0518307-9150-49df-a996-025b8665ed12 · outbound

This paper cites Diffusion Models Through a Global Lens: Are They Culturally Inclusive?.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Diffusion Models Through a Global Lens: Are They Culturally Inclusive?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.679724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.679724Z digest=sha256:2761b7914ea1f12d5a1c1b338b743a169cd2685a1a2ce4124930c3ec669964a9

Observation dd4f7c4f-44d3-44d5-844a-01e6252cd2f6 · outbound

This paper cites Adabins: Depth estimation using adaptive bins.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Adabins: Depth estimation using adaptive bins

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.684520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.684520Z digest=sha256:769d77ce441bfe21285819b887a5a4de6de4a199ed265e88e209fa73ee80d4de

Observation d20c8806-8848-479d-becf-562849163015 · outbound

This paper cites Ledits++: Limitless image editing using text-to- image models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ledits++: Limitless image editing using text-to- image models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.688621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.688621Z digest=sha256:320cbaaeadd8f267adf42d68b8261fda87534334c68ad9df337dadf4bfaf6c62

Observation 1099dc0d-9e68-429b-bead-154f303e1082 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Instructpix2pix: Learning to follow image editing instructions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.692633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.692633Z digest=sha256:c1e359722c91a8090dcef57842865ecd2d3378cecf680994c03909c59e61c270

Observation 18e18725-8c52-4fe0-9862-9367d3146fb8 · outbound

This paper cites Learning to generate realistic noisy images via pixel-level noise-aware adversarial training.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning to generate realistic noisy images via pixel-level noise-aware adversarial training

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.696479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.696479Z digest=sha256:366256695b8b8f01550da9d54576d20a17618566fab6dffe838143054abda39e

Observation 79f1762c-f203-41b7-b1d1-344874892682 · outbound

This paper cites Decoupled Textual Embeddings for Customized Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Decoupled Textual Embeddings for Customized Image Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.918935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.700930Z digest=sha256:91041f9cc09de8d58e66eebab7230ef53b5e75c11b97906aafe9779d20c7ea34

Observation 1dd26251-8747-4f02-b62e-4449266e6874 · outbound

This paper cites Controllable generation with text-to-image diffusion models: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Controllable generation with text-to-image diffusion models: A survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.705414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.705414Z digest=sha256:4f1fe6c5ef916fe8b4f861445d699228a42fc0295914fa9125cdd77a8c70bb99

Observation a28b3082-2557-4e4f-a468-60d249ca8477 · outbound

This paper cites End-to-end object detection with transformers.

Preliminary Explorations with GPT-4o(mni) Native Image Generation End-to-end object detection with transformers

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.709903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.709903Z digest=sha256:e27c75f08d1f959f74a243c9b558a9e778d4891762ece1c8654f0660821f5ff7

Observation fde5357f-4cf0-474c-a43b-7b0ad7bacbd5 · outbound

This paper cites Generative novel view synthesis with 3d-aware diffusion models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generative novel view synthesis with 3d-aware diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.714366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.714366Z digest=sha256:ab42b8664243998484ffc3729654b1d3a149c4ebe1b6a6b7d3c9431e702febf5

Observation adf74f5c-5fb7-462f-915c-658c64b779ec · outbound

This paper cites Spatialvlm: Endowing vision-language models with spatial reasoning capabilities.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Spatialvlm: Endowing vision-language models with spatial reasoning capabilities

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.718644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.718644Z digest=sha256:7f8df0dcf51d71797b0e05f6aa90ba43084b6b5c2c851d1af40e100c24d1b17f

Observation 73c824a5-04fa-4b49-9113-54dc16c64d83 · outbound

This paper cites ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit Adaptation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit Adaptation

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.832209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.722900Z digest=sha256:07493a30a42b8ccdc1675900f9a462dbf3d2ae9f3cf37df8d5d66725009307d3

Observation ad939b62-f772-4283-b493-c7287b43361b · outbound

This paper cites DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.727261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.727261Z digest=sha256:4d48996cb2f167284938c0544bdc22ee1eb782e05652aa6b721ed9d7bbfc22f4

Observation 59912438-66d6-4ff5-bb62-79e632c62f3d · outbound

This paper cites TextDiffuser-2: Unleashing the Power of Language Models for Text Rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation TextDiffuser-2: Unleashing the Power of Language Models for Text Rendering

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.731711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.731711Z digest=sha256:18452a7be9ffd1ea87aa58aecc237daef4b5c8a926f98017960c280a338fd556

Observation 37680e1e-3a6c-4d2b-b740-7fd746b94214 · outbound

This paper cites TextDiffuser: Diffusion Models as Text Painters.

Preliminary Explorations with GPT-4o(mni) Native Image Generation TextDiffuser: Diffusion Models as Text Painters

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.736690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.736690Z digest=sha256:4465e5081e36206b7f0a32fbf04b35fcb4235329b84aa2db2d80799292a36085

Observation d264e7f9-0cd0-4d58-bffa-56c9768eb36f · outbound

This paper cites Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.741344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.741344Z digest=sha256:06a9d872a592967c4ca288b4360c621c285c924a5a37ac2c7129c29cd097d71d

Observation 144adc4e-f03e-4738-885a-6ecade1076d2 · outbound

This paper cites An Empirical Study of GPT-4o Image Generation Capabilities.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.746041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.746041Z digest=sha256:5382ee44a472a463bdd3bd5fccda0aacc404adc8901441b2fab7d508739917d7

Observation ad71bfe2-44f0-49f3-8971-e9aa8831e946 · outbound

This paper cites Manga Generation via Layout-controllable Diffusion.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Manga Generation via Layout-controllable Diffusion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.750697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.750697Z digest=sha256:66ef3384d4397298869fb2f62481c74868f3fb853e79d4c2b8fed623278971f6

Observation 1bf12b18-de75-432b-b3ff-dac921afb197 · outbound

This paper cites All snow removed: Single image desnowing algorithm using hierarchical dual-tree complex wavelet representation and contradict channel loss.

Preliminary Explorations with GPT-4o(mni) Native Image Generation All snow removed: Single image desnowing algorithm using hierarchical dual-tree complex wavelet representation and contradict channel loss

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.755132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.755132Z digest=sha256:5d7d77b3f2d95bca7a2db083577e17db27326e0e10752029d209f6cf233dff5b

Observation 0e070bc9-c686-4a74-990b-ba419fdfd1d0 · outbound

This paper cites DreamIdentity: Improved Editability for Efficient Face-identity Preserved Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DreamIdentity: Improved Editability for Efficient Face-identity Preserved Image Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.759651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.759651Z digest=sha256:28e432dae34098ada0c6ddbfbed583b1f7c1e04913461ae41ac48cfc826295b1

Observation 771d326f-3b91-491c-8196-7d2d99db9355 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.764185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.764185Z digest=sha256:48eb9b8a2f68310052ab9ac8becd683f8e10986115898237e40abfb7b9562f36

Observation c0a62679-c050-4b44-aeb2-565d6e1f5ac8 · outbound

This paper cites Snow mask guided adaptive residual network for image snow removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Snow mask guided adaptive residual network for image snow removal

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.768725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.768725Z digest=sha256:04f88d32ae39505de4c1612e54ef9f03da433b170457bb7b45af3f091fe47d13

Observation 81906c70-4c57-4731-9167-1584c2f88e7c · outbound

This paper cites Masked-attention mask transformer for universal image segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Masked-attention mask transformer for universal image segmentation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.773805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.773805Z digest=sha256:1b142df48c31c771d8460432e8e5aa1e6d56586d8322b3c9da002b9006b3b666

Observation fb144973-a305-4ff1-9e7b-53ed587bd2a9 · outbound

This paper cites Per-pixel classification is not all you need for semantic segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Per-pixel classification is not all you need for semantic segmentation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.778216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.778216Z digest=sha256:2f5aac03e218389d8b0b8f8d538e649f980ebe27743e4f9ca7b9e00e8bab23a5

Observation aa57dda6-cfb8-4972-822e-7452d1676e04 · outbound

This paper cites Object counting and instance segmentation with image-level supervision.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Object counting and instance segmentation with image-level supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.782665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.782665Z digest=sha256:2ca977097c5ca37d347d39e990a504e1c14832a98edbc74f0dec656a5b5b8782

Observation efb880de-d4ed-4f27-a38b-dd511d5a384e · outbound

This paper cites Generating diverse agricultural data for vision-based farming applications.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generating diverse agricultural data for vision-based farming applications

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.787222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.787222Z digest=sha256:ce4fd063b498a8904b7778683894a81d3173a2f453a3b3c4c7ca86674c799eaa

Observation 842a30dc-5498-40c7-a371-30b5d1dfa2e1 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Preliminary Explorations with GPT-4o(mni) Native Image Generation The cityscapes dataset for semantic urban scene understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.791409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.791409Z digest=sha256:58241d6486e4589bb76d55d81c0eabef5459be90d6502ab57d1fe2672dbe7c09

Observation 9dc81256-0f63-4a7c-974e-cb0e54898864 · outbound

This paper cites Latentpaint: Image inpainting in latent space with diffusion models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Latentpaint: Image inpainting in latent space with diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.795343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.795343Z digest=sha256:7a712eca4da63eae2e3fb7903bec46b769d8747013032c62b0f8ea330e63e840

Observation 35d55de4-b92f-443f-b11b-34312494d835 · outbound

This paper cites Deep learning based 2d human pose estimation: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Deep learning based 2d human pose estimation: A survey

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.799583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.799583Z digest=sha256:10fb979001e39899547277f92e83b0ddfaac525bf1713929afef6f7b3208ae6c

Observation c94acfc6-3603-4f26-b205-f1706bac0966 · outbound

This paper cites 3d-aware conditional image synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 3d-aware conditional image synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.803714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.803714Z digest=sha256:57bfc45db9a2c119bf8706380a569ef775bc53ee55522cbfdd58d3fa0775aada

Observation f1049d54-fde6-4363-9609-7d23088a7cec · outbound

This paper cites Towards intelligent design: A self-driven framework for collocated clothing synthesis leveraging fashion styles and textures.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Towards intelligent design: A self-driven framework for collocated clothing synthesis leveraging fashion styles and textures

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.808389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.808389Z digest=sha256:9c2fba0fe94d2c795573e7b3b69582da10ba715e0f0b4fee456c397331e47256

Observation b9e021c8-a6b4-42b7-9377-b0e7818556dd · outbound

This paper cites DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.812799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.812799Z digest=sha256:d178ad531b2799ae196f2a1860228eaef336822032c2ab5ced9bb066ed1716e5

Observation cc8644ff-a1cf-45ff-b6d1-c9264d3d3e35 · outbound

This paper cites Discovering novel biological traits from images using phylogeny-guided neural networks.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Discovering novel biological traits from images using phylogeny-guided neural networks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.817421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.817421Z digest=sha256:7f410b3d85d429b521cc72a554befa224bd2896551bc7237cd811f05a30f9693

Observation 3b902b97-9c4c-4c7b-9c9c-1dcb46892af8 · outbound

This paper cites Everingham, L.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Everingham, L

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.821802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.821802Z digest=sha256:27a64a4210a39677f6edddcd439819cf1e1e663c2739a347e0d98f99219eed0e

Observation 795e3154-f504-4469-86ce-559b3197ed52 · outbound

This paper cites Everingham, L.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Everingham, L

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.825899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.825899Z digest=sha256:358acd68f88ef2d44e3cdf49923dfa68920098ae95a0ad5c609551da7d6d37a1

Observation cfea3f1b-a514-4cd3-814a-cfaa7b5a2dd8 · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.830158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.830158Z digest=sha256:91f1b89fb8684a0b7c7e889fe01bae587fb70baacec22310da79828e14614e58

Observation 3579c4b0-bea9-41da-a16e-4b22a62178bc · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.834485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.834485Z digest=sha256:870a3e402c8d282d005e96b78a9e19274b8033cff85fa7d9505bf159a61dd562

Observation 7d6494be-079c-44b0-af90-c7f83bcdd069 · outbound

This paper cites CascadedGaze: Efficiency in Global Context Extraction for Image Restoration.

Preliminary Explorations with GPT-4o(mni) Native Image Generation CascadedGaze: Efficiency in Global Context Extraction for Image Restoration

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.838880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.838880Z digest=sha256:d5302b0228ea192fb12c9f244dc1fb923e87b008df03ce442b4fde3ee73805fc

Observation 038a5edf-4984-4878-aae7-9146e0f93afe · outbound

This paper cites Fast r-cnn.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Fast r-cnn

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.843410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.843410Z digest=sha256:823c231f636be55b91689516a96bbed62c64a1c6fb268ae2ba8c2dfbab09ab4c

Observation 44601787-8338-494b-bfa9-f15a2d3eff96 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Rich feature hierarchies for accurate object detection and semantic segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.847545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.847545Z digest=sha256:eaa01b30f8f528ab44431df8a89c4ce24f0fc17942740f111911ddb3620ea075

Observation f1bfd7b5-69d9-495f-abe3-3c61337ed373 · outbound

This paper cites Talecrafter: Interactive story visualization with multiple characters, 2023.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Talecrafter: Interactive story visualization with multiple characters, 2023

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.851631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.851631Z digest=sha256:47e80feced086183562a745b94c80817d318c07dc27c6eebe07b6110f4477386

Observation 27b339b5-6703-467f-8f51-62bbc13b7210 · outbound

This paper cites DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.855641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.855641Z digest=sha256:72c08d001704de06d50b89dca975fda283bf596ffc64aaa9108cad56514b16d6

Observation 719a037f-9686-4e88-a9f5-1b7e6a4fe024 · outbound

This paper cites Modulating Pretrained Diffusion Models for Multimodal Image Synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Modulating Pretrained Diffusion Models for Multimodal Image Synthesis

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.628774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.860484Z digest=sha256:fb9cf26f9b6740cfe2e815db341f0a35c21bd8de9debd88d0b1dd2b29558ca68

Observation da1b9c2a-45b8-49c6-920b-b19e59bd6b0c · outbound

This paper cites Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.864916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.864916Z digest=sha256:452c9030196dfa28a1a99974cda0cc2b8dadd7465238bdf1217172bbe2fe51da

Observation 530f8f8b-2a42-4e3a-af28-5c699994804b · outbound

This paper cites an unresolved cited work.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.869304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.869304Z digest=sha256:27aad224fe5e4f5107591c6643145279b5bfe56b75290e056ad2c4a328cb4e8d

Observation c14053b7-d2c7-4b55-8457-9cce2b8d67ad · outbound

This paper cites an unresolved cited work.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.873445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.873445Z digest=sha256:9fcf610918da42a391798e0b70950e524ac1d6b7c2b9cfd57e06c8c9b72e4c25

Observation 1785a553-6d7c-464b-85fc-eb57d11e5020 · outbound

This paper cites Dreamstory: Open-domain story visualization by llm-guided multi-subject consistent diffusion, 2025.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dreamstory: Open-domain story visualization by llm-guided multi-subject consistent diffusion, 2025

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.877554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.877554Z digest=sha256:c37e3c6faf7b0ca1fd482a205ad0f3e7b9fc71f400f4f515957530f1cd7fa37c

Observation 37d306bd-4176-4657-add1-b92b5857256c · outbound

This paper cites Styleposegan: Pose-consistent virtual try-on via pose-guided style transfer.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Styleposegan: Pose-consistent virtual try-on via pose-guided style transfer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.881449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.881449Z digest=sha256:f92b14371419d1f95435102732777788de48e163e67c7138e60b249416c6a793

Observation 27fc4bf7-d9d2-4bfb-b952-2f079060611c · outbound

This paper cites Dresscode: Au- toregressively sewing and generating garments from text guidance.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dresscode: Au- toregressively sewing and generating garments from text guidance

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.885537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.885537Z digest=sha256:fbe09afe87500074427844c6bb3f49b27cabd1e3a13745dcbe259a5c06a98189

Observation 3b483f99-2dc4-4887-bb4f-0d86f3b9a311 · outbound

This paper cites SynthSet: Generative Diffusion Model for Semantic Segmentation in Precision Agriculture.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SynthSet: Generative Diffusion Model for Semantic Segmentation in Precision Agriculture

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.889670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.889670Z digest=sha256:20597ef93c548b194c878207fa001aabd025b981b4781dc7d2238c147f3535e6

Observation c3b79798-25d7-44b9-b6a9-eb4f08b76709 · outbound

This paper cites Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.893977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.893977Z digest=sha256:ee58d1b067ab9e69113fe3933377b6398a73b038c37aefd435648b635a5a72bf

Observation 9cb529da-72ca-48b1-8cae-28ce021f1691 · outbound

This paper cites Composer: Creative and Controllable Image Synthesis with Composable Conditions.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Composer: Creative and Controllable Image Synthesis with Composable Conditions

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.898202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.898202Z digest=sha256:c38421659ebd04183df91ca74718c636950cd47032bb17d5bc8cb2d304191923

Observation f996091a-1cd1-40ce-96fb-96f203d1ad48 · outbound

This paper cites Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.555542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.902399Z digest=sha256:68c19dbef73a6260773ec5e16dcd8057916840fe0facf4f4ea88427a24096916

Observation 0659121a-ae89-4980-a33c-7b0621a351bb · outbound

This paper cites AutoGeo: Automating Geometric Image Dataset Creation for Enhanced Geometry Understanding.

Preliminary Explorations with GPT-4o(mni) Native Image Generation AutoGeo: Automating Geometric Image Dataset Creation for Enhanced Geometry Understanding

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.906527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.906527Z digest=sha256:59717d9ca78a4c9caafb34f3c086945e4f0f8ad2a40f409b3bca8a7db77f962b

Observation da0625f8-2d9a-43a4-bd93-3b8c12fade59 · outbound

This paper cites ReVersion: Diffusion-Based Relation Inversion from Images.

Preliminary Explorations with GPT-4o(mni) Native Image Generation ReVersion: Diffusion-Based Relation Inversion from Images

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.910917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.910917Z digest=sha256:514541a5d2e2ab6fd2912cf5995e78afb2a93821f44e82ef03840a1808d39389

Observation 8f837654-7f41-417b-a035-eea6134b0574 · outbound

This paper cites Liteflownet: A lightweight convolutional neural network for optical flow estimation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Liteflownet: A lightweight convolutional neural network for optical flow estimation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.916052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.916052Z digest=sha256:13212e568528b825222b83c3a541aa021ec6082fb9688e86de6654a01c7f0caf

Observation f9e57c1f-b25e-47a8-be60-48d7052a8e4c · outbound

This paper cites Vision transformer in industrial visual inspection.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Vision transformer in industrial visual inspection

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.920065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.920065Z digest=sha256:c85d2dad2577b7a1186791afa41737de254752b635a8e047c8557f3d874e6db8

Observation 63a02054-893f-420e-a2da-d9a46e4bf82a · outbound

This paper cites Flownet 2.0: Evolution of optical flow estimation with deep networks.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Flownet 2.0: Evolution of optical flow estimation with deep networks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.924276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.924276Z digest=sha256:357bcb1ff9eec00c442274745ccda7807fcd92af2efffed0366d6bb448247344

Observation 5c9a6014-9305-4151-b7aa-9856591d1143 · outbound

This paper cites Desnowgan: An efficient single image snow removal framework using cross-resolution lateral connection and gans.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Desnowgan: An efficient single image snow removal framework using cross-resolution lateral connection and gans

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.928240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.928240Z digest=sha256:b7546369f0e3072617ea383ebcaa6b21d28264202156fee27a4bd4599ffa8816

Observation bf5fb43c-4d2b-4f47-9e21-a10750de7164 · outbound

This paper cites Remote sensing change detection in urban environments.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Remote sensing change detection in urban environments

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.932328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.932328Z digest=sha256:0f5fa2029b65054f6cd86f10ad0a887e695fa293c9deea6182c40e1ea8cae538

Observation cb8dcaf4-c535-472d-b974-0afacab16b0b · outbound

This paper cites Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.936502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.936502Z digest=sha256:7eab132615e82b80acb7c6203af795d3d7a16853effbd4f49c71b1e5e1215cfb

Observation 66f47646-64f3-4701-8919-89ac17e4823f · outbound

This paper cites SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.494729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.940679Z digest=sha256:6f67f9d5a3ca44dd2a97ad4de9fe38d24454c2c9d2a779d34644d9ee22691bf8

Observation c98d2440-8761-4f94-aacc-064690ab2357 · outbound

This paper cites Lumen: Unleashing Versatile Vision-Centric Capabilities of Large Multimodal Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Lumen: Unleashing Versatile Vision-Centric Capabilities of Large Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.945352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.945352Z digest=sha256:fc8ea4c90d6d3fb20e8d99e8b78f9dfee192104106cbbebea51558f605937d6b

Observation 7a29afa5-3cec-408f-99d2-d47299c4a2f5 · outbound

This paper cites Beyond Aesthetics: Cultural Competence in Text-to-Image Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Beyond Aesthetics: Cultural Competence in Text-to-Image Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.949916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.949916Z digest=sha256:ad6a4c7bc5071c196d8bc82759a6d8a6e07ac4cdaca043670b91a16cb97653bd

Observation 557623ef-19d9-4a24-aaf0-6202808f6732 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 3d gaussian splatting for real-time radiance field rendering

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.954334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.954334Z digest=sha256:ac94ec14306473317a33123521a56790be21a00c35e3c21432e6a1742c9c885e

Observation 6a5b976d-9aed-453c-b850-aa770c613df1 · outbound

This paper cites DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.958763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.958763Z digest=sha256:6fab4610a506ff8dd66f898f57c7f124e9d15f79648470f71bf5a30ca96a353a

Observation e722c22a-6969-44e7-8556-d2540879507f · outbound

This paper cites Probabilistic modeling for human mesh recovery.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Probabilistic modeling for human mesh recovery

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.963353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.963353Z digest=sha256:2631d85020820b0690bc83a8bfc52ec3f0f14dc5301bd5800a884c33bcaa6b29

Observation 2a3ae3de-47d3-4b5e-850a-3139a0907cae · outbound

This paper cites Raindrop-removal image translation using target-mask network with attention module.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Raindrop-removal image translation using target-mask network with attention module

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.967398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.967398Z digest=sha256:2f9d6907e3115d37feb9d666af1a20aafc36caeb5b740f582f6506565d8b8eba

Observation bc586949-05d3-4f58-a8c3-67e684898143 · outbound

This paper cites Dicti: Diffusion-based clothing designer via text-guided input.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dicti: Diffusion-based clothing designer via text-guided input

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.971373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.971373Z digest=sha256:93ea36ce7062f5f1dbf596ae37f4db72ad995326a913749eee63a2320a70b72b

Observation 5d7bd6e6-486c-4f55-ae82-b522f0b294ec · outbound

This paper cites Physics-based shadow image decomposition for shadow removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Physics-based shadow image decomposition for shadow removal

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.975366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.975366Z digest=sha256:94ec9e5046e2610a6acf40e56e38bcc8fa12bae7ba4f756a9e67a5fe1fd59b50

Observation 60341660-f287-4588-858c-a35a363177c4 · outbound

This paper cites From Text to Visuals: Using LLMs to Generate Math Diagrams with Vector Graphics.

Preliminary Explorations with GPT-4o(mni) Native Image Generation From Text to Visuals: Using LLMs to Generate Math Diagrams with Vector Graphics

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.434842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.979690Z digest=sha256:ed8e594af09ce033ca6488e8bf41caf343937a4537f6f048319603e762fa3f8e

Observation 9d49abaa-04a5-4f99-adba-a64617eb183d · outbound

This paper cites An underwater image enhancement benchmark dataset and beyond.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An underwater image enhancement benchmark dataset and beyond

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.984144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.984144Z digest=sha256:2f15a1598df6875a93f528e1ac62406e51c042748c42fbe44fefe320ce77be55

Observation 6c8fa3de-94c1-46cb-ae53-a46c538b7287 · outbound

This paper cites Real-world deep local motion deblurring.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Real-world deep local motion deblurring

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.988400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.988400Z digest=sha256:bcad00545a50219eb57d2894db34987f5904f98a86822961add9672a75179a33

Observation f4738e64-9a59-44a8-b5a6-8aa89ae0307a · outbound

This paper cites Cheffusion: Multimodal foundation model integrating recipe and food image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Cheffusion: Multimodal foundation model integrating recipe and food image generation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.992330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.992330Z digest=sha256:b29618a65e98af0453cab7ba952b0be006d05c7169f2df8224d5f89190c23310

Observation 3c9fd2a9-a618-4930-9f8a-7e0113366c74 · outbound

This paper cites Image content generation with causal reasoning.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Image content generation with causal reasoning

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.996386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.996386Z digest=sha256:79783278bea1eb970ec917fbdf04aa7c1380cd457ee0706c3592dfdd63873820

Observation e0b3a3b6-8b60-4731-979f-4eb3fca83de0 · outbound

This paper cites Exploiting reflection change for automatic reflection removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Exploiting reflection change for automatic reflection removal

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.000457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.000457Z digest=sha256:c483e9e19581c5344ea6572aa19ea522cd578df53288b9ff10de698dc4b153e7

Observation ff29fc25-f131-4327-8a2c-6f0ede33dd9d · outbound

This paper cites Generate Anything Anywhere in Any Scene.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generate Anything Anywhere in Any Scene

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.004601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.004601Z digest=sha256:79927a193b95997611ce594ad9478657a55f0553c78ac219d7b36b22862258d6

Observation 309c2a77-bd40-462b-9d10-8de4fcfb5d38 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Gligen: Open-set grounded text-to-image generation

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.009540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.009540Z digest=sha256:ac6c355c1104cbc61fe63cbf5a69c9217a0de69f30a9f4b22d5fdeaa3e61229d

Observation 17d7d127-b48e-41a2-818d-979823c3d95e · outbound

This paper cites Swinir: Image restoration using swin transformer.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Swinir: Image restoration using swin transformer

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.013949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.013949Z digest=sha256:11c81b2c2fb90e274a192cc7de6a4f4c94bdb3f4ea6cd2cf2debfbc462d7a91d

Observation 3dcaaf6c-f497-4a4b-b579-c84e6b9b14f9 · outbound

This paper cites Object Counting: You Only Need to Look at One.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Object Counting: You Only Need to Look at One

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.017973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.017973Z digest=sha256:3332e7c449fd59a55ec6b156f5bd9e7826c4be56545ca4c9f5d3a80c8f420bf7

Observation ae60a7ab-10be-4e17-ae25-0d380b23dc62 · outbound

This paper cites Phys4dgen: A physics- driven framework for controllable and efficient 4d content generation from a single image.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Phys4dgen: A physics- driven framework for controllable and efficient 4d content generation from a single image

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.022186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.022186Z digest=sha256:c9ba7cb0227e28dda72bfcd677ef345a93b60c2831c4c0d6980f1027ca4baaac

Observation be5619c1-a47e-42dd-b6c4-f69a150adc2a · outbound

This paper cites Microsoft coco: Common objects in context.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Microsoft coco: Common objects in context

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.026366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.026366Z digest=sha256:e302214feab5f0ba3633b265caf490a73a9c4c060b109ff50efb10144b22a7b7

Observation 630eb4f5-70be-4fbf-9da6-1b9f6b190442 · outbound

This paper cites On the cultural gap in text-to-image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation On the cultural gap in text-to-image generation

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.030724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.030724Z digest=sha256:169f2a5e1546347f7e9336cd6506c01a6cfb308ca3e4f851fb16942bce2f1aa7

Observation dfce535c-f1ea-494f-9e9e-cffa8163f5d0 · outbound

This paper cites Generative Physical AI in Vision: A Survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generative Physical AI in Vision: A Survey

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.035004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.035004Z digest=sha256:b7d5352fab25ef72c849b1186576afb51a58b4ca7df23d2caaf9f2e2683c75e2

Observation e2282a71-9935-46f7-a685-e1547c4a58b8 · outbound

This paper cites StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter.

Preliminary Explorations with GPT-4o(mni) Native Image Generation StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.039629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.039629Z digest=sha256:780db9492c1d9143b44314931304c822a016e29d2ecb3a575402c0b4a737c7b6

Observation b3e7ee1f-17d1-4baf-a65f-e0c854a84b01 · outbound

This paper cites Structure matters: Tackling the semantic discrepancy in diffusion models for image inpainting.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Structure matters: Tackling the semantic discrepancy in diffusion models for image inpainting

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.044224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.044224Z digest=sha256:e9c5b04200c71b5b9bdb1313014c31ea095a76a9f24f82e91d51233d3bd9da10

Observation 54215261-fb6c-42ec-a503-37627b0ccd0b · outbound

This paper cites One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.

Preliminary Explorations with GPT-4o(mni) Native Image Generation One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.048723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.048723Z digest=sha256:e58ed03356877de4dc4d4e0d0e7e2b7d0396daceb18cad6688bbd14b30ecb488

Observation c419dc49-cd53-4dba-b0cd-23f047d9b483 · outbound

This paper cites Git-mol: A multi-modal large language model for molecular science with graph, image, and text.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Git-mol: A multi-modal large language model for molecular science with graph, image, and text

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.052810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.052810Z digest=sha256:78e68c2b82101fb6f3ad0bd2d2de62235484ae7e541349ff5253cb1508df2ae1

Observation be9f98fa-9853-4479-ba25-35e94013e4ce · outbound

This paper cites Character-Aware Models Improve Visual Text Rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Character-Aware Models Improve Visual Text Rendering

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.056819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.056819Z digest=sha256:a6ee62c7128d4a41b7f638b1e1b9f979b778ee8f4f963961a4b881e1d8872c1e

Observation 6a51503c-58dc-48ce-a7c9-8abaa2446b08 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Zero-1-to-3: Zero-shot one image to 3d object

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.061601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.061601Z digest=sha256:30dbc77777613c795c636bd6d81f0ee26aa9aa4074ad335609b62c06838cb340

Observation eb0c1020-6dcf-4345-bf62-bd5e33e9411e · outbound

This paper cites Physgen: Rigid-body physics-grounded image-to-video generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Physgen: Rigid-body physics-grounded image-to-video generation

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.065904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.065904Z digest=sha256:4f87faab4244b54e628f9890845e1c64cf6928f8655acdf34343259bf30baefc

Observation b08867eb-ef58-45ba-a0d0-4e02857f91d1 · outbound

This paper cites Ssd: Single shot multibox detector.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ssd: Single shot multibox detector

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.070172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.070172Z digest=sha256:48ccfa20427535dc8a8762bd090c4819a30496a197768bf054fc44666237e5c5

Pith citing papers

Observation f09b4bd0-cc60-4329-b3ee-9ccb02314bf8 · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:10:59.092148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T17:46:59.581486Z digest=sha256:858768c3705d1ef7ad3d380574cb03c15893fa7b1ed1910e143d954ede6d65c7

Observation 2727aeba-fb0f-4b09-af38-c50b6465e54c · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:34:57.141773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T09:34:16.292323Z digest=sha256:eff663f7247b47f1e7c29300e750c8f19bc62eb9d21856119d7cc7e1699ab3d5

Observation 92c38713-8390-44c4-b780-4939cd4b66a9 · inbound

PosterHarness: Turning Scientific Poster Generation into an Auditable Instruction-Following Benchmark cites this paper.

PosterHarness: Turning Scientific Poster Generation into an Auditable Instruction-Following Benchmark Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T05:30:32.163273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:30:32.163273Z digest=sha256:3898864d373bb1ec4ef67bf550fd48a7091c70dcb43e7be5fd9c35e5617c83eb