Pith. sign in

Paper Citation Record · LEDGER

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

As of 15 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 2 inbound Pith citation observations for arXiv:2412.05722.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05722 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:28:49.296167Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:10:34.521533Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T13:10:35.138865Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 151650cd-83d3-4abf-96eb-57aed0039964 · outbound

This paper cites Image quality metrics: Psnr vs.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Image quality metrics: Psnr vs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.761721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.150491Z digest=sha256:373f542ba3ebb7eb3522313c64689a6c0584df089dd16122f73cbe69223ec4cc

Observation 16a584f9-0b6a-46c1-bb53-3290b3b91e63 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Learning transferable visual models from natural language supervision

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.745500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.155674Z digest=sha256:e1afc49b64ef7ab722f8a34e2615506b0e6be7c1e10ab715b40256fe75a0436a

Observation 21723be7-aca4-4330-b655-629700778397 · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.729697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.160468Z digest=sha256:9338d46442cb6c913015126fc6f2b7c940c655fe5a3112c77f042ec5540527b7

Observation 61d19daf-2581-4681-a84a-3a632c9f8207 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.165301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.165301Z digest=sha256:f4dedf678c0009fba806cb891be478507ffbf2f7d0045ad9540015e9672984fd

Observation 7f9a7d9d-96ab-4de3-aaec-cd9e8c6a90f8 · outbound

This paper cites TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.170451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.170451Z digest=sha256:3fea1931d610494bb6bd9c398dc0f129d3623d8b8453675be51915b95be3d69e

Observation 511d82ce-1bc4-442b-b7ca-fa927803ab0f · outbound

This paper cites Holistic Evaluation of Text-To-Image Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Holistic Evaluation of Text-To-Image Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.176106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.176106Z digest=sha256:77e6e2957905cc8836ef0a3cefd8bd0a6e44e2a4b68930ce68e71f53bdfafe82

Observation 05c34299-15c2-4996-928b-9db49254ef86 · outbound

This paper cites Attribute2image: Conditional image generation from visual attributes, 2016.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Attribute2image: Conditional image generation from visual attributes, 2016

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.182060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.182060Z digest=sha256:d7678eb39dbaa223e9e9a117a4ae2f1e68958053726f338710f59d53e9ad5bd5

Observation d0f756cc-015b-4b78-afef-4fe2baf8090a · outbound

This paper cites Neural discrete representation learning, 2018.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Neural discrete representation learning, 2018

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.186958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.186958Z digest=sha256:39c8c4eecdb8eaa3bd2b508763f04b190d7dec1e3db150db3932ab286da13b9e

Observation 7b9d38fd-ea08-40a7-b16d-d866dc62e916 · outbound

This paper cites Generating diverse high-fidelity images with vq-vae-2.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Generating diverse high-fidelity images with vq-vae-2

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.191869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.191869Z digest=sha256:9dbb56e5b8d2df8da2781fb09586c49b13fd29f094c25c73e85d3f932b6d6f0a

Observation b53a17eb-e00d-4b7e-9bf9-5e76e35b8cad · outbound

This paper cites Zero-shot text-to-image generation, 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Zero-shot text-to-image generation, 2021

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.196844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.196844Z digest=sha256:56aebad753c5c34e85621e4a272e2b2d9216bf65d1c4bc05bac0b11c6157b5c8

Observation a06f047c-90ff-4521-ab8d-e113cc4dc6a1 · outbound

This paper cites Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.201609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.201609Z digest=sha256:22fe7a72b0d76fa4f0ea064e7b23ecaf84a21d08be7dbb8458de83df823a7c87

Observation 101a44eb-8dbe-41ae-9a2a-3aab9813f416 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.206582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.206582Z digest=sha256:c326f773072aff08eecc78e9af2d52f5300bd39c96dc6f6380facea465722bd3

Observation aa35b2b2-f0dc-4077-8006-315d50566203 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.211973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.211973Z digest=sha256:33c4b6b1c173ab69f5dd4bdcbc362dc1934336dd4ad149e1754eaa2a265ef042

Observation 3edfb075-b63a-4127-b96d-f3c069ba6168 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Denoising Diffusion Probabilistic Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.217561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.217561Z digest=sha256:841607ca7cbca9ac9b5f8644c596fdffea3f3864658c98ddeea31b9dfa8189f1

Observation 27181258-3808-469a-9805-ae6b09b43a77 · outbound

This paper cites Openclip, July 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Openclip, July 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.223090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.223090Z digest=sha256:f969da1edda8b6103a0c21d64f7dd9d5a750063aa611bfea4cd1279690219f00

Observation 2b495b58-f14c-4303-af11-5ba88ff02f66 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.227867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.227867Z digest=sha256:955fe7f1d0a7d8895b048f3c14428421f11f3f381d80239abeee3dd332d1181c

Observation ab0e4879-c23b-4f78-acc7-04c988262364 · outbound

This paper cites Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.232721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.232721Z digest=sha256:c4ef2c8e1b693dc73001fd3958a15ab890227cec4a03472070d2ebfeb36a29f0

Observation 85ca351c-4b89-4525-95a5-0abd231b05de · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.238323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.238323Z digest=sha256:f23c2d746fcca75c77c4308d329116e30fc9802afcdd3ce4d0eed7ba57883b63

Observation 64b3ca76-c1eb-4e71-a13a-5292fc5e873d · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.243359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.243359Z digest=sha256:a157f488886e297ae504d5f74e6e2e8230603a278fafe9adf8a0337b2f674104

Observation 2a2f85bc-529d-4b91-825f-f2932054bee8 · outbound

This paper cites Effectively unbiased fid and inception score and where to find them.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Effectively unbiased fid and inception score and where to find them

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.631619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.248470Z digest=sha256:0f85a7bfeef07b21d267be2a30085447bcddd8d4077d7a5a433c7f70a39da3cd

Observation 76a53b11-a4ed-4a60-8dc9-630a6b048a33 · outbound

This paper cites A Note on the Inception Score.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent A Note on the Inception Score

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.253323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.253323Z digest=sha256:89940b725e772aaa30a2ac793c23ed3218011daa86b60cc52ebff51509576089

Observation 2a8d3119-2974-495f-b6c9-558251bf0009 · outbound

This paper cites Improved Precision and Recall Metric for Assessing Generative Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Improved Precision and Recall Metric for Assessing Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.258616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.258616Z digest=sha256:61b418e1934384da66c46d87d31d976f8c5fc1061767ddff659a97d465bcb01a

Observation f09ccf2e-cc4b-4385-add2-64cd45acaf48 · outbound

This paper cites Benchmark for compositional text-to-image synthesis.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Benchmark for compositional text-to-image synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.615268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.263655Z digest=sha256:e410d244f4f1f09dcaf70f327f9ab5b2c9d1abe6c7b6b12b67e4aea7dbbd954a

Observation 89e3bb8b-afe0-435b-b10d-7d7528d42cf9 · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.268128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.268128Z digest=sha256:4c4d53e58ccba36b999d212c20b016adac0ab7dd88d6f02c14045e98e2464f78

Observation 8c014d5a-66f2-45f8-8f09-8f782e8a36ba · outbound

This paper cites LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:28:49.360323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.272738Z digest=sha256:f4b34d1616a94c84ed0c59db56ba8eeca507dbd7d30eec011b57e61c97252c25

Observation bb5e3565-f2b6-4a75-818f-6e543cf9006d · outbound

This paper cites DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.277601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.277601Z digest=sha256:8756fcc1be4a7f95d0cb9bed6ab29f72fe7115864c8ad7003893d4bb833d924f

Observation 150ae64c-8f65-4f22-a4ba-05c0a27ccc91 · outbound

This paper cites Grounded-Segment-Anything, April 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Grounded-Segment-Anything, April 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.599989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.282635Z digest=sha256:3922f16ca64122d47845eecfe1f430bbae9d1798a612805eb4c224e92f606ee7

Observation e7611f13-3427-4ffc-95e9-1f934b951d0a · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.584119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.287077Z digest=sha256:7fa365220881281e2439c6c1236ada8f0682607cd926cdaea24e0fd0630a519b

Observation e2db01af-d7e7-4e13-ab4f-7da514d8336c · outbound

This paper cites spaCy: Industrial- strength Natural Language Processing in Python.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent spaCy: Industrial- strength Natural Language Processing in Python

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.291456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.291456Z digest=sha256:69b2bff40e5ca1d4e5557820d6f4fd81b42a604faee4dacb1501deb9aec918b3

Observation e976bf36-8cf8-4430-a2c0-099f8301380e · outbound

This paper cites LangChain, October 2022.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LangChain, October 2022

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.559163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.296167Z digest=sha256:5256ea8b1e4370aa8fb4ccb1da83aaac360434f9914f90aa7f6efc98c49421da

Pith citing papers

Observation 21829e93-2341-4698-bfb8-1df7c5a62bdf · inbound

Mitigating Diffusion Model Hallucinations with Dynamic Guidance cites this paper.

Mitigating Diffusion Model Hallucinations with Dynamic Guidance Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T11:24:16.280713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:24:16.280713Z digest=sha256:67a066ff39eac5e0aa682d6909d5b1fc0754200e5b721ac7870bd3b2780956ec

Observation 1c1d5067-e92d-446a-b227-13c7d55f2d7b · inbound

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models cites this paper.

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:10:35.145783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T13:10:34.521533Z digest=sha256:32d24b85b021473e6e362b20591c2ef186e7a54f8d944cab3d6a4fcd8c1c6e6c