Pith. sign in

Paper Citation Record · LEDGER

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

As of 18 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 2 inbound Pith citation observations for arXiv:2412.05722.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05722 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:28:49.296167Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:10:34.521533Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T13:10:35.138865Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 151650cd-83d3-4abf-96eb-57aed0039964 · outbound

This paper cites Image quality metrics: Psnr vs.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Image quality metrics: Psnr vs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.761721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.150491Z digest=sha256:790a5bb15bd6aa84b6211b653d48d988c3dc49569b2b26fea867aa5f58e36286

Observation 16a584f9-0b6a-46c1-bb53-3290b3b91e63 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Learning transferable visual models from natural language supervision

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.745500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.155674Z digest=sha256:361113de66c20b34fcc6980122632cc3ec8e5bc46ff2c7510be76ae68a377d9b

Observation 21723be7-aca4-4330-b655-629700778397 · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.729697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.160468Z digest=sha256:1ab9c5bb1bdd52cd8cda591822d922ff338636c0292954de54cacb84abf6c05b

Observation 61d19daf-2581-4681-a84a-3a632c9f8207 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.165301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.165301Z digest=sha256:3460539851e8b876946b91997a0046036c9656ca720bc9534f25ae0c3b09e09e

Observation 7f9a7d9d-96ab-4de3-aaec-cd9e8c6a90f8 · outbound

This paper cites TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.170451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.170451Z digest=sha256:5cefd2128c5c563d343a13558b8e1340a119e62f22bc4856a74b85834bb1bc22

Observation 511d82ce-1bc4-442b-b7ca-fa927803ab0f · outbound

This paper cites Holistic Evaluation of Text-To-Image Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Holistic Evaluation of Text-To-Image Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.176106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.176106Z digest=sha256:4aaac806679cfcc9a6aae726e4c1da1ab2fcb11dd7751e1608be44306d93a461

Observation 05c34299-15c2-4996-928b-9db49254ef86 · outbound

This paper cites Attribute2image: Conditional image generation from visual attributes, 2016.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Attribute2image: Conditional image generation from visual attributes, 2016

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.182060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.182060Z digest=sha256:eda11aa4b98a0fe7db64f845e1dd08b00ad359ca3712bf405d9a9eb8dd5ab2c5

Observation d0f756cc-015b-4b78-afef-4fe2baf8090a · outbound

This paper cites Neural discrete representation learning, 2018.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Neural discrete representation learning, 2018

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.186958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.186958Z digest=sha256:914a1c366647d0e57d48374b11f281b89570700dc684b4e000ff5cc08ae6e6b8

Observation 7b9d38fd-ea08-40a7-b16d-d866dc62e916 · outbound

This paper cites Generating diverse high-fidelity images with vq-vae-2.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Generating diverse high-fidelity images with vq-vae-2

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.191869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.191869Z digest=sha256:713dd9b04ab4a3672085ffac755982a17896060ccb846e344e88005e30ce1bdc

Observation b53a17eb-e00d-4b7e-9bf9-5e76e35b8cad · outbound

This paper cites Zero-shot text-to-image generation, 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Zero-shot text-to-image generation, 2021

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.196844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.196844Z digest=sha256:7a8d6a9eeff57c5c0723c501779a264f7bd2d8602902bbf5caaa97d02ef3c6a3

Observation a06f047c-90ff-4521-ab8d-e113cc4dc6a1 · outbound

This paper cites Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.201609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.201609Z digest=sha256:97aa52f00873691c5d8765ee38694c3bec789605e3dcb5b89197e9d113b660fa

Observation 101a44eb-8dbe-41ae-9a2a-3aab9813f416 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.206582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.206582Z digest=sha256:9e72938d9dc9c82e0fedddf879f75ced3e548979340c13ee10d1297dd2deb89e

Observation aa35b2b2-f0dc-4077-8006-315d50566203 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.211973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.211973Z digest=sha256:05adfcb76a202ef9b7523a66b8bc94166b769d201481233a8130caa36f73514d

Observation 3edfb075-b63a-4127-b96d-f3c069ba6168 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Denoising Diffusion Probabilistic Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.217561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.217561Z digest=sha256:f533c0d551fdcba41cef0d84266bf172037b42baacd189cd5f1a3a38d4eb2595

Observation 27181258-3808-469a-9805-ae6b09b43a77 · outbound

This paper cites Openclip, July 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Openclip, July 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.223090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.223090Z digest=sha256:f2c5f4cb72891d012865ff54c6dedd8aeddd3b0e6c87734c5e20d3d263a93bbe

Observation 2b495b58-f14c-4303-af11-5ba88ff02f66 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.227867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.227867Z digest=sha256:2da171fcd0477f1b55974af26515d3f664e1b0166fa5ba2b4380e0ed4365c1f8

Observation ab0e4879-c23b-4f78-acc7-04c988262364 · outbound

This paper cites Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.232721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.232721Z digest=sha256:4d6f31f1fd293bd1e8cd7796b4d76bbc444cd5148c3d8fea0b24aa01f12bc251

Observation 85ca351c-4b89-4525-95a5-0abd231b05de · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.238323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.238323Z digest=sha256:1f8cb6948a51e76fc4d09b42d71bb1b508917917db2a43cf2ef401caafc47b04

Observation 64b3ca76-c1eb-4e71-a13a-5292fc5e873d · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.243359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.243359Z digest=sha256:7f4b9a75eaa7be5fcf99fbe5043b4ac2c9f095f75f032a6d491e936c772b3cb0

Observation 2a2f85bc-529d-4b91-825f-f2932054bee8 · outbound

This paper cites Effectively unbiased fid and inception score and where to find them.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Effectively unbiased fid and inception score and where to find them

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.631619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.248470Z digest=sha256:fd2451ee91ddb5d69ce7dbfe59036796e673f76cc724efbf7df5228ce6d27c16

Observation 76a53b11-a4ed-4a60-8dc9-630a6b048a33 · outbound

This paper cites A Note on the Inception Score.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent A Note on the Inception Score

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.253323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.253323Z digest=sha256:0d4386f7f34f6e3d4a9af6cf2301bc0aee57f8da4f0545cc8e0d51545e36ce34

Observation 2a8d3119-2974-495f-b6c9-558251bf0009 · outbound

This paper cites Improved Precision and Recall Metric for Assessing Generative Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Improved Precision and Recall Metric for Assessing Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.258616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.258616Z digest=sha256:448e3cada9c3d344781509601bc132f7ca1fc4a63298f627b4e78262406235e2

Observation f09ccf2e-cc4b-4385-add2-64cd45acaf48 · outbound

This paper cites Benchmark for compositional text-to-image synthesis.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Benchmark for compositional text-to-image synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.615268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.263655Z digest=sha256:cb45981559284d0fd2fe02eade512e5d043dc1f56efa486aa903b0e34c0ea526

Observation 89e3bb8b-afe0-435b-b10d-7d7528d42cf9 · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.268128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.268128Z digest=sha256:58e44cb6c4407795d3f4d539bd1c6d3edcb916344b6dcf5e7f649c25b961b84c

Observation 8c014d5a-66f2-45f8-8f09-8f782e8a36ba · outbound

This paper cites LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:28:49.360323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.272738Z digest=sha256:d6f8f258adc3cf76439db7ecd7987f18eebf484875a29265c83e1478a57b7f8f

Observation bb5e3565-f2b6-4a75-818f-6e543cf9006d · outbound

This paper cites DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.277601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.277601Z digest=sha256:f28a45a80b66bcc8dfe57a66d3b7e07042dc2c6f4cd9811057d82dbbfd4a427a

Observation 150ae64c-8f65-4f22-a4ba-05c0a27ccc91 · outbound

This paper cites Grounded-Segment-Anything, April 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Grounded-Segment-Anything, April 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.599989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.282635Z digest=sha256:0db0fad0467cfbb74e480521419be7bc316c7549ac4ff7f813f77f22de2e123f

Observation e7611f13-3427-4ffc-95e9-1f934b951d0a · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.584119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.287077Z digest=sha256:ace1a98fadc7d950b2b05260ec91cb284a5e9533d796844b6670c698f2eb9511

Observation e2db01af-d7e7-4e13-ab4f-7da514d8336c · outbound

This paper cites spaCy: Industrial- strength Natural Language Processing in Python.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent spaCy: Industrial- strength Natural Language Processing in Python

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.291456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.291456Z digest=sha256:f91447726f9df0fdb74ffb95c03b7432ba8f8b7e1cba632dc8cd8b0c92b40394

Observation e976bf36-8cf8-4430-a2c0-099f8301380e · outbound

This paper cites LangChain, October 2022.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LangChain, October 2022

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.559163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T20:28:49.296167Z digest=sha256:d37bab0b2322d7ab1bfa2ff6bae93ad8683ae585df6e3f372276ce21d8d25c16

Pith citing papers

Observation 21829e93-2341-4698-bfb8-1df7c5a62bdf · inbound

Mitigating Diffusion Model Hallucinations with Dynamic Guidance cites this paper.

Mitigating Diffusion Model Hallucinations with Dynamic Guidance Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T11:24:16.280713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:24:16.280713Z digest=sha256:3cdebf847b43631e583e2e193b001831091d172711b1c65ce77a0a7f658cc831

Observation 1c1d5067-e92d-446a-b227-13c7d55f2d7b · inbound

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models cites this paper.

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:10:35.145783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-11T13:10:34.521533Z digest=sha256:5fd4ec17b5b3d9e5881b753e051323cbb370c027e579f646f74ff34a6ab5206e