Pith. sign in

Paper Citation Record · LEDGER

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

As of 14 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 2 inbound Pith citation observations for arXiv:2412.05722.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05722 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:28:49.296167Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:10:34.521533Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T13:10:35.138865Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 151650cd-83d3-4abf-96eb-57aed0039964 · outbound

This paper cites Image quality metrics: Psnr vs.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Image quality metrics: Psnr vs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.761721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.150491Z digest=sha256:39476a4152c05a094049ce74d88b0d58b82036cc583c6381d021543a8dcae4b4

Observation 16a584f9-0b6a-46c1-bb53-3290b3b91e63 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Learning transferable visual models from natural language supervision

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.745500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.155674Z digest=sha256:29d1a3150a93abac390c243ea10ef4e8ad284cb88557eceb784e2824bc52ea46

Observation 21723be7-aca4-4330-b655-629700778397 · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.729697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.160468Z digest=sha256:4ace26c1ee711a0a7ce28a947ddf26c02e03bac1dbd35377e2f63bb42fee2018

Observation 61d19daf-2581-4681-a84a-3a632c9f8207 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.165301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.165301Z digest=sha256:89b801d136ab7507849e3175b9d4c3a831ec2d78ebd0eafa8e20fb68643209b1

Observation 7f9a7d9d-96ab-4de3-aaec-cd9e8c6a90f8 · outbound

This paper cites TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.170451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.170451Z digest=sha256:8ba15222dee14e757c70f040341264378ab0471a29e0c88c347625e69d1befe6

Observation 511d82ce-1bc4-442b-b7ca-fa927803ab0f · outbound

This paper cites Holistic Evaluation of Text-To-Image Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Holistic Evaluation of Text-To-Image Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.176106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.176106Z digest=sha256:e97d4808a7606c29434a28ba10f2e7269f83de22e09db0afca1a4d0a31198549

Observation 05c34299-15c2-4996-928b-9db49254ef86 · outbound

This paper cites Attribute2image: Conditional image generation from visual attributes, 2016.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Attribute2image: Conditional image generation from visual attributes, 2016

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.182060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.182060Z digest=sha256:2692979e500303a6029c979f61b04c31ed7bef05dca41679d2d7f77b34038456

Observation d0f756cc-015b-4b78-afef-4fe2baf8090a · outbound

This paper cites Neural discrete representation learning, 2018.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Neural discrete representation learning, 2018

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.186958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.186958Z digest=sha256:3c4f1159a86fd7ff1bdc29e2f5e089cc6dcccbcf4e0b7ff4329c35680922a378

Observation 7b9d38fd-ea08-40a7-b16d-d866dc62e916 · outbound

This paper cites Generating diverse high-fidelity images with vq-vae-2.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Generating diverse high-fidelity images with vq-vae-2

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.191869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.191869Z digest=sha256:4c1d12b79b1c3ad6f10bb46732b4031a64338a7feb48c984983093a285aa3fa6

Observation b53a17eb-e00d-4b7e-9bf9-5e76e35b8cad · outbound

This paper cites Zero-shot text-to-image generation, 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Zero-shot text-to-image generation, 2021

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.196844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.196844Z digest=sha256:526afbb86da54a70fb9714f6e026d3d6b41642813a4fb8b15f0a8a122ebc16b9

Observation a06f047c-90ff-4521-ab8d-e113cc4dc6a1 · outbound

This paper cites Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.201609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.201609Z digest=sha256:92dcf019ce70a4889a63dadab531c73f3ad582c0ac2cbac5fa63c793e5238c45

Observation 101a44eb-8dbe-41ae-9a2a-3aab9813f416 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.206582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.206582Z digest=sha256:2107c907b54ad91077c4a832268b6eb623d8a0fd37ac43ab84ee4f6ea7705712

Observation aa35b2b2-f0dc-4077-8006-315d50566203 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.211973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.211973Z digest=sha256:f0a5e4f5afc4f54578f0534815f4851103bee337ea15c4b6ffc37fe142401336

Observation 3edfb075-b63a-4127-b96d-f3c069ba6168 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Denoising Diffusion Probabilistic Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.217561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.217561Z digest=sha256:a0ab5a4003138a44f5b53388d6d0450864d1e3796f1f19fb624f6b38a67ff5cc

Observation 27181258-3808-469a-9805-ae6b09b43a77 · outbound

This paper cites Openclip, July 2021.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Openclip, July 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.223090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.223090Z digest=sha256:93bd7af5384a9cdf60ff3211f2c1c6f654dd0b5f00852f42cb2df1627f9d5d2a

Observation 2b495b58-f14c-4303-af11-5ba88ff02f66 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.227867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.227867Z digest=sha256:767d26a11ce9753203d705bd29c2709224806171f8b90459e3b83710b244a4e4

Observation ab0e4879-c23b-4f78-acc7-04c988262364 · outbound

This paper cites Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Siren’s song in the ai ocean: A survey on hallucination in large language models, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.232721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.232721Z digest=sha256:9d11c56cd4cf0c6a4357beb9fb63c25513ed4e516651f3cfaa0bd2b69daafb72

Observation 85ca351c-4b89-4525-95a5-0abd231b05de · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.238323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.238323Z digest=sha256:eefc353bc69c0fbaf3e1cdf10202de930ae78ee34763b4500f4b3b241bdf6f07

Observation 64b3ca76-c1eb-4e71-a13a-5292fc5e873d · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.243359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.243359Z digest=sha256:15790845ee929b8dc2c2f78641d5f8db6e72a6c2cc23f096c428fd32bb0c1105

Observation 2a2f85bc-529d-4b91-825f-f2932054bee8 · outbound

This paper cites Effectively unbiased fid and inception score and where to find them.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Effectively unbiased fid and inception score and where to find them

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.631619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.248470Z digest=sha256:5040725280a58780ebad6bf7b4c9529d2048a40a1e888ebaf702dd90ed57cd39

Observation 76a53b11-a4ed-4a60-8dc9-630a6b048a33 · outbound

This paper cites A Note on the Inception Score.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent A Note on the Inception Score

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.253323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.253323Z digest=sha256:b5ca774118158df2735cd46a511460067012eeca9ae5376f1394aa77a9461b8d

Observation 2a8d3119-2974-495f-b6c9-558251bf0009 · outbound

This paper cites Improved Precision and Recall Metric for Assessing Generative Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Improved Precision and Recall Metric for Assessing Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.258616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.258616Z digest=sha256:dab18b0a76939529065ccb633ba9269fbd037606c223d924f31582bc2c5392e7

Observation f09ccf2e-cc4b-4385-add2-64cd45acaf48 · outbound

This paper cites Benchmark for compositional text-to-image synthesis.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Benchmark for compositional text-to-image synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.615268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.263655Z digest=sha256:34346155944196c6c3c4b57106da52dddc5d82e138f0958a3cf92591443ed193

Observation 89e3bb8b-afe0-435b-b10d-7d7528d42cf9 · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.268128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.268128Z digest=sha256:e1222dae77a1702853061633f8ce1f5d425efd0e18d81be6cef5d8b744e599d2

Observation 8c014d5a-66f2-45f8-8f09-8f782e8a36ba · outbound

This paper cites LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LLMScore: Unveiling the Power of Large Language Models in Text-to-Image Synthesis Evaluation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:28:49.360323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.272738Z digest=sha256:98e09ab02338ddbfff3d3efd9efdb6918afcf34fa703b196faf2a1cf47bd0570

Observation bb5e3565-f2b6-4a75-818f-6e543cf9006d · outbound

This paper cites DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent DALL-Eval: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.277601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.277601Z digest=sha256:41a02dcdbe973a46854ec0a6936cfe89d8d826798d9ff3974c4856e969bc295c

Observation 150ae64c-8f65-4f22-a4ba-05c0a27ccc91 · outbound

This paper cites Grounded-Segment-Anything, April 2023.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Grounded-Segment-Anything, April 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.599989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.282635Z digest=sha256:4611b458274421fab304cd04dae7d99810b8bf18c5353373dc05a23a762d2403

Observation e7611f13-3427-4ffc-95e9-1f934b951d0a · outbound

This paper cites an unresolved cited work.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-11T20:28:49.584119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.287077Z digest=sha256:deef1590bbb1a1873156d90c94772a7b3d8e14111d23327f6596ccde60690500

Observation e2db01af-d7e7-4e13-ab4f-7da514d8336c · outbound

This paper cites spaCy: Industrial- strength Natural Language Processing in Python.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent spaCy: Industrial- strength Natural Language Processing in Python

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:28:49.291456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:28:49.291456Z digest=sha256:f2f87a003a73eb5f98ecee5cfc5f91c61749021537774f2e50a57c1668e5918a

Observation e976bf36-8cf8-4430-a2c0-099f8301380e · outbound

This paper cites LangChain, October 2022.

Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent LangChain, October 2022

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:28:49.559163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T20:28:49.296167Z digest=sha256:b108befe0d309e731aebd9167965b12d0640b58f958aeb5915cde25e3e033b72

Pith citing papers

Observation 21829e93-2341-4698-bfb8-1df7c5a62bdf · inbound

Mitigating Diffusion Model Hallucinations with Dynamic Guidance cites this paper.

Mitigating Diffusion Model Hallucinations with Dynamic Guidance Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T11:24:16.280713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:24:16.280713Z digest=sha256:96153bb56a90fcecedffb5d880671dc9489eec6b4ee4de3cf5a6e53a9c4a0502

Observation 1c1d5067-e92d-446a-b227-13c7d55f2d7b · inbound

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models cites this paper.

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:10:35.145783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T13:10:34.521533Z digest=sha256:f4e6a9f63f77a94466fa044d09c70b979d17e18834c189eb90710f6a68d40e9f