Pith. sign in

Paper Citation Record · LEDGER

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models

As of 9 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2506.08480.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08480 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:14:43.033978Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T21:41:10.852265Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T14:25:46.015089Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d04e99d2-52be-42c6-94ac-a5110609acfc · outbound

This paper cites online" 'onlinestring :=.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.534295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.534295Z digest=sha256:e85181cb99b0f022c1b8dce8114e04cfbdae42110f778f595909c69c6348f351

Observation 343fa16a-e992-48e1-a8e0-44f6ae564951 · outbound

This paper cites write newline.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.623732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.623732Z digest=sha256:b5e95f0d04013c26436507b96eb51ef252c53fe553b94104225aac2fd9676b6f

Observation 722e85eb-e0e0-49f8-8272-117a6434adc6 · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:44.606532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:40.739452Z digest=sha256:832be3a24e9d5bf9ba94c8c51e7f5b04d02ffd5172183e4e645cec7c73713145

Observation 00f7d635-966c-4406-b3de-4ab59a097c14 · outbound

This paper cites PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.852888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.852888Z digest=sha256:b72c217722dbaed85715fb5a10cf924bc4d5b5c1f7166491bc18f801646f5e56

Observation 1eaedb72-0415-4070-ab85-23d82c30fe60 · outbound

This paper cites Unified Hallucination Detection for Multimodal Large Language Models.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unified Hallucination Detection for Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.971289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.971289Z digest=sha256:396a48b164884a74a6cea129c95f1718ae5cfbb24ff44ea32940e5e501616e64

Observation f3a1352f-1c68-49c4-9b6d-c80204ad3006 · outbound

This paper cites Davidsonian Scene Graph: Improving Reliability in Fine-grained Evaluation for Text-to-Image Generation.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Davidsonian Scene Graph: Improving Reliability in Fine-grained Evaluation for Text-to-Image Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.111755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.111755Z digest=sha256:9a6ee843db3c5426a763c70e9193e0039289ff18bb4b674141d8f9c51eaccf16

Observation d1163b47-90fd-4dce-bc08-9f769c0caa6c · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:44.449158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:41.224534Z digest=sha256:ca1a86ce8b8fbc6a7ab1dbff10baf1551631cd78164d76c511528b72dc9948e4

Observation 8c0c5ef2-2649-4650-bc27-89785d034008 · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:44.286273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:41.339032Z digest=sha256:231f3529a0617eff46ac447fac419ab0a35fcb8af1ac1381ace4c25dc2c4ae9c

Observation fbaee488-b011-4168-9d0d-9bf5f8ac655b · outbound

This paper cites Benchmarking Spatial Relationships in Text-to-Image Generation.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.429534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.429534Z digest=sha256:2520b37f0d929f86503dd223e58e0e5ff4c5449ef3746a48d217594ad3785fb9

Observation 8300ec3c-da9d-49ed-8dee-fbd00f8a1a0d · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:44.114397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:41.491640Z digest=sha256:ebe1bb11ed957018789131dcd68fa67fa079183b3f1397055fbf86370020ede4

Observation c6ee1981-6af1-4362-a710-44f1bc5f232b · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:43.921102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:41.603219Z digest=sha256:462e9a0c054fb5432cd0aab13e4f77f97515c039af3b9409b0cbcd4d7b61e27a

Observation 1240962d-56f1-466e-9851-e69f16a28d70 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.673571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.673571Z digest=sha256:cc8777b8d121ed3b2b1935661108d362c0367ca54feb7b7da7a462abf6e13483

Observation a232d6a6-d3c5-4df7-b19f-ef5a65166b19 · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:43.769144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:41.769405Z digest=sha256:1a483bf35e1554b5882de89a9411d60ad86a350f28dac4a66d07595c036e2996

Observation e9c4f316-7fac-4712-857c-a97cf24242ca · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.914104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.914104Z digest=sha256:233cf872199b3e432ee31590f56f63ee62392ccc399931b1ba0bb4c993260b73

Observation 65141a82-4e53-4557-8949-302640b40b8b · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.025600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.025600Z digest=sha256:ba7c6369212d8a186240e276506cb04b8ea23013d89fd0759f6cf642225bc9a4

Observation 703f16e1-cf9f-40a6-b78c-b678b5ea79e7 · outbound

This paper cites Evaluating Text-to-Visual Generation with Image-to-Text Generation.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Evaluating Text-to-Visual Generation with Image-to-Text Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.136106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.136106Z digest=sha256:9c5ffe5f38ca9bab56451db6e27b860e6f6a0938f47b854b274c97feb2bd4a00

Observation d594fff6-ce5a-4e96-a3a5-c3439ade17c1 · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.211336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.211336Z digest=sha256:d3dfc2c147a9931a9a530e802809a4dd41ddfebd3999d8f4ca5edfdc608c7c8a

Observation 0c16d08f-ed4d-4751-909a-13ebc3bd7053 · outbound

This paper cites ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:14:43.248885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:42.297692Z digest=sha256:2e904d33b5c3fe98f64c234b759fa77bf9544eac9ce19ea719cc498e07a591e3

Observation 29200d13-37c8-4cb1-ab83-e358a11f303b · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.399345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.399345Z digest=sha256:37cfdad8aec0e0735421a1d14fcadd386494509df71a497afc2fc2570ab378df

Observation e8109674-c308-4a2f-8541-44336b9f7ace · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:43.624612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:42.514345Z digest=sha256:aeea2a524e3f6aa17feea96e81756803728aa75d26b6ab9cdbf4c31169df2b96

Observation 76a2795c-931c-498f-8ff0-7b2d9ac01e45 · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.598591Z digest=sha256:982fb40ec5d24f84a8e26fb311d9534ee3858db5279566700e4c13b904921458

Observation 36fb1309-36bf-4877-aa34-1faa11774360 · outbound

This paper cites Revisiting Text-to-Image Evaluation with Gecko: On Metrics, Prompts, and Human Ratings.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Revisiting Text-to-Image Evaluation with Gecko: On Metrics, Prompts, and Human Ratings

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.724690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.724690Z digest=sha256:98a6c9673b59191e26811bf793117c5be33867eb0c28759b9c151a0597460f67

Observation 3e61048c-bee1-4d80-a03a-b7b088b2dfe9 · outbound

This paper cites ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:42.807933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:42.807933Z digest=sha256:87a08ecfb9a3dce429c7aaa70ce8bb9bda0aa9e35a101d86dd0944ebd131d2d1

Observation 147871f9-ef9e-4b52-9c68-930433ea2c6d · outbound

This paper cites an unresolved cited work.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:14:43.490295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:14:42.939120Z digest=sha256:438dc87e2e647909120be13e37a765f139f165613b8dc3f9ae2d1f511b778f6d

Observation 740ac670-3217-4895-b2a1-23c8ae371472 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:43.033978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:43.033978Z digest=sha256:0b0fabab668a25bde1811be13af5f6442bdc2e1ba957e0ef2dcb1d1a85fe2b53

Pith citing papers

Observation c3748a19-91d3-4940-a10f-6e09799c29a7 · inbound

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis cites this paper.

Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.016575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T21:41:10.852265Z digest=sha256:5269b309cdf9d710dba9f28c88be782cf0aa13e8e537785a1abfb11db37dfa05