Pith. sign in

Paper Citation Record · LEDGER

Test-time Prompt Refinement for Text-to-Image Models

As of 19 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 3 inbound Pith citation observations for arXiv:2507.22076.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22076 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:01:52.609484Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:13:10.574530Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:29.278949Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43569dd6-df28-472c-b09b-7dd3ef991d24 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Test-time Prompt Refinement for Text-to-Image Models Flamingo: a visual language model for few-shot learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.176538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.176538Z digest=sha256:8925c7788161b64a40a3ca20bc0991bdc0f5cdfe529b6b077963ab97bc936870

Observation d086ba3e-d8bd-4d9a-b8bc-3c593c135bf0 · outbound

This paper cites Blended latent diffusion.

Test-time Prompt Refinement for Text-to-Image Models Blended latent diffusion

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.721694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.182180Z digest=sha256:1b2435343b515af863fba9be381686085d346f36606ac97fea76c65c06f6181d

Observation 2e92cf0d-7eeb-4d4c-9d31-50ebba10271f · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Test-time Prompt Refinement for Text-to-Image Models eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.186385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.186385Z digest=sha256:5552f0f4aa489c554e69197d57aefe7b0f453e63bcc0c0bf0e11abfc764f1bad

Observation 7f4b6d50-f2df-4c29-b53a-be447b2216c3 · outbound

This paper cites Multidiffusion: Fusing diffusion paths for controlled image generation.

Test-time Prompt Refinement for Text-to-Image Models Multidiffusion: Fusing diffusion paths for controlled image generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.700213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.191212Z digest=sha256:8db32f642406f6dc36b428ceea235b20f81273c2beee28ee2b6e307d5a193811

Observation 9c19095c-5274-4f7e-8dbd-ea1d0be6be35 · outbound

This paper cites Improving image generation with better captions.

Test-time Prompt Refinement for Text-to-Image Models Improving image generation with better captions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.679210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.195527Z digest=sha256:12c65bcf9a33aea5253bf861a93545f31db96dde172597355bcafa13bfaaa282

Observation f1065118-1e2e-437d-8dbc-4419629261b4 · outbound

This paper cites Training-free layout control with cross-attention guidance.

Test-time Prompt Refinement for Text-to-Image Models Training-free layout control with cross-attention guidance

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.663088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.199854Z digest=sha256:3d34915cf8ea6705280e8e1e86dc93a881979975702ab0f338ca6676fc091df6

Observation e5531e3a-38e1-423a-8b42-ad44ed35c59a · outbound

This paper cites Masked-attention mask transformer for universal image segmentation.

Test-time Prompt Refinement for Text-to-Image Models Masked-attention mask transformer for universal image segmentation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.645208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.204947Z digest=sha256:784b768ddf331f61fc7fabf5b2e72f1f33101e7cdf007b782613a541aea24870

Observation 7a840b82-e008-4775-abc9-d7e62f1a7d72 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

Test-time Prompt Refinement for Text-to-Image Models Cogview: Mastering text-to-image generation via transformers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.627074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.210094Z digest=sha256:1fcbc51f54f837104614bfd26aa5a942bcb72a598ef45e7e7b02ab7a35044555

Observation 8cad3c74-3d26-436f-ac96-f036b47e3f99 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Taming transformers for high-resolution image synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.215333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.215333Z digest=sha256:4d642a81759c80cc9133e99a656951383d7d5092ac6a75500a23c32a1577daba

Observation 41cb1387-aa9d-4c76-ae4e-f2046d7fd1be · outbound

This paper cites DPOK: Reinforcement learning for fine-tuning text-to-image diffu- sion models.

Test-time Prompt Refinement for Text-to-Image Models DPOK: Reinforcement learning for fine-tuning text-to-image diffu- sion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.598235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.220132Z digest=sha256:d8f50f4a448294799d789a4b4c0c6619dc92fe328b136ac3bb587bb7c48ccf64

Observation 2ec0318e-7b3d-4610-9ee5-19cc19ca7bb2 · outbound

This paper cites Layoutgpt: Compositional visual plan- ning and generation with large language models.

Test-time Prompt Refinement for Text-to-Image Models Layoutgpt: Compositional visual plan- ning and generation with large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.224431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.224431Z digest=sha256:96c68ff671ba7be5c6ae8e1e0d84df336a38c1871359fb53be4f0e5bc0478c88

Observation b668eda1-6ab9-4385-b393-4e9b5541dd2f · outbound

This paper cites Layoutgpt: Compositional visual plan- ning and generation with large language models.

Test-time Prompt Refinement for Text-to-Image Models Layoutgpt: Compositional visual plan- ning and generation with large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.570232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.228582Z digest=sha256:5b67cfa6b46cdec30df039858e968dff8a7ea99d72c4b08658dabe047da333a9

Observation d2feb80e-38fb-46c3-bf8f-41a51c693c73 · outbound

This paper cites LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts.

Test-time Prompt Refinement for Text-to-Image Models LLM Blueprint: Enabling Text-to-Image Generation with Complex and Detailed Prompts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.232968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.232968Z digest=sha256:3d46b1aee9757966dcc715fae68d1c9b583f1fe0cc660424aae238d0e6886629

Observation 7fb07d2c-bc98-4ba5-a989-8c9a404d06ba · outbound

This paper cites Geneval: An object-focused framework for evaluating text- to-image alignment.

Test-time Prompt Refinement for Text-to-Image Models Geneval: An object-focused framework for evaluating text- to-image alignment

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.551291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.237769Z digest=sha256:708454ea80f6a76c58552b521a25d2143aa4c9afbe40611ce0277db14f4710c9

Observation 607496b9-0335-4ba5-8b1e-70b831dbe442 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

Test-time Prompt Refinement for Text-to-Image Models Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.241657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.241657Z digest=sha256:7602bf04a04f53ca8928aca14db60e955b55dee7f1859a4f416006ed358908d7

Observation 5d66487a-27c2-485b-a521-03b2ffe11767 · outbound

This paper cites GPT-4o System Card.

Test-time Prompt Refinement for Text-to-Image Models GPT-4o System Card

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.246022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.246022Z digest=sha256:aad2517d9fa014271ce64ee0500d68314533446c28a536f9badb080301582450

Observation 7b7040ec-03fc-4003-bbd9-b6762caf11d4 · outbound

This paper cites Few-shot classification and anatomical localization of tissues in spect imaging.

Test-time Prompt Refinement for Text-to-Image Models Few-shot classification and anatomical localization of tissues in spect imaging

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.529471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.250084Z digest=sha256:958fe61197990c03759101be975ff2ccfa9682ac8750cfa7fe63107a1028c2ef

Observation f334e179-6c0b-4836-805a-a3a84e487e3a · outbound

This paper cites Clas- sification of microstructure images of metals using transfer learning.

Test-time Prompt Refinement for Text-to-Image Models Clas- sification of microstructure images of metals using transfer learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.510571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.254159Z digest=sha256:2176e8de8ed4029a336a145dba3512927ebd9bb085f7518c1a8b61da8c765e07

Observation ac1bfb0e-b880-4f05-a352-0d20f21c9884 · outbound

This paper cites Alina: Advanced line identification and notation algorithm.

Test-time Prompt Refinement for Text-to-Image Models Alina: Advanced line identification and notation algorithm

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.492598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.258108Z digest=sha256:9bf51c546eb141d4d3f91deb73190c2865257aedd2a3d2564c87b39f2eaddeb0

Observation 4fbd46a4-5c4d-46b5-96cd-ec8e50de5cb6 · outbound

This paper cites Gen- erating images with multimodal language models.

Test-time Prompt Refinement for Text-to-Image Models Gen- erating images with multimodal language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.476091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.262537Z digest=sha256:49afa20649a598e90f893ee8c278b716367a8b1f3a89c0c92fe5a092ee3fb6e3

Observation 1e8729aa-64a5-400f-8ff6-9fa8c1c86155 · outbound

This paper cites Zero-shot Text-guided Infinite Image Synthesis with LLM guidance.

Test-time Prompt Refinement for Text-to-Image Models Zero-shot Text-guided Infinite Image Synthesis with LLM guidance

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:01:52.927761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.267531Z digest=sha256:6dbf824186af8aab627c414550e4b0474b0017094d8c1005f0686f47621fa6a7

Observation 0b04754e-ca52-4c10-965f-b0fc759b9aa5 · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.458307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.273106Z digest=sha256:a0b8bbafc98207d428843eaa2c503c57f9ac2204cff31dd7c88bc468d3d29d45

Observation c458aae0-ddd2-42cd-8521-ce422abd5c84 · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

Test-time Prompt Refinement for Text-to-Image Models Aligning Text-to-Image Models using Human Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.277213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.277213Z digest=sha256:e72ef0e08b946764d4e06a7d70329475994a94fa743d352e97edc5fbafad69ae

Observation ed207dd5-73fc-477d-b770-76d4b4edfa32 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Test-time Prompt Refinement for Text-to-Image Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.281822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.281822Z digest=sha256:ab652e9947cef60d09bf73847d948f0cfafd1c8d023b3198909ff5b0b5052a6a

Observation b870a995-e2de-4ba2-9621-f8a92b33a46a · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Gligen: Open-set grounded text-to-image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.429341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.285748Z digest=sha256:8329bf4bda811843fe3b35e83e4513eccc96d47cbc35fdba4c35e2faeb602467

Observation 0f586462-6019-4bfd-a6d6-84efb249685e · outbound

This paper cites LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models.

Test-time Prompt Refinement for Text-to-Image Models LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.290435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.290435Z digest=sha256:b466e4db8cd18dd58e4f87d2a89dc915a500fdcc1b7e5ba1502f720ff3b9235e

Observation 666964f2-506d-4ca3-860c-6dda68be8baa · outbound

This paper cites VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning.

Test-time Prompt Refinement for Text-to-Image Models VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.294824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.294824Z digest=sha256:b25af6f53cc31d0027fd584b19f2c4c90e5b67f19a134b1df892af47e9d797ed

Observation d2b90bad-7946-40f5-a4b1-644a8afd93a2 · outbound

This paper cites Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps.

Test-time Prompt Refinement for Text-to-Image Models Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.299646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.299646Z digest=sha256:e8c4e099c5481356ef02c5c2ba694aef1b53744815b31faeac68cc68a572950a

Observation c69a07c1-9aec-482b-b49f-5660feabae24 · outbound

This paper cites PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models.

Test-time Prompt Refinement for Text-to-Image Models PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.303932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.303932Z digest=sha256:e025f20fa9aa2a482b9013c5120f8c39258991bfb93485df442155c128e45270

Observation 0827f241-f3f2-4252-abbf-8283292521de · outbound

This paper cites Simple open-vocabulary object detection.

Test-time Prompt Refinement for Text-to-Image Models Simple open-vocabulary object detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.409155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.308732Z digest=sha256:7c37cff918123d3d441dc12942137131ccf815422bad334f0ef107db8e9fb805

Observation 7e5399fb-d84e-42d3-b10c-bd914fde2f4b · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Test-time Prompt Refinement for Text-to-Image Models GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.313124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.313124Z digest=sha256:1b590ad77d1c4e246d0bb02e176c099986bd2944fe550070ecb69f6fbc8c236b

Observation 3edf1190-2cbd-4ab6-9b5a-52d72c946208 · outbound

This paper cites Localizing object-level shape variations with text-to-image diffusion models.

Test-time Prompt Refinement for Text-to-Image Models Localizing object-level shape variations with text-to-image diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.390860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.318058Z digest=sha256:dd4c2f20c409d7ae0fe92a1e8463ea8c1d1373c1e0c8a57c65871902c084093a

Observation 7e66b17b-116f-467a-a772-e431decfacdc · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Test-time Prompt Refinement for Text-to-Image Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.322258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.322258Z digest=sha256:dea73a406e6018e1c4d260e907bc074a0ddc7597b1105ac8f289658e8fd0e29a

Observation 8fc2749c-911d-4d41-a400-bcca2f8dbe4f · outbound

This paper cites Diffusiongpt: Llm-driven text-to-image generation system.

Test-time Prompt Refinement for Text-to-Image Models Diffusiongpt: Llm-driven text-to-image generation system

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.327594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.327594Z digest=sha256:c9cc09bb4291771c97d7d3e82f2b4809a31ccd0e2e9abe8899699282472aeb05

Observation 4be411d5-7a06-439d-8ffb-7f5fd95a4a65 · outbound

This paper cites Layoutllm-t2i: Eliciting layout guidance from llm for text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Layoutllm-t2i: Eliciting layout guidance from llm for text-to-image generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.373242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.331685Z digest=sha256:3a90af3df56da0c21d5e3e42ba44f1701bab9153ee74d3c9a74d10e2ba19c216

Observation 95ae111d-5157-41a2-bd76-2f672cfde516 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Test-time Prompt Refinement for Text-to-Image Models Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.356271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.335654Z digest=sha256:9f2b8266f280cae68948200c44820d79b82caec81f4bfeb39ab121a4b9b57221

Observation dbe6e7c9-470f-4e2d-94b5-667c65a5c52f · outbound

This paper cites Zero-shot text-to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Zero-shot text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.335735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.339583Z digest=sha256:e67aac1f235ec5cecd2c5a8cee63094164aebacc618a51dfacfb03b28cea4811

Observation f24677d5-041d-4729-8ad2-7a39540dd740 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Test-time Prompt Refinement for Text-to-Image Models Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.343453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.343453Z digest=sha256:4da3138b48ca002088587816fc1d1e1e9a2da4de4b93091ce5ead4bc17530397

Observation 5b6bbfd6-dd56-4e85-bb72-145d4820cb02 · outbound

This paper cites Generative ad- versarial text to image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Generative ad- versarial text to image synthesis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.347957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.347957Z digest=sha256:fa37f1d050cb8fda095b8b87862dbc941e81dd9a6789940e95fb10bc366a6c0a

Observation 054a60bb-6615-44f5-bcb2-ca5f51cd2a55 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Test-time Prompt Refinement for Text-to-Image Models High-resolution image synthesis with latent diffusion models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.307768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.352213Z digest=sha256:31ecef716d2ecb23460c86b7391fb185551a061bb66bf667dfd048265ad7c5da

Observation 34456057-989b-4c33-95f7-f18177b0478c · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

Test-time Prompt Refinement for Text-to-Image Models Photorealistic text-to-image diffusion models with deep language understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.357931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.357931Z digest=sha256:a6526cc5681b4ee5b846bcad6852eb9274648c284cf60507cc2b4a22529872ad

Observation a3b624fe-b3e8-4325-9b8a-77b0efa9309a · outbound

This paper cites Benchmarking awesome diffusion mod- els.

Test-time Prompt Refinement for Text-to-Image Models Benchmarking awesome diffusion mod- els

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.281385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.362885Z digest=sha256:7fb00885c363a738210ed273ff8f7f1e46cbe15e515c6fc33df666b3eef5f4dc

Observation 6f5a8638-aeef-488c-9eb8-5dd93c618c76 · outbound

This paper cites Df-gan: A simple and effec- tive baseline for text-to-image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Df-gan: A simple and effec- tive baseline for text-to-image synthesis

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.249741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.371093Z digest=sha256:af30b8e7e85c2432f9dc8f9bf9b52ab80921427c56b7afe42d8cfa2f5f2b5aee

Observation 52b951c6-ecf2-45a9-ab21-7763cce1a45b · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Test-time Prompt Refinement for Text-to-Image Models Qwen2.5: A party of foundation models, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.233186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.375107Z digest=sha256:08ac4f8563fb7a9a7467d20b47bc2cb4c40402d7f447bfc2ee84835eb6340579

Observation 8757f061-212b-41f5-9a4d-2d7b2673e23a · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

Test-time Prompt Refinement for Text-to-Image Models Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.379789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.379789Z digest=sha256:f20e578098e3a331b1f4da510ae095425cfc87e7db9930095e6795d43998a9d0

Observation 306d0874-42d8-4f2c-8762-a23c7dabc5b3 · outbound

This paper cites Self-correcting llm-controlled diffu- sion models.

Test-time Prompt Refinement for Text-to-Image Models Self-correcting llm-controlled diffu- sion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.218936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.384252Z digest=sha256:6987a359a8e77f38aba6edf7f1ac505eaef02e13158921657fb2277645ca53e6

Observation c40d7134-6a2b-4b31-99d2-0c92c2e7e2f6 · outbound

This paper cites Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion.

Test-time Prompt Refinement for Text-to-Image Models Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.201848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.388650Z digest=sha256:589bf1c176562eaeaaf55fc84b86840f5894971e802af6a7b7a767843c1b67a2

Observation 585525c1-dfa8-4590-8882-11ae613bb283 · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

Test-time Prompt Refinement for Text-to-Image Models Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.187417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.393591Z digest=sha256:da671d5cec2f499be97c7d27de26e37489631d05be47dc2dcfb9d4b2d85e4615

Observation e1077ab1-cc15-4f07-adb7-24b287ce6577 · outbound

This paper cites Qwen2 Technical Report.

Test-time Prompt Refinement for Text-to-Image Models Qwen2 Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.540099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.540099Z digest=sha256:6c3c9bda77e0f1d835f8d2c8d81da4b0438449059eb43f93889337aba797fa2a

Observation de47594c-35f6-4206-8575-02af20cd8924 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion mod- els.

Test-time Prompt Refinement for Text-to-Image Models Paint by example: Exemplar-based image editing with diffusion mod- els

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.172630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.547403Z digest=sha256:d344fb0bb55083551a8c059840e62db2afbb51b343d2139bfd58a3b3d08fe7cb

Observation f64aab34-d6b6-48cc-8bc1-1d97b6cceb17 · outbound

This paper cites Reco: Region-controlled text-to-image genera- tion.

Test-time Prompt Refinement for Text-to-Image Models Reco: Region-controlled text-to-image genera- tion

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.554275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.554275Z digest=sha256:049f19757784f4cf9a88ed570d77982bd35c1f9c8da206cbcf333a55a6f526d7

Observation 958ad8bd-cfe6-412a-8dfe-7f255a1b844e · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Test-time Prompt Refinement for Text-to-Image Models Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.148697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.562603Z digest=sha256:3ecb2cb9671c4d74c93a428d854a8fcc115c85dd38bde4dc4fb9b90c1e669196

Observation 85a3587b-fc2c-4b79-928e-7d51e7409aa8 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Test-time Prompt Refinement for Text-to-Image Models Adding conditional control to text-to-image diffusion models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.569855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.569855Z digest=sha256:7a2d2b7f8bcd6aa7a8be57317330860407586bc7652aa329626aafa54ba5a5c5

Observation 2513d08e-c38d-43c6-8d7f-5e9f15a9e864 · outbound

This paper cites Controllable text-to-image generation with gpt-.

Test-time Prompt Refinement for Text-to-Image Models Controllable text-to-image generation with gpt-

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.123489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.576873Z digest=sha256:1ce79681823ee32d3c9f82dfdca7fba089e5c6fd06f3e46d1d392da6797c9243

Observation e7e3f439-9ce0-49a1-891c-c8ddd560bf47 · outbound

This paper cites Dm-gan: Dynamic memory generative adversarial networks for text- to-image synthesis.

Test-time Prompt Refinement for Text-to-Image Models Dm-gan: Dynamic memory generative adversarial networks for text- to-image synthesis

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.108208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.593493Z digest=sha256:8b67e9b9c25d39350ac9e0d27708d5cf7cac58ff5b535759824f732b014bfb12

Observation 6c7405c6-5644-493e-8975-3742ad478835 · outbound

This paper cites Controllable Text-to-Image Generation with GPT-4.

Test-time Prompt Refinement for Text-to-Image Models Controllable Text-to-Image Generation with GPT-4

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.585344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.585344Z digest=sha256:95480941deb12dfcbf5a31554577a1eab7ddc02027995eba760dc5d154bbd0a9

Observation b152ec5c-59c1-4354-b5b8-6dd4a6c28df3 · outbound

This paper cites Benchmark Datasets We use three benchmark datasets to assess compositional fidelity, prompt comprehension, and generalization: 9.1.1.

Test-time Prompt Refinement for Text-to-Image Models Benchmark Datasets We use three benchmark datasets to assess compositional fidelity, prompt comprehension, and generalization: 9.1.1

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:01:53.093043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.601255Z digest=sha256:6b6a046683d87f5b87a10af35ca7b5372aa4795f8853e563f443896fc7a48c5b

Observation 0c053d58-a99d-4125-a188-9f923a76db6d · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.078443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.609484Z digest=sha256:83f7986342ceee618af5103f70733178f8983e4989cfc1d02532a2c78a9b6075

Observation 7bb962e7-2f81-48b0-a76f-fcc33561f370 · outbound

This paper cites an unresolved cited work.

Test-time Prompt Refinement for Text-to-Image Models Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:01:53.266861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T15:01:52.367049Z digest=sha256:ed6cacfcb0da49d54c49e5bf0c2c9fd4a0397cb396697764273ae609e128a987

Pith citing papers

Observation 5af3e4d7-3f17-45fc-acab-134cacffe0e0 · inbound

Evolutionary Token-Level Prompt Optimization for Diffusion Models cites this paper.

Evolutionary Token-Level Prompt Optimization for Diffusion Models Test-time Prompt Refinement for Text-to-Image Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:40:57.856590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T17:05:00.723170Z digest=sha256:72902badf8fd50c997e431f8ee77f85abde898539fafddf2694f2b40d0addc31

Observation 5a7d0d4a-ab9d-4063-99a1-367dc1aa8057 · inbound

DuET: Dual Expert Trajectories for Diffusion Image Editing cites this paper.

DuET: Dual Expert Trajectories for Diffusion Image Editing Test-time Prompt Refinement for Text-to-Image Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:29.280305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T06:58:39.254427Z digest=sha256:9242da32fd8d04f71887712dce41434cd7fcb7a5079236aab0ae0f2e4a597c0d

Observation ee1caf41-9d5b-4243-9317-9ad86911a2f3 · inbound

DuET: Dual Expert Trajectories for Diffusion Image Editing cites this paper.

DuET: Dual Expert Trajectories for Diffusion Image Editing Test-time Prompt Refinement for Text-to-Image Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T02:13:10.574530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:13:10.574530Z digest=sha256:6d348503eb252cdeea3675cf7002d41f63a575811b30d98474637f8c408901e1