Pith. sign in

Paper Citation Record · LEDGER

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models

As of 23 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 0 inbound Pith citation observations for arXiv:2412.13195.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.13195 v2

Coverage vector

measured 85 of 85 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:24:10.973352Z

measured 85 of 85 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

85 of 85 outbound references displayed

  • verified exact2
  • verified fuzzy55
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation af09a962-410c-4633-baf7-fa361a82a668 · outbound

This paper cites Hrs-bench: Holistic, reliable and scalable benchmark for text-to-image models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Hrs-bench: Holistic, reliable and scalable benchmark for text-to-image models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.448204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.448204Z digest=sha256:200aa5831985f38909edaa2b506f1d03be1055ed5e2bec571111a26143935a30

Observation 8c18e47f-eaf8-408d-a348-aee5213e7c4a · outbound

This paper cites Improving Image Genera- tion with Better Captions.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Improving Image Genera- tion with Better Captions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.454506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.454506Z digest=sha256:ffbb3fb2a83f0f28fd4e56d01a1d8805ca91f1af44308637da9f2a6cd84928d3

Observation 0de1cc4c-a751-4bb9-8d41-f9b0bc750704 · outbound

This paper cites Training diffusion models with reinforce- ment learning.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Training diffusion models with reinforce- ment learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.460598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.460598Z digest=sha256:463b5a3eef039cb8779824caabb9010cf6046ed9721797812406c2d4a46f1412

Observation 3d84197d-9f93-4fdf-a324-4bcea850789f · outbound

This paper cites FLUX.1-dev.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models FLUX.1-dev

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.466530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.466530Z digest=sha256:00c0229394b14e3c1814ceba9076a954ec97e41d8734e7ddf14ce5789f373da4

Observation cb2c6247-805d-40ac-9042-dfc3950a54fc · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.472366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.472366Z digest=sha256:391a82121e3e1f87f9342176f24720798813ca90b3e9947529196b4560f0ee01

Observation 3363fdc0-0f7c-416f-b1ef-9ca6e21ecde5 · outbound

This paper cites Getting it right: Improving spatial consis- tency in text-to-image models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Getting it right: Improving spatial consis- tency in text-to-image models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.796356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.477951Z digest=sha256:1b7282e786eee49619ffae3614167abefd2fee62d2a79e0887bd80177ee23314

Observation b9fa0cb9-321a-4865-8d00-b8c45627b3c6 · outbound

This paper cites Attend-and-excite: Attention-based se- mantic guidance for text-to-image diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Attend-and-excite: Attention-based se- mantic guidance for text-to-image diffusion models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.778961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.483179Z digest=sha256:faf55646a9082fcba4c63e267307c2c66f3b5b02f9ed841e5e7cf983b9802d89

Observation 102c9a39-a694-47f6-a416-0473d30b262a · outbound

This paper cites Cohn, Dayou Liu, Sheng-Sheng Wang, Jihong Ouyang, and Qiangyuan Yu.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Cohn, Dayou Liu, Sheng-Sheng Wang, Jihong Ouyang, and Qiangyuan Yu

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.760862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.488978Z digest=sha256:9ba6983d807e96b72bec7a285d196f497b4c83e8ffd082402084766c69978ae5

Observation 8da81bf0-c142-4669-a306-17dce0b2cba5 · outbound

This paper cites Cohn, Dayou Liu, Sheng-Sheng Wang, Jihong Ouyang, and Qiangyuan Yu.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Cohn, Dayou Liu, Sheng-Sheng Wang, Jihong Ouyang, and Qiangyuan Yu

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.738552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.494138Z digest=sha256:e8689cec82aa5c21294e381d8ef3d8ad1c7bd34a7eaf1ea465e2c3be76628899

Observation bca47ade-3f40-4842-acb0-e5282a6a1d3c · outbound

This paper cites Pixart- Σ: Weak-to-strong training of dif- fusion transformer for 4k text-to-image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Pixart- Σ: Weak-to-strong training of dif- fusion transformer for 4k text-to-image generation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.719647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.498844Z digest=sha256:d48386e8523046fbb02b01cd34666ebbde7080071cff41df6bf98b7dbb87ab12

Observation 7a101917-882d-4623-85e1-d04b43433a7d · outbound

This paper cites PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.503518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.503518Z digest=sha256:e6ae8c950d0a53485fc598b1e39af99327aabc930a8e95f5c82cc8620051e816

Observation 4d672181-2995-4cbc-94dc-4420a8875e78 · outbound

This paper cites Kwok, Ping Luo, Huchuan Lu, and Zhenguo Li.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Kwok, Ping Luo, Huchuan Lu, and Zhenguo Li

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.695268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.509234Z digest=sha256:1031522da8bca20c3d1c45b1dfa4772c5c61b90765e1ffbd8567903d46312c26

Observation d42d775a-fb74-4ae0-9514-4c1a903ad14f · outbound

This paper cites Training-free layout control with cross-attention guidance.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Training-free layout control with cross-attention guidance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.666729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.515304Z digest=sha256:9403f6a0d3117999849b5557b917f7f5728b1ba7ffafa45d4bfc14a904983066

Observation 96f2c72e-714b-4ec6-ab8a-9f2773381b35 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Reproducible scaling laws for contrastive language-image learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.648255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.520846Z digest=sha256:57a67ef22ad66156750df9ef8f205c1480039d173755cfd69b9d4408bfcf83eb

Observation 091e9028-2ce8-483b-886b-04b61b8c425a · outbound

This paper cites Visual pro- gramming for step-by-step text-to-image generation and evaluation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Visual pro- gramming for step-by-step text-to-image generation and evaluation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.625385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.525705Z digest=sha256:2f28013b5e0df572ef5f9028d69ebc445471b368d572070da4be2ce4a4852184

Observation c93a0399-66fd-49be-9fef-7ed72e56cea7 · outbound

This paper cites Coventry, Merc `e Prat-Sala, and Lynn Richards.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Coventry, Merc `e Prat-Sala, and Lynn Richards

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.606074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.532228Z digest=sha256:bce96544a965dcc558c2e036167a362ec0211aa92d1a432345cc7a65b3e92bb3

Observation 63e64da8-b7bf-4218-90ac-21cbe38ba662 · outbound

This paper cites Dall·e mini.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Dall·e mini

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.588078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.540376Z digest=sha256:03c4f8df44400285d5a308453b932fbbe5dcb48393271d20cb2901d718cdb32a

Observation 132aeb3f-f4b2-45d5-80ce-8ce3e606c3ad · outbound

This paper cites Diffu- sion models beat gans on image synthesis.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Diffu- sion models beat gans on image synthesis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.546909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.546909Z digest=sha256:5408960d82f1c93c61e8ebdabe74869ffe30862cc9a41d969f469400f50e6fa1

Observation 26028663-c670-4f62-82b9-c724ffb288d9 · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Cogview2: Faster and better text-to-image generation via hierarchical transformers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.552414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.552714Z digest=sha256:5030f8467eb3aaea5c78459cdae0bc9962581b4ccf6f52e785fd826533215db3

Observation f16bbd7e-d132-4e80-a5c7-f744889a73dc · outbound

This paper cites Scaling rec- tified flow transformers for high-resolution image synthesis.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scaling rec- tified flow transformers for high-resolution image synthesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.531405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.559790Z digest=sha256:d84df8249a600beba53ff32ec017f7a42c979a813dae0a4c7ba2da6eaa5ed14b

Observation 3fc00ec8-d18d-44b1-a426-64a9359772f3 · outbound

This paper cites Re- inforcement learning for fine-tuning text-to-image diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Re- inforcement learning for fine-tuning text-to-image diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.505806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.566012Z digest=sha256:e87c8a77069631b4503321bbb40262a6ef943c4f4e6d449d8bee28aaf47d9680

Observation 15d09b38-ff1a-49f3-bb54-645892f8cabe · outbound

This paper cites Akula, Pradyumna Narayana, Sugato Basu, Xin Eric Wang, and William Yang Wang.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Akula, Pradyumna Narayana, Sugato Basu, Xin Eric Wang, and William Yang Wang

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.478478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.572203Z digest=sha256:279cddf7bccf6e8e272f7c20290dacb2e3407c126f458b207c3beb13a8001db0

Observation 8c3e2e20-3aff-46e8-8794-f14fb1036e13 · outbound

This paper cites Akula, Xuehai He, Sugato Basu, Xin Eric Wang, and William Yang Wang.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Akula, Xuehai He, Sugato Basu, Xin Eric Wang, and William Yang Wang

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.452703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.577769Z digest=sha256:c8eb264e2940a077a68919a21731c486ca528b7186978873badf8c67e2cf1316

Observation 91ef17fe-07c7-4694-8578-06be7ef9ddfa · outbound

This paper cites TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.583658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.583658Z digest=sha256:9550d5fd162bf75cea0d77963d150393bbbf250931a76604157b9b611e336f66

Observation 267c75fe-1f7f-4b97-9340-31f22bde8b18 · outbound

This paper cites Geneval: An object-focused framework for evaluating text- to-image alignment.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Geneval: An object-focused framework for evaluating text- to-image alignment

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.428912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.590061Z digest=sha256:c93e9b7e0b8f6b401e6b62de177ed6ad0012fd3853a5e85ef917b58ac35ad59c

Observation dfd366b9-6386-4ffd-bffe-58f836eda018 · outbound

This paper cites Benchmarking Spatial Relationships in Text-to-Image Generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.596080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.596080Z digest=sha256:8779cbf3b358f92c42a0c8b6108aa51715849b34230d45ea1a44db8d8eddd49c

Observation 03b32048-f7ac-474a-9738-c5fb36383d8c · outbound

This paper cites Diffusion-RPO: Aligning Diffusion Models through Relative Preference Optimization.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Diffusion-RPO: Aligning Diffusion Models through Relative Preference Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.602422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.602422Z digest=sha256:c342544c9ac8dddf0aeb3d632d6dd89b7ebce78c61b49de87193f6fa23ba23d1

Observation 950d7731-c9b4-4de6-8665-0e2eef36c52b · outbound

This paper cites an unresolved cited work.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:24:12.406299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.613575Z digest=sha256:1d175dbc0cf07c55b57b42404ed6fdb38ab3fcabce5510fd610c3517b4e9491b

Observation aac30f3d-c1d3-40c7-901d-ff32e305d6b4 · outbound

This paper cites Ganspace: Discovering interpretable GAN controls.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Ganspace: Discovering interpretable GAN controls

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.386207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.621527Z digest=sha256:f50db18f0c0037b3cecf98b8128bd6d16b2e6d668ead480f527ed0f15e1679fb

Observation b9a90481-b722-45f0-9f08-c19164f43025 · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Prompt-to-prompt image editing with cross-attention control

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.360600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.628033Z digest=sha256:d20c0484a2e6bd460e18c4b554a2282b2e5720bb325cd7865d891d2898988ca3

Observation 620e772f-7014-41ef-a31a-07780e57dd31 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.338897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.633696Z digest=sha256:00d33f9cfd4951a97b7c6fbd778630c9789003aaa1e4101673cff8f2be961818

Observation ec42c18d-9b43-4b46-b951-62d2d7ec4e26 · outbound

This paper cites Denoising dif- fusion probabilistic models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Denoising dif- fusion probabilistic models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.639534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.639534Z digest=sha256:155aa189dd0f308f34827b17ef3b90cb2b816eec7bec5d1d07532a5340c02222

Observation cf04431c-2031-4c3d-a723-48a21ca52776 · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.644541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.644541Z digest=sha256:51ac383d786926313c0db8ed6d310ca01e98ce827dd101e442d5a6866170782f

Observation 92168039-98be-4c48-8a39-d198b6f5546f · outbound

This paper cites an unresolved cited work.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:24:12.305667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.649920Z digest=sha256:eff8650a4390eb6f0e7377c87286e6396c1653262965d0b36e8d7104592e8c1d

Observation 59756ed9-5ff3-4488-929d-139381a3cf15 · outbound

This paper cites T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.285816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.654641Z digest=sha256:6f29aa717787bf630c40a14d8975929d699aad749ff7ca9e66bf0a29c8435bc1

Observation c254fee4-ec18-482b-9b2d-f093ee7f6e6e · outbound

This paper cites Re- thinking FID: towards a better evaluation metric for image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Re- thinking FID: towards a better evaluation metric for image generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.258315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.661068Z digest=sha256:718cb49a2f7b51c564bb6ee89b80ebcdce6652964eaf942f8affd261d60bee35

Observation beeb2cfe-e62b-4b66-b014-2070001f75ce · outbound

This paper cites Comat: Aligning text-to-image diffusion model with image- to-text concept matching.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Comat: Aligning text-to-image diffusion model with image- to-text concept matching

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.226180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.665565Z digest=sha256:3e08e1867763632921738ebeaa9bca19aee708be855551c8f3a24c62ee60d2c7

Observation 095f94bf-d488-4cb7-abd2-fba298eefc5a · outbound

This paper cites Scalable Ranked Preference Optimization for Text-to-Image Generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scalable Ranked Preference Optimization for Text-to-Image Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.670874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.670874Z digest=sha256:4f44dc1a9957af794d7cf634d2b6f80dff0f6e809a31929315c588a17fafc8a8

Observation 69d327ea-1d7e-4a90-bc86-9a689801b250 · outbound

This paper cites Evaluating and improving composi- tional text-to-visual generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Evaluating and improving composi- tional text-to-visual generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.207257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.676906Z digest=sha256:ef2ac074acfcbc0626fd8e7ff36c91611f8c126a0b6a72e426cce7b88d88902f

Observation f90974a1-72a5-4783-99a5-3f2b4038da5a · outbound

This paper cites PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.681938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.681938Z digest=sha256:7f3e800f4945009b887ae77453a95d21e7ae954e8c48c94c9fc47fc6d7239ec0

Observation 6b2af98c-1698-4016-a66d-b7bf95c046dc · outbound

This paper cites Llm- grounded diffusion: Enhancing prompt understanding of text-to-image diffusion models with large language models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Llm- grounded diffusion: Enhancing prompt understanding of text-to-image diffusion models with large language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.189468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.689503Z digest=sha256:6f5bec25ef0b9c6742d060d5c8271b040dcf48a1feb71eba94652f7bfec2d6f4

Observation c9aca626-6551-44c8-a593-4bf3ac87b3f7 · outbound

This paper cites Collins, Yiwen Luo, Yang Li, Kai J.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Collins, Yiwen Luo, Yang Li, Kai J

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.169740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.695624Z digest=sha256:ef0cdc305b4dbaeabf5da0ba2c2157a45250b04ce649b9702e8a432be3dd3677

Observation 765c9c1e-fcd3-4715-9282-cd6e1e1300a2 · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.144736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.701011Z digest=sha256:3d5a53a06a4cf2d048a459cea560054d2f7cdfcce26cd35be9c7d52c0af13797

Observation 54e9693f-1402-4ca4-8c74-b0f2a6506ee2 · outbound

This paper cites Tenenbaum.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Tenenbaum

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.123852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.707640Z digest=sha256:80fa7282c5edf03c6f5556e00c27ee40f746e6f67aef3e34f214ecf67dcf1ee1

Observation 1f5aaf0f-4191-40e3-a279-e9a4f943f8a8 · outbound

This paper cites Repaint: Inpainting using denoising diffusion probabilistic models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Repaint: Inpainting using denoising diffusion probabilistic models

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.104822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.713183Z digest=sha256:0b22c0829799f8851d90ab35b4706e7b0e14409c1178707a3640764e06425888

Observation d0320800-6ae1-466a-adfc-0c3e2a2713ca · outbound

This paper cites Pick-and-draw: Training-free semantic guidance for text-to-image person- alization.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Pick-and-draw: Training-free semantic guidance for text-to-image person- alization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.086178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.718984Z digest=sha256:0d86d90ac4a1b47a027f65e051b5ebb6a9f18f27947ccaedbd7c8c383e3b76fe

Observation 8ef50172-4d70-4481-ad23-6c70ca93dae9 · outbound

This paper cites MidJourney.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models MidJourney

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.063572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.724530Z digest=sha256:9afbafc30c62d0077a64f6bc496e44ca709e07ea918669d7c1fcea504899f9f8

Observation 29ea6c02-6f9a-469c-90b0-022c3e0b4f4d · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.041000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.730516Z digest=sha256:fb6ffa3ca299fe5b0fe30ce5c7f2bb6c84d59f6d65f67151bf7ef623489e5362

Observation e86a4a1c-254c-4a14-bb37-eb7655bbd2d8 · outbound

This paper cites GLIDE: towards photorealis- tic image generation and editing with text-guided diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models GLIDE: towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:12.017373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.735735Z digest=sha256:075d0f694a90540fc3cc215dfafda11a7129f8196759fd50b4ab080283786e72

Observation 926a0225-76de-4d83-8fbf-e0947af5978c · outbound

This paper cites Drag your GAN: interactive point-based manipulation on the generative image manifold.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Drag your GAN: interactive point-based manipulation on the generative image manifold

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.994288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.741274Z digest=sha256:b6984a4ba2905d66b78fd97b6b23e300e6b73c31fcfa95917b830f46c48cc6be

Observation e23bb6c6-2188-407b-89b3-ffa53ff711c9 · outbound

This paper cites Scalable diffusion models with transformers.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scalable diffusion models with transformers

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.973689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.747315Z digest=sha256:a3dfe151a5d867135790430d983fdeca54b0e4002a7a559afd9120723aec9791

Observation aa52c4ff-ddd0-4596-ac1b-3b74d8be1610 · outbound

This paper cites Grounded text-to-image synthesis with attention refocusing.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Grounded text-to-image synthesis with attention refocusing

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.956466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.754972Z digest=sha256:0d0a9f90b2cb3d830a8242b02a58047fb6a029eea827c92d32996f3278fbdc27

Observation e01d5566-530d-4474-a642-06b447fecaa6 · outbound

This paper cites SDXL: improving latent diffusion models for high-resolution image synthesis.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models SDXL: improving latent diffusion models for high-resolution image synthesis

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.934740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.761651Z digest=sha256:63e22f4e69518edd2b56873fc4dfff9e1c891144fb8eafe412b98c7001b54bb9

Observation 7080e902-5def-4ce3-a9ca-78d6b19aec07 · outbound

This paper cites Barron, and Ben Milden- hall.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Barron, and Ben Milden- hall

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.916358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.767976Z digest=sha256:d3905eb9029318e20e8eed50cb25239db563369d31222d14ec013e0f089f5515

Observation 0d0c5842-a78b-4d3c-850b-4fdda161ffb6 · outbound

This paper cites Learning transferable visual models from natural language supervision.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Learning transferable visual models from natural language supervision

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.773083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.773083Z digest=sha256:eea048d82a9d80e47b733f0d2cee705316030ab6e5b55dca00c6ba1ac2a52e80

Observation dd5e34d9-b678-4588-b97c-a933d210b9d9 · outbound

This paper cites an unresolved cited work.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:24:11.886403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.778341Z digest=sha256:336c70206a0bdbc54c11f4330e06f4b23cc3a8871c9a19df5fd476a70e840f72

Observation 2984c77a-f6cf-4c2e-8e3e-3e6545ad89cb · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.783583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.783583Z digest=sha256:9430cdb6b672ea727459f7c60f2714767d355e2b1d492db8d7a7228d2cb6e390

Observation 7687f769-764f-4579-9e80-72d47737ed45 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models High-resolution image syn- thesis with latent diffusion models

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.868151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.789097Z digest=sha256:602c11fbd4bd80be6fdf3523043fa4c0ea4ef6c664e78155351573e0520fa8bd

Observation b7e4b89f-92ce-4ee5-bffa-374d6ac74fd1 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.794549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.794549Z digest=sha256:a28705658ef462d35a69b3bcc1740a4dd4f870cd0523650078adb019a0fac115

Observation 10611063-b259-48da-8249-1026876cdb59 · outbound

This paper cites Runway AI.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Runway AI

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.835898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.800035Z digest=sha256:bb2a3f0b1441a12ccc6d41124aea4b5545cfe6dcff9a614a088aa6916d279137

Observation c9594d61-30ab-482b-a943-4fca15c1330b · outbound

This paper cites Dual caption preference optimization for diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Dual caption preference optimization for diffusion models

Reference 61

Resolution
verified exact
raw_fallback, observed 2026-08-11T13:24:11.309326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.805071Z digest=sha256:194d5a44cc6fdeac2da66e7ce0f5eac485eeebf85767e4da35c4cd8d8be446d6

Observation 49d2fb5e-11e9-428b-b0fb-1e47a34d58a5 · outbound

This paper cites Denton, Seyed Kamyar Seyed Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, Jonathan Ho, David J.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Denton, Seyed Kamyar Seyed Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, Jonathan Ho, David J

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.816335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.810558Z digest=sha256:141eb2d411b5dfd30c265e873550726dcef22d3a2f180e551ea727a420b66bc9

Observation 638370b8-2bd4-4f66-8c89-162c47bd72d5 · outbound

This paper cites Scribbler: Controlling deep image synthesis with sketch and color.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scribbler: Controlling deep image synthesis with sketch and color

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.799628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.816766Z digest=sha256:dbad46bc354e479d8b688b2a61b4d0ef024a7e8aa7a827fecac5766e8a1a8271

Observation 445fbd8f-510d-44c1-b4ed-c9510c3693fd · outbound

This paper cites LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.822470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.822470Z digest=sha256:d84bc46dc2c5b67b596f05b68965f8bdc5d315d45c22e89325c5e457d4d0d372

Observation 37a46d67-7e22-41ae-ba4c-e2feb7305527 · outbound

This paper cites LAION-5B: an open large-scale dataset for training next generation image-text models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models LAION-5B: an open large-scale dataset for training next generation image-text models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.781666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.828253Z digest=sha256:84216761ff54fb8e620331cbfa8298adf5c90494ea6980f47e787a044c5b7622

Observation ab00d0d8-669d-4b10-8593-ba5be5ffd414 · outbound

This paper cites A Picture is Worth a Thousand Words: Principled Recaptioning Improves Image Generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models A Picture is Worth a Thousand Words: Principled Recaptioning Improves Image Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.841427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.841427Z digest=sha256:bf2c8d1aad177928c2d3a06347c1e808470dd4cefa698b430685e07798a6d004

Observation 48833b0b-4ae8-44d3-bfc6-8257ab996257 · outbound

This paper cites Box It to Bind It: Unified Layout Control and Attribute Binding in T2I Diffusion Models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Box It to Bind It: Unified Layout Control and Attribute Binding in T2I Diffusion Models

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:24:11.115958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.856007Z digest=sha256:3b4373f3222d6899671c4b8afe264ee6bcf3465754deed97b709ef084f043794

Observation c4077533-ba0d-4185-9691-e1d9fb5b37e2 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.763704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.862045Z digest=sha256:0ed7f3500c646fbd46ec7fd7f07aebc888658012f28bf3e8a6a4185baa975f13

Observation 2af58fe0-a2bb-49ff-a225-1ebb4ae18b1e · outbound

This paper cites Diffusion model align- ment using direct preference optimization.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Diffusion model align- ment using direct preference optimization

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.744672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.867393Z digest=sha256:f3f612f2d1fe6d6aee4c664a023f0c49201c979f53f94a6210d56ea707ad8bc8

Observation c18779da-4b1e-4b41-8dc5-041ad58897d8 · outbound

This paper cites InstantID: Zero-shot Identity-Preserving Generation in Seconds.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models InstantID: Zero-shot Identity-Preserving Generation in Seconds

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.872998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.872998Z digest=sha256:2ee872264ba4c2fc9df03578007ea8447e9158ed07d10ce3bb2990c5382d4b85

Observation b06c5731-26e4-485c-b211-79679011c61a · outbound

This paper cites Tokencompose: Text-to-image diffusion with token-level supervision.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Tokencompose: Text-to-image diffusion with token-level supervision

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.726553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.879033Z digest=sha256:ce7f8cedf20330c8b474a5242ce6ab9402441647c8edcc387ce20888338245d8

Observation 86825bcd-31b4-4bb2-bab5-29cee8c5d78e · outbound

This paper cites Wang, Evan Montoya, David Munechika, Haoyang Yang, Benjamin Hoover, and Duen Horng Chau.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Wang, Evan Montoya, David Munechika, Haoyang Yang, Benjamin Hoover, and Duen Horng Chau

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.707263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.885079Z digest=sha256:bf4a089f7caf2efae5cb8908cb7a0f2d1cb50060fafed0042d6866a160f425af

Observation 38b1eb26-daf8-4186-8483-866e0cc81d01 · outbound

This paper cites Seesr: Towards semantics-aware real-world image super-resolution.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Seesr: Towards semantics-aware real-world image super-resolution

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.683919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.892075Z digest=sha256:95e0c4af7ca8e247a2b6c50eceff33e3410303ec8fc4c4fe4af17272d0567f70

Observation 9e03bc59-2d8d-4e12-9b38-77f82f6d8d2a · outbound

This paper cites Paragraph-to-Image Generation with Information-Enriched Diffusion Model.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Paragraph-to-Image Generation with Information-Enriched Diffusion Model

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.898243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.898243Z digest=sha256:fc3a41930ff74640c39aeba0612e633c8ea7f96c84fed72065f15813b66114e0

Observation 2252543c-929d-4168-9b12-723ad639e611 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.904255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.904255Z digest=sha256:a29b4eb529cbac934b3934f48bfa118d1c4e40ff3a524f6be9292cd59fc1e11e

Observation 4a59abcc-7d48-4683-af46-379e42f8f5b9 · outbound

This paper cites Human preference score: Better aligning text-to- image models with human preference.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Human preference score: Better aligning text-to- image models with human preference

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.658991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.912130Z digest=sha256:ae144188290162bf3acc53c4b8b47bd54b838c137d90dca096300687e94be6e2

Observation 83b53798-2f3a-4c30-8aee-6fc1295ada32 · outbound

This paper cites Stylespace analysis: Disentangled controls for stylegan image genera- tion.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Stylespace analysis: Disentangled controls for stylegan image genera- tion

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.637957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.923913Z digest=sha256:767481f38a82ed4dd5914189ff9d3a1dd07d8fdaf2529d6f4829a3a2ec88c1e7

Observation afce0546-38c0-430b-bb9f-d0089fb09d6a · outbound

This paper cites Freeman, Fr ´edo Durand, and Song Han.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Freeman, Fr ´edo Durand, and Song Han

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.617144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.932289Z digest=sha256:a07c2cb9f5ecdf18d47c49ead90840dc0076d4e881276838b270c4d1e24263e6

Observation d8b95086-b201-40a2-b605-d57702f4fcad · outbound

This paper cites R&b: Region and boundary aware zero-shot grounded text-to-image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models R&b: Region and boundary aware zero-shot grounded text-to-image generation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.598441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.938638Z digest=sha256:e8e56c8ffd0a7f54bcc3fa17b1ee826d57e2ebcaf06f096f86df5f3ac3baf505

Observation f5fb7fd4-67d3-4ae2-9623-e723dd1b9320 · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.944286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.944286Z digest=sha256:8778f27b8034685ed164d7e98072fa3f7058e2982fac7ec1c365270ca13e1765

Observation fc44424b-9431-4c8a-952b-f6c06be816fd · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.951063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.951063Z digest=sha256:27192a7c3981df7dea99376b62a8d29220ff6b1c74e828c5c7742ce5c9cc7797

Observation cd07df9d-f452-454e-bd4e-39cf1b910cbb · outbound

This paper cites Scaling up to excellence: Practicing model scaling for photo- realistic image restoration in the wild.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scaling up to excellence: Practicing model scaling for photo- realistic image restoration in the wild

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.563663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.957503Z digest=sha256:fd5c2ed17d832197c9906788334b7fbab5fda5a88ada50876543b2679c5c9f8e

Observation b441700e-f494-4f96-b689-c2cb32f0e77e · outbound

This paper cites Scaling autoregressive models for content-rich text-to-image generation.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Scaling autoregressive models for content-rich text-to-image generation

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.541812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.962644Z digest=sha256:ab5f503d86c0302d7aff9666b19639b4a6887dc19ca2194b03840c1ab3e6bdec

Observation 21779ad1-bf30-4bcf-b2da-782c78f30706 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T13:24:10.968045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:24:10.968045Z digest=sha256:3379eb3cebe4a97c84b11f5239df5557f86ca09947db0db0b40d9fe187aaf5fc

Observation 9835b348-4bca-4b18-80fa-b8d68909d851 · outbound

This paper cites A horse to the left of a bottle.

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models A horse to the left of a bottle

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:24:11.509598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T13:24:10.973352Z digest=sha256:31b53964cff05fde41b4cebcc67cb61b39540b582cde38cbb9f766576e91b35f

Pith citing papers

No inbound Pith citation observations are available.