Pith. sign in

Paper Citation Record · LEDGER

VSC: Visual Search Compositional Text-to-Image Diffusion Model

As of 21 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2505.01104.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01104 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:32:37.842627Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8940cc4c-a0d3-4fdb-b82b-557647e0941a · outbound

This paper cites A-star: Test-time attention segregation and retention for text- to-image synthesis.

VSC: Visual Search Compositional Text-to-Image Diffusion Model A-star: Test-time attention segregation and retention for text- to-image synthesis

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.241280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.706662Z digest=sha256:046ac4e21e37ef01f3343a612a28c00340e877184a8c4a2f2083ae2590bccaf6

Observation 40f90421-1480-41d0-9504-1c7d6e5060ac · outbound

This paper cites Spatext: Spatio-textual representation for con- trollable image generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Spatext: Spatio-textual representation for con- trollable image generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.229245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.711004Z digest=sha256:c538b926c51e909e2f676c1ca99c701b478c34fa49907a7f2820e4da4336aa01

Observation e0b7dfa2-60d5-4e53-a0ce-66d4141327e5 · outbound

This paper cites Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.714812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.714812Z digest=sha256:abe15bec50c18a3cdac9293d25cb2cd790e46495670f59928e568c5a483b120b

Observation d832ea00-2b9a-40b3-bde5-3b59a0b73948 · outbound

This paper cites Attend-and-excite: Attention-based semantic guid- ance for text-to-image diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Attend-and-excite: Attention-based semantic guid- ance for text-to-image diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.218008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.719179Z digest=sha256:ef2cad5d0bdb509e310691248d28c9af538952c16714ca496fff284c9affa344

Observation e8d53753-aff8-4545-a0f2-ca88086d14b7 · outbound

This paper cites Schwing, Alexan- der Kirillov, and Rohit Girdhar.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Schwing, Alexan- der Kirillov, and Rohit Girdhar

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.207142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.722873Z digest=sha256:50213725c4a8c3f31d565224d8e20f17ec1e4d119a5bb7f0ff1669873f6c5eca

Observation 483d8d9b-178f-4a4d-91b9-4447ffee0951 · outbound

This paper cites Aspects of the Theory of Syntax.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Aspects of the Theory of Syntax

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.195479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.726658Z digest=sha256:bb8062e766ad8cecb42be49beb690f6de37d001c8dcaecd49050453a0cb96bd6

Observation c9bb5b28-0d28-45dc-97ee-b863c4ba3028 · outbound

This paper cites Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.730423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.730423Z digest=sha256:b03e9907417204f7c4a73c21e3885a73ea39b621d9b36b7776c158a217f5527f

Observation 78d82129-7784-4850-8598-c8833d7669c2 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Scaling rectified flow trans- formers for high-resolution image synthesis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.183999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.734752Z digest=sha256:4a07e3d869c97f6e0a5192d13d9a51443c4604d5d9696b43aba04c761f027290

Observation ce26b3e5-917a-49be-9563-a3ed55ade6c9 · outbound

This paper cites Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.738628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.738628Z digest=sha256:b4a9a9471e9bdc9394b54e1dc7d69c89ae4f66c56b2189705ec444e35f85d880

Observation fb3419f3-042e-46b5-88bb-7ed76213febd · outbound

This paper cites Layoutgpt: Compositional visual plan- ning and generation with large language models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Layoutgpt: Compositional visual plan- ning and generation with large language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.742553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.742553Z digest=sha256:1f189237af5e365a2c4e3e30aaf29d5552ed53210dd3f20327df3ca439f9f094

Observation 2217b315-d14f-4d4e-9673-840da372ba0e · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

VSC: Visual Search Compositional Text-to-Image Diffusion Model An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.746049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.746049Z digest=sha256:5268d6c821109facf5b6574f56ccbbf2eb03821343c358e2ccf0d2f48dd81c3c

Observation e15fdd26-5ed9-4661-bea6-fd7ccb957a32 · outbound

This paper cites Encoder-based domain tuning for fast personalization of text-to-image models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Encoder-based domain tuning for fast personalization of text-to-image models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.164486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.750541Z digest=sha256:35441d5f337015400c69c4121e691571a1a7467e66a45417bb4248cf193139b6

Observation 0a187407-5d32-4ea9-af7e-70fbac53d3d0 · outbound

This paper cites T2i-compbench: A comprehensive benchmark for open- world compositional text-to-image generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model T2i-compbench: A comprehensive benchmark for open- world compositional text-to-image generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.754114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.754114Z digest=sha256:09d77c36670dc570cc925e129c69b18d337ded8888d8007eabc3028bcf2c0a61

Observation 9fdf6f7e-e5d1-4629-a575-6631904d0fbc · outbound

This paper cites Openclip, 2021.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Openclip, 2021

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.145998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.758118Z digest=sha256:50848f98980293ec9c9af957649fdc5d4ec5f5a3282929a467cf7c3dcee6e577

Observation fc05b7c0-5f3c-47e6-a232-3c9c525e0c48 · outbound

This paper cites X&Fuse: Fusing Visual Information in Text-to-Image Generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model X&Fuse: Fusing Visual Information in Text-to-Image Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:32:37.924264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.762395Z digest=sha256:7625e71740f6dffd79de1ca606ea0be54fc9aaf5da632d112a18fb70a9be7172

Observation 475967bc-e754-42a9-b8cb-ec6b382ef0a8 · outbound

This paper cites Multi-concept customization of text- to-image diffusion.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Multi-concept customization of text- to-image diffusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.766454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.766454Z digest=sha256:04eafbc94914dc5c1b8eaf4a01d56033ccd28fa294a3633fa5026ef91aa7cf66

Observation 27cc0c02-f10f-4d5c-9aac-99aecd694353 · outbound

This paper cites Does CLIP Bind Concepts? Probing Compositionality in Large Image Models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Does CLIP Bind Concepts? Probing Compositionality in Large Image Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.770193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.770193Z digest=sha256:25ef396a53edde7f5253be55013a06bc56338ef243780857b18f099fdfcd0d95

Observation eba71ed0-9fec-446b-8835-8cfc7bbad605 · outbound

This paper cites Compositional visual generation with composable diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Compositional visual generation with composable diffusion models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.127426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.774118Z digest=sha256:cdcfde5652d4a222e033beedaf90d078bb7e04ae8e13a39f26a40dfedf47d162

Observation 246ef3c6-f8aa-415c-95d5-4755765e9599 · outbound

This paper cites Conform: Contrast is all you need for high- fidelity text-to-image diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Conform: Contrast is all you need for high- fidelity text-to-image diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.115080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.777873Z digest=sha256:b4ca6183e628f0c64628af6d936a93da731142989b657e3e9e4373e8d346efaf

Observation 3effc674-5468-43ba-b74a-f63f11fe3166 · outbound

This paper cites Null-text inversion for editing real im- ages using guided diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Null-text inversion for editing real im- ages using guided diffusion models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.781845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.781845Z digest=sha256:da18c6c8fe4b6af7f5f625fc6d41ee62a047c754793fdff54375f3b7842fee89

Observation bfdfa16b-5383-4db0-8855-825f07bbc94d · outbound

This paper cites Compositional Text-to-Image Generation with Dense Blob Representations.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Compositional Text-to-Image Generation with Dense Blob Representations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.785519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.785519Z digest=sha256:2589ecc6e10ed744537b1f42751b5bfce4ed505dd2fa68811f4d083015b136ae

Observation 593e2bfe-744e-49ab-b89f-0c891c861d1c · outbound

This paper cites Compositional abilities emerge multiplicatively: Ex- ploring diffusion models on a synthetic task.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Compositional abilities emerge multiplicatively: Ex- ploring diffusion models on a synthetic task

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.095525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.789348Z digest=sha256:6154ac5a26d9e8c1c65739f137b05ce1bcd2d3bd42f689ef1c590254c387a9b6

Observation 60bc3daf-34c9-453c-99e6-2b33424ecf8a · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Learning transferable visual models from natural language supervi- sion

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.792906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.792906Z digest=sha256:9ccae3f1c21d44d3dd1f838f08d2a274c669153ab1432804f44e07b63fb90d84

Observation 80927ef5-59b0-4991-8688-f531ccadf660 · outbound

This paper cites Linguistic binding in dif- fusion models: Enhancing attribute correspondence through attention map alignment.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Linguistic binding in dif- fusion models: Enhancing attribute correspondence through attention map alignment

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.074977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.796732Z digest=sha256:d498dd5d00fb0d5c29afa8b5f2a01c7b225b3c227e2b3babf14bef038294814c

Observation b666c1f3-c711-4fda-aaa9-29d28451b183 · outbound

This paper cites High-resolution image 9 synthesis with latent diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model High-resolution image 9 synthesis with latent diffusion models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.063720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.800192Z digest=sha256:d19079463749fe337590c01493a697c8d16de4022dfc10424af80d17223ee42f

Observation a3633a04-5f50-4a87-b22d-b3a43a6cff16 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven gen- eration.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Dreambooth: Fine tuning text-to-image diffusion models for subject-driven gen- eration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.803812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.803812Z digest=sha256:a8f6bd14c53cdcc5c7c7d6dd5567872fa231891ef9cb9c2f2655c532d20df0f2

Observation 2180b882-8dc7-4b3e-8605-5a8c3baf1bd8 · outbound

This paper cites Collage diffusion.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Collage diffusion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.046303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.807535Z digest=sha256:2518ddf6746a497b7451b7d16c086b520f1f6d88c7fbb563c3d9c6447d6518af

Observation 9673ca5c-6ba6-4a31-9fc1-0a6269c44097 · outbound

This paper cites In- stantbooth: Personalized text-to-image generation without test-time finetuning.

VSC: Visual Search Compositional Text-to-Image Diffusion Model In- stantbooth: Personalized text-to-image generation without test-time finetuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.811840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.811840Z digest=sha256:9a9dcd1534058b6bed90ad194a9112e77cb2ae012ebe5a7403899313191b8606

Observation 25ae24a8-e593-4180-a8e5-ab39b7a7fa8c · outbound

This paper cites Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.815468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.815468Z digest=sha256:ff9ca2ef2c17a2fee9060e8905ccca2b5924ac9447a8de0d7e11ef0389ea0d50

Observation 44ce8b81-3659-4245-bbe4-75c113bf497e · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.819233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.819233Z digest=sha256:458226b72551d33ac193b9a4244a34e6afc2b6c773e93a12a9ff2cf14417fd6e

Observation 8a9d1631-6d7b-4ee4-8f20-e7560056296a · outbound

This paper cites A feature-integration theory of attention.

VSC: Visual Search Compositional Text-to-Image Diffusion Model A feature-integration theory of attention

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.021965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.822804Z digest=sha256:fa69537ada3193e57a4a1ae06a587439d73eb16fb380a49056883cbd3eb5c6ef

Observation a69a5f3f-ef06-47e8-84a7-f895f0dfb716 · outbound

This paper cites Compositional text-to-image synthesis with attention map control of diffusion models.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Compositional text-to-image synthesis with attention map control of diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:38.010243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.826762Z digest=sha256:33c417a4e8048bbeca167ae04ba65e5c1d27ccd1f20ac5877481e2454c8f5e7d

Observation 3a6d0582-4e8f-4dc2-af51-7b889b5ccbcb · outbound

This paper cites Elite: Encoding visual concepts into textual embeddings for customized text-to-image generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Elite: Encoding visual concepts into textual embeddings for customized text-to-image generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:37.997285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.830414Z digest=sha256:34040848d45928ab5fd0af19f0d83eb7f872838904c9fc3c095b91525378e0c7

Observation 77391611-dd56-4e12-8280-20fc649ef992 · outbound

This paper cites Fastcomposer: Tuning-free multi- subject image generation with localized attention.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Fastcomposer: Tuning-free multi- subject image generation with localized attention

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.834554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.834554Z digest=sha256:3e3bd94b98d0aaa3e1187799fcdfd7adff997622564663d4d22a32c2b8c76022

Observation bd17cd23-2685-4f28-83f3-2800da2cfaaa · outbound

This paper cites Improving Compositional Attribute Binding in Text-to-Image Generative Models via Enhanced Text Embeddings.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Improving Compositional Attribute Binding in Text-to-Image Generative Models via Enhanced Text Embeddings

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:32:37.838448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:32:37.838448Z digest=sha256:61828a2e2ed507585bfe410766c1c512c5e726066e7af948ef5762ab26f827bb

Observation dcafdac5-7e56-46c7-beb9-7d52ccc63e47 · outbound

This paper cites Layoutdiffusion: Controllable diffusion model for layout-to-image generation.

VSC: Visual Search Compositional Text-to-Image Diffusion Model Layoutdiffusion: Controllable diffusion model for layout-to-image generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:32:37.978537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T04:32:37.842627Z digest=sha256:9336667913009667b3ef4a19f743d90097b57a10ecf17149e4f60cd926221ae7

Pith citing papers

No inbound Pith citation observations are available.