Pith. sign in

Paper Citation Record · LEDGER

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation

As of 7 August 2026, this Paper Citation Record lists 100 of 116 outbound references and 0 inbound Pith citation observations for arXiv:2507.02792.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02792 v5

Coverage vector

measured 100 of 116 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:24:55.510895Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 116 outbound references displayed

  • verified exact1
  • verified fuzzy56
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e266709e-355e-45ea-b8dc-3d50e8ec0c04 · outbound

This paper cites GPT-4 Technical Report.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:49.497593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:49.497593Z digest=sha256:92d1c7b4125a01c05959d8a7434a7a0091dd972f4e44306721298754ab6d68b6

Observation c1329692-d50b-4e13-a0cf-1f83fb474ee8 · outbound

This paper cites Cross-image attention for zero- shot appearance transfer.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Cross-image attention for zero- shot appearance transfer

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:49.607298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:49.607298Z digest=sha256:69d59db747528d86155135df9ec7a98913ab8c7d58e26942e3a304d346b84efb

Observation 485a6cdb-665a-42f1-a0c5-6030b9be67c9 · outbound

This paper cites Break-a-scene: Extracting multiple concepts from a single image.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Break-a-scene: Extracting multiple concepts from a single image

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:49.757789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:49.757789Z digest=sha256:befd5567d7e7036e3fa3dc6cdaddb3a6e33ae1aca763b633e29a2d037bb77963

Observation b819adf7-286a-4a83-9a3c-abdf38596b1a · outbound

This paper cites Spatext: Spatio-textual representation for con- trollable image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Spatext: Spatio-textual representation for con- trollable image generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:49.872123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:49.872123Z digest=sha256:5f7d25f37382b82f48604e341fb2c1023e0ab4fc2a9f3199beb42f65886bf514

Observation b118e7f7-8aee-428d-8afb-e2fde7339934 · outbound

This paper cites Stable flow: Vital layers for training-free image editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Stable flow: Vital layers for training-free image editing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:49.996971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:49.996971Z digest=sha256:a0c3d3987e0fd36cad763f1348f26ae5a8cf1705b0e1d8b8297b8fdfd15c1ed1

Observation 055b30d6-7348-4fe1-a69c-def567508852 · outbound

This paper cites Universal guidance for diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Universal guidance for diffusion models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.115532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.115532Z digest=sha256:c3b5c0225f707f50ed4f71862c417f5c356bf0c2d9205488c8e90ba1bf906358

Observation 1f1c8c7c-d4c0-46ea-adec-9eed02116cbe · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation In- structpix2pix: Learning to follow image editing instructions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.240522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.240522Z digest=sha256:1789d992de6818d1cc697e546ec45556d5ee2c62ed00a06cba30e7667e4aa636

Observation 5bfa0b53-e0bb-4826-b429-e5f657d84ff9 · outbound

This paper cites Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.349140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.349140Z digest=sha256:e57de9410b72e62e21a74fe5baad6913658f99ace4f1a39066d7739ae52cfb27

Observation 3923c41a-307f-4c86-ae6b-b195fe61ff3e · outbound

This paper cites an unresolved cited work.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.481192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.481192Z digest=sha256:6d5f058ace961e5619c917654c12884a45a2d841e8be7cccabcfdbe578fd1e91

Observation 4220aeca-190d-4366-a172-67ef3fbafbe4 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Emerg- ing properties in self-supervised vision transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.583071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.583071Z digest=sha256:d421c3bea9f235e233e3fb66eaaf27f9298aa53e463114a6b1c8aeaf3bb640a8

Observation 055474b9-420c-4920-9826-d8f7ff26d599 · outbound

This paper cites Training-free layout control with cross-attention guidance.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Training-free layout control with cross-attention guidance

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.713859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.713859Z digest=sha256:3a9f1c8d54dc6d2c0550454bcc43fd0b4303ad541dd3528293c7c89b2979517e

Observation 847263c7-9b2d-4a62-8291-c8557a096088 · outbound

This paper cites Unireal: Universal image generation and editing via learn- ing real-world dynamics.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unireal: Universal image generation and editing via learn- ing real-world dynamics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.814774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.814774Z digest=sha256:de575376d82c0e2d44dd9d63ba3bf028df5ba5a7034b96b49c330659068a4d14

Observation de956882-21f6-4208-906a-6f4e5e8a8827 · outbound

This paper cites Style in- jection in diffusion: A training-free approach for adapting large-scale diffusion models for style transfer.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Style in- jection in diffusion: A training-free approach for adapting large-scale diffusion models for style transfer

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:50.913306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:50.913306Z digest=sha256:8c68c0e3f458acd34c280530d3581b1690fce3c8ea070443a7c1e163a8dc7fc8

Observation 0368fc1c-f792-45bb-a362-91658d163183 · outbound

This paper cites Diffedit: Diffusion-based semantic image editing with mask guidance.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Diffedit: Diffusion-based semantic image editing with mask guidance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.007576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.007576Z digest=sha256:78846366800e8e00e51f27cdea11f9e7f3c2f9a9e40eba32ea1b30790f2784d2

Observation 17f87e2d-1548-48fe-af4c-648d84336dec · outbound

This paper cites Fluxs- pace: Disentangled semantic editing in rectified flow trans- formers, 2024.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Fluxs- pace: Disentangled semantic editing in rectified flow trans- formers, 2024

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.106831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.106831Z digest=sha256:1c51fe1acb338e6c56aba32c8d13bac6ce1657bca0f967406bb1970ac79de2e7

Observation 533a0ae2-12b7-4073-8309-e4f94d059899 · outbound

This paper cites Freecustom: Tuning- free customized image generation for multi-concept compo- sition.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Freecustom: Tuning- free customized image generation for multi-concept compo- sition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.228684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.228684Z digest=sha256:afd2bd911cf230236b99bbd85593d6dfa668fea1e2addd31038f67396ecf485f

Observation e62d770b-8e43-4065-a31b-013a6ccf51a5 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation An image is worth 16x16 words: Transformers for image recognition at scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.367662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.367662Z digest=sha256:0d45decae31233710706f983711bb000efb221911d9c2b96d6e7410e3bf4e475

Observation c374d79d-171e-41fc-94c7-0f961b4497d4 · outbound

This paper cites Efros, and Aleksander Holynski.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Efros, and Aleksander Holynski

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.463412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.463412Z digest=sha256:1e708560f209573f7908f7cbd30cf6f0536b525a5bf7e17891ebefaf0c765d67

Observation b9bd1199-f3ce-4610-8ef3-f314c34b6883 · outbound

This paper cites Scaling rec- tified flow transformers for high-resolution image synthesis.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Scaling rec- tified flow transformers for high-resolution image synthesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.604238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.604238Z digest=sha256:8c4397970d06baf47a969516918aedfb937fbd7458e26c8e172bade90bb242dd

Observation 220c09bf-79dc-43bc-b604-5052a5e62ae2 · outbound

This paper cites Personalize Anything for Free with Diffusion Transformer.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Personalize Anything for Free with Diffusion Transformer

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:24:55.670646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:51.690968Z digest=sha256:603dc5a2825ae4dc71422ac2b1ee4d692a7194dc6f1e20988189e0ccdceb8ea4

Observation e3bf6a9d-8b53-47ec-8d19-d235ae651301 · outbound

This paper cites Dit4edit: Diffusion transformer for image editing.Proceedings of the AAAI Conference on Artificial Intelligence, 39(3):2969– 2977, 2025.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dit4edit: Diffusion transformer for image editing.Proceedings of the AAAI Conference on Artificial Intelligence, 39(3):2969– 2977, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.786540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.786540Z digest=sha256:6b72a4ed53313d898fcaf6bce60efdb6f89f408db7b2ad9868b4c9aba1d208d7

Observation 59eddcce-5385-4164-aa45-28cdd65f634e · outbound

This paper cites Dream- sim: Learning new dimensions of human visual similarity using synthetic data.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dream- sim: Learning new dimensions of human visual similarity using synthetic data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.895105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.895105Z digest=sha256:149b35aafd125ba6b68bed9b629b2586cf5a9bbdbb8787d8954c218062e19971

Observation dd5cc86a-a996-4b63-988f-6390265380f7 · outbound

This paper cites An image is worth one word: Personalizing text-to-image gener- ation using textual inversion.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation An image is worth one word: Personalizing text-to-image gener- ation using textual inversion

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:51.966682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:51.966682Z digest=sha256:2164dd168fd1adc6762526efe91d52bfd0a7413001cb0a89f06997cb5c3d10f9

Observation 74a73dd8-7a0e-4b3a-9595-ad86050dfd48 · outbound

This paper cites Instructdiffusion: A generalist modeling interface for vision tasks.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Instructdiffusion: A generalist modeling interface for vision tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.053278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.053278Z digest=sha256:408f25cd446d1335cd174cb4f5fb4d0dde7d862cfc5a6b096af9f80aa7c00e6a

Observation 7b2706ca-8c6b-476a-ab0c-20984ababe23 · outbound

This paper cites Eye-for-an-eye: Appearance transfer with semantic correspondence in diffusion models, 2024.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Eye-for-an-eye: Appearance transfer with semantic correspondence in diffusion models, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.115912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.115912Z digest=sha256:592724bde53d8f8fcae6580d3c470fd69227294e346ad44f17a8a6f6a9e8e431

Observation 8131ffcc-72c9-4787-9074-53554c20a5ae · outbound

This paper cites Pair diffusion: A comprehensive multimodal object-level image editor.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Pair diffusion: A comprehensive multimodal object-level image editor

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.230225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.230225Z digest=sha256:82151805768b987b7f07bfc1fbb6f6598c862d51db66501facfce8b19921898f

Observation a75d61d2-ab5c-4f72-98ac-b48999b4e8a8 · outbound

This paper cites Ace: All-round creator and editor following instructions via diffu- sion transformer, 2024.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Ace: All-round creator and editor following instructions via diffu- sion transformer, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.346305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.346305Z digest=sha256:4bd03fc65beab7ca0f85a904ffb8bc6809a8df02efdee61a5b8a3e42a9e00e7d

Observation 5e8552ed-c88b-486c-a147-3b3df9138c98 · outbound

This paper cites Prompt-to-prompt image editing with cross attention control.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Prompt-to-prompt image editing with cross attention control

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.446284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.446284Z digest=sha256:3941f623b30abbd76fdc5db72a342f0bc797ecd0045cafc1f2893ca9e3bdfd5f

Observation 2935d79f-702f-4c04-9f23-097ffb9e1623 · outbound

This paper cites Classifier-free diffusion guidance.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Classifier-free diffusion guidance

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.586230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.586230Z digest=sha256:15c7aa3ee22e2b9cbb641555937090fab85c0f02680ae383181cf44c75668469

Observation 9c4ba7a9-91be-4c4b-90c3-8672b09f3cef · outbound

This paper cites Denoising dif- fusion probabilistic models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Denoising dif- fusion probabilistic models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.720503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.720503Z digest=sha256:320639a88463d98eafb4b96d021f4e9f478cdfbd8a53dcf93d65a7ebf31acecf

Observation 91f10378-36e5-41da-a7b9-dcbc154e9264 · outbound

This paper cites Cascaded diffu- sion models for high fidelity image generation.Journal of Machine Learning Research (JMLR), 2022.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Cascaded diffu- sion models for high fidelity image generation.Journal of Machine Learning Research (JMLR), 2022

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.846727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.846727Z digest=sha256:846fcbf7e61df498b9bcab318b24424da41243239417b6104bf2ddbb5d4e61de

Observation 7014afd6-c2ba-42d9-b7ab-b9ff4e827e3e · outbound

This paper cites Anchor token matching: Im- plicit structure locking for training-free ar image editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Anchor token matching: Im- plicit structure locking for training-free ar image editing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:52.968113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:52.968113Z digest=sha256:8ac2a41061c2a024871c78033acae48936ad18d92e43d4ac50a58eb63f4c4b92

Observation 4f5f8f89-791b-4a2f-8469-4841e2776558 · outbound

This paper cites Attenst: A training-free attention-driven style trans- fer framework with pre-trained diffusion models, 2025.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Attenst: A training-free attention-driven style trans- fer framework with pre-trained diffusion models, 2025

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.107448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.107448Z digest=sha256:b816bb3a1026b17254caf5d4a89ff153a52801a11763ed8b5c4f98ea24c6fbd8

Observation d417ddfe-01a1-46a2-bd0b-401645070aca · outbound

This paper cites Image-to-image translation with conditional adver- sarial networks.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Image-to-image translation with conditional adver- sarial networks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.242525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.242525Z digest=sha256:88e97c6a860dfb35cdd41dad4b5c48a0e780bdfc5c2a7cef736a394dae4ea000

Observation f4f2f04b-4e9d-4f60-834d-f2445ee30fc8 · outbound

This paper cites an unresolved cited work.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.392055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.392055Z digest=sha256:b1850877d8c76bb59dc4bcf2c056d5b246bfca803260536acda43c40df0c63ec

Observation 9dd0bda8-899a-44b4-b418-c4e48a7c5bb5 · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Elucidating the design space of diffusion-based generative models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.500171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.500171Z digest=sha256:ffad9ce59facfc62a5f69681158fb2fa052d5039d4bc3a0fed5bf9e3f23c8a48

Observation 61aa3e54-646c-4d8a-bf29-25922a5651a3 · outbound

This paper cites Dif- fusionclip: Text-guided diffusion models for robust image manipulation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dif- fusionclip: Text-guided diffusion models for robust image manipulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.620035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.620035Z digest=sha256:820b1c708fdb65d357aa84bd723d5df785e27ec975f35edd3431341687ba6214

Observation b88d9e29-608c-4d98-97db-77b32fa67cd2 · outbound

This paper cites Dense text-to-image generation with attention modulation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dense text-to-image generation with attention modulation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:53.722395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:53.722395Z digest=sha256:5727599680dc1d95527092a976c987826678752dd4146905dc1f26a056d38d01

Observation 48a8d58d-4137-4c06-a0ed-d044a1fa04f2 · outbound

This paper cites Diffusion-based image translation using disentangled style and content representa- tion.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Diffusion-based image translation using disentangled style and content representa- tion

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.621882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:53.871092Z digest=sha256:5816995d682ef70f046a34e8946ba1be0c24dbbfc6b8a57dbc9156d8c2ea6f9d

Observation 9e0c75a1-a7ea-4ddf-85f7-b35930176e25 · outbound

This paper cites an unresolved cited work.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:56.610261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:53.989615Z digest=sha256:4d08d983d1a92715c47865809e9e4db6fe568e4e07e51f35894a265d2c9a033c

Observation 81b9b68e-b717-4d50-a74a-37720962dc1d · outbound

This paper cites Flux.1 kontext: Flow matching for in-context image generation and editing in latent space,.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Flux.1 kontext: Flow matching for in-context image generation and editing in latent space,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.598062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.120312Z digest=sha256:47571b823cf78c1b35f5c15070f679db7aaebb5b4f10a29688e043949f8f15d3

Observation f516df46-1b36-4140-91cb-31377f0446fe · outbound

This paper cites Le, Tuan Pham, Sangho Lee, Christopher Clark, Aniruddha Kembhavi, Stephan Mandt, Ranjay Krishna, and Jiasen Lu.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Le, Tuan Pham, Sangho Lee, Christopher Clark, Aniruddha Kembhavi, Stephan Mandt, Ranjay Krishna, and Jiasen Lu

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.586454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.272804Z digest=sha256:e869f32b6ada550830e2feaf06a3baae647f154cdfc3d7f0ad5ce1b9b507f555

Observation 13fb5b3a-65d1-4568-ade7-d84ab7f10f58 · outbound

This paper cites Scribble-guided diffusion for training- free text-to-image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Scribble-guided diffusion for training- free text-to-image generation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.573719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.413782Z digest=sha256:ca374817af94a404f3041ea59a38fb844be9bea9ed2f20011215e80206d1fc26

Observation 95650980-fe03-421c-b60c-6394730e1a59 · outbound

This paper cites Control and realism: Best of both worlds in layout-to-image without training.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Control and realism: Best of both worlds in layout-to-image without training

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.560246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.541193Z digest=sha256:60ef113f93e27e07c37dff1e59820eee17bd44bf73281b0dc0d9642ffddb6f3a

Observation 2790d098-00aa-493c-9028-770bc66b8674 · outbound

This paper cites Blip-diffusion: Pre-trained subject representation for controllable text-to- image generation and editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Blip-diffusion: Pre-trained subject representation for controllable text-to- image generation and editing

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.546732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.667169Z digest=sha256:6f50efeb9f13fa26a0193e3a4a0e1716bda74fd962f1c2cce605a8085fd7b30b

Observation c09a79db-1b83-4a3d-b1ae-3a4cb9a3e29c · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Controlnet++: Improving conditional controls with efficient consistency feedback

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.532230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.815482Z digest=sha256:f8d47ec610aeaa67b8250e06c5d5093fc3860a0a690863ce93d51b8e96b19ad0

Observation 960a6e60-2c2d-4e2d-8ead-6536d116ef13 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Gligen: Open-set grounded text-to-image generation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.519194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:54.965417Z digest=sha256:cf7a37a57e65d1bac40fec6b43ae4715f280082f82ada905d614bbe776c7d188

Observation 20c52763-bc6f-44e0-b33b-c82b4074c6b2 · outbound

This paper cites Freecontrol: Efficient, training-free structural con- trol via one-step attention extraction.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Freecontrol: Efficient, training-free structural con- trol via one-step attention extraction

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.506538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.088700Z digest=sha256:7352ccf3e1419697204586fbb8eea52d67ed35186ddcb3b7985d728251e30dad

Observation 6a26d6cb-a69f-4d0c-af21-91379ce08e41 · outbound

This paper cites Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.493565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.092548Z digest=sha256:45fcdfde00eb68a5aac618e43923f6b44049946ead89cba4b3e8680ac5b96428

Observation 624db162-15ca-4b87-9a98-44d161ffc86b · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow, 2023.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Flow straight and fast: Learning to generate and transfer data with rectified flow, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.480812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.120924Z digest=sha256:5e597adc04263e4fe01c049ece099e18a6abf9976a72aac82354027acee6f087

Observation 92d9a21c-d500-47dc-a4ec-6f284458b49f · outbound

This paper cites Dpm-solver: A fast ode solver for diffu- sion probabilistic model sampling in around 10 steps.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dpm-solver: A fast ode solver for diffu- sion probabilistic model sampling in around 10 steps

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.468254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.175981Z digest=sha256:7938562c950a9ba5f35a8d4654f584eb13e8014e20de21642587f2aa19569069

Observation cdd062f3-4148-4142-8f47-d8c8b8cc626b · outbound

This paper cites Latent consistency models: Synthesizing high- resolution images with few-step inference, 2023.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Latent consistency models: Synthesizing high- resolution images with few-step inference, 2023

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.456444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.224034Z digest=sha256:37da68b292b928e5ee43937848d218fab3b90df6b5b94d5d2aacbb264c7c515d

Observation 029bdc42-df15-40f5-9af2-1b4ee6d90719 · outbound

This paper cites Hpsv3: Towards wide-spectrum human pref- erence score.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Hpsv3: Towards wide-spectrum human pref- erence score

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.443958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.254247Z digest=sha256:0ca4cc496860222998c149575d788b98d3225f05d95d43e75f9f4c82a0a024fa

Observation cd6fd799-6b59-4c5e-9cf0-3e5a4ee95f9a · outbound

This paper cites Sdedit: Guided image synthesis and editing with stochastic differential equa- tions.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Sdedit: Guided image synthesis and editing with stochastic differential equa- tions

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.431156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.258418Z digest=sha256:0c077ff9dd05b48a5c6eeb163787ea0c1b3799f97dc5b36926fe425e3f3bfb34

Observation 1405355d-3c07-4a1b-9126-56f32216a6d1 · outbound

This paper cites Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.418526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.306480Z digest=sha256:292ac4f6300c96857ce0bb44ea453183a28f1f1c09021987b618c0920e20f68d

Observation fedd7da2-9b7d-4ec3-8029-8d09c57a5bee · outbound

This paper cites T2i-adapter: Learn- ing adapters to dig out more controllable ability for text-to- image diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation T2i-adapter: Learn- ing adapters to dig out more controllable ability for text-to- image diffusion models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.406467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.326191Z digest=sha256:8a0636ac240977d1d8a9a7a0a938ffcf4bb2eefdd313e45e7f1ae55e43fd4636

Observation 4babbe34-66e4-4906-9c59-7551354d83f3 · outbound

This paper cites K-lora: Unlock- ing training-free fusion of any subject and style loras.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation K-lora: Unlock- ing training-free fusion of any subject and style loras

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.392800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.331027Z digest=sha256:2cb4ff0ffb281667fc2f18144c57c58639ba25752e4e3d9027498c58cdda1588

Observation 47cf10d6-8c6b-436e-9101-2703203fac3e · outbound

This paper cites Semantic image synthesis with spatially-adaptive normalization.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Semantic image synthesis with spatially-adaptive normalization

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.379168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.335041Z digest=sha256:46a665e928d15ab2db3ba0807812797d1da4343f7a049a6d2c09abd3035168bb

Observation 89bd3eb1-85c4-43f4-9d94-80b24ac4806b · outbound

This paper cites Zero-shot image-to-image translation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Zero-shot image-to-image translation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.365639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.339841Z digest=sha256:8d34db30ffe516a1fd6278f1099a5afae7fc61c5f62001a269fe1b90c3d91240

Observation bef0e505-d58c-4249-b168-27a2004c3275 · outbound

This paper cites Scalable diffusion models with transformers.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Scalable diffusion models with transformers

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.352087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.344874Z digest=sha256:c72c4d04bd57275e1c954db4b56063e18715b2a36780b3d8d3af38cacde60c6e

Observation fa5a37cc-e600-47bf-b8eb-e8fd20de525d · outbound

This paper cites Pham, Jingye Chen, and Qifeng Chen.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Pham, Jingye Chen, and Qifeng Chen

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.339443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.349070Z digest=sha256:dd644812fea8195f11ae89547119285fd58b125930a6233ece0bd03ba169e96d

Observation c9cfd424-ec8e-4fd9-a5b0-97aa92f65932 · outbound

This paper cites Orthogonal adaptation for modular customization of diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Orthogonal adaptation for modular customization of diffusion models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.327339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.353017Z digest=sha256:8631a84c7c46e28975c0b6e00a07fb12e852685ac4e402c3df14b741f1561926

Observation 9eef02a9-41c2-4367-ba2a-de320f2c1c73 · outbound

This paper cites SDXL: Improving latent diffusion models for high-resolution image synthesis.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation SDXL: Improving latent diffusion models for high-resolution image synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.315568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.356925Z digest=sha256:fb1b075ed1349242ecb819166801955b40629497320a284daa2569e4ff892082

Observation 87ff7e98-e8a8-48e2-bc91-3b10a5c1607f · outbound

This paper cites Learn- ing transferable visual models from natural language su- pervision.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Learn- ing transferable visual models from natural language su- pervision

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.303233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.360891Z digest=sha256:3d014aa55d520b1bac02d6fb586b13b63de2af845a7188cec55240758b2babab

Observation 47adc399-c8fa-42b4-bf51-3737dbba611c · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:55.364539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:55.364539Z digest=sha256:ef98faffe9f5ffdf6d8c5c1bd06d670c742e6f118ed70b083b751d1d884232e8

Observation 6c76d348-7d10-444f-a921-e0491374ae6f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation High-resolution image synthesis with latent diffusion models

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.291610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.369174Z digest=sha256:e45c5e1df5c4b5a5fb7f9300465a042ce656b8b249d8c5b8cc29512c87963141

Observation eae5be58-37b1-47f5-bfb8-f776d2ebfc94 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation U-net: Convolutional networks for biomedical image segmentation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.278787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.373557Z digest=sha256:33741aaac149dca08bba5b241883ed3ff89af1e5fd1cabd35d0302bfff8a8818

Observation aec79a41-6ec4-4772-ba48-515d69891350 · outbound

This paper cites Rb-modulation: Training-free stylization using reference-based modulation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Rb-modulation: Training-free stylization using reference-based modulation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.265352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.377718Z digest=sha256:f45edfe24e5c321e69b498eb73be84488ac4e8f1e54b4007a5b99da4b5349b49

Observation a2166379-bc7b-48e3-b6bf-b447ee331d5e · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.252920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.381993Z digest=sha256:cb1aad538ab1c888b5a76b0e4236f8ca294ca0230dfbeb6b29c59db3f5e5e1ce

Observation 537bf2ba-c88d-4fee-8c56-feda37266a81 · outbound

This paper cites Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.241435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.386573Z digest=sha256:a4fa7009b082b4504fd454aebd788f6172b7598dd9b08126722195f6637f9243

Observation 0d832540-4ce0-4f70-ac7a-63d09e6fe44e · outbound

This paper cites Palette: Image-to-image diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Palette: Image-to-image diffusion models

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.227872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.390578Z digest=sha256:23694e4338bb4f33b1812cf73550065822f3f878b92e935feaefe1bf6e235bb2

Observation 9c3f62d6-e031-4f35-b656-47b1d7698bc4 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Photorealistic text-to-image diffusion models with deep language understanding

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.215554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.394553Z digest=sha256:fa975bd133c2cea27e1173d0aff1f7136cb58f75f5b97a455477b645461db17e

Observation 69f8e184-7542-414e-83a8-dad600acfb3f · outbound

This paper cites Emu edit: Precise image editing via recognition and gener- ation tasks.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Emu edit: Precise image editing via recognition and gener- ation tasks

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.203089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.398602Z digest=sha256:fa66e89216b73eae4ef05be497d177fee4c50e652a320ab1b0e812262419b98d

Observation 9843ef30-ec77-43c7-b96d-bf12ef3c653c · outbound

This paper cites De- noising diffusion implicit models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation De- noising diffusion implicit models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.189793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.402785Z digest=sha256:1fa53c156f9cce4121bfdbfd604a34f9c9bbd07d54b5b8eca27783e00e81e8e4

Observation 5b9da0b5-8910-4774-8c02-a17891e457bb · outbound

This paper cites Consistency models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Consistency models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.176988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.407089Z digest=sha256:f06de31e89d927decf888ba45fff5cbc498a4a3109b32ac1d4e3d5f681c6b5de

Observation fd29e967-4114-4e4d-b2de-d264cf3648c7 · outbound

This paper cites Dual diffusion implicit bridges for image-to-image translation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dual diffusion implicit bridges for image-to-image translation

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.164931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.411072Z digest=sha256:96bf8929a368bf4fa37f0dc6252de16b26d43d75ea031071c3d551cc7096870a

Observation 1bcbe839-d806-47ad-8353-1327d10a478d · outbound

This paper cites Ominicontrol: Minimal and universal control for diffusion transformer.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Ominicontrol: Minimal and universal control for diffusion transformer

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.152905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.414750Z digest=sha256:cafaa1262d6a6c5cfa9d9563c46761a4b3075a362fa93cfca5c147097b7132c9

Observation e9a960a3-7627-4871-88ce-b2da58e6929c · outbound

This paper cites OminiControl2: Efficient Conditioning for Diffusion Transformers.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation OminiControl2: Efficient Conditioning for Diffusion Transformers

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:55.418820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:55.418820Z digest=sha256:60462189c853d8bb4a9d8a65ab4a778f7b67d61fc25e3df696cb2a44aecab760

Observation a1006352-9c35-416f-8b68-3b934f468526 · outbound

This paper cites Add-it: Training-free object inser- tion in images with pretrained diffusion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Add-it: Training-free object inser- tion in images with pretrained diffusion models

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.141405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.422944Z digest=sha256:dd32ea626df4a636b21da1034ff3d0e00837e9a0f1f8f9feade8007d1a10eb10

Observation ab3a586d-1d3f-4d8b-aba0-80ba618e2810 · outbound

This paper cites Guide-and-rescale: Self- guidance mechanism for effective tuning-free real image editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Guide-and-rescale: Self- guidance mechanism for effective tuning-free real image editing

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.128545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.427394Z digest=sha256:bab31162e333079bd7588e3a0b4d9e31f3faa92f607f8f4fc6e1790bf740c5a2

Observation a378d218-8ff4-4248-833b-564ab939b5d1 · outbound

This paper cites Splicing ViT features for semantic appearance transfer.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Splicing ViT features for semantic appearance transfer

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.115871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.431794Z digest=sha256:7f524fcf6c85bb4c9432f515201483b35ba6a3dd62c6fd40fe51f8debc7bbd85

Observation accdbfcd-734f-4fa6-94ec-83f4126ab00d · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to- image translation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Plug-and-play diffusion features for text-driven image-to- image translation

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.103280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.435872Z digest=sha256:a07488762a49f5a490a4430faf2e1440119215237c476490821a2ac327fb74b3

Observation cc3161d0-5cef-4c4b-9dc0-6abf21f778f4 · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Diffusers: State-of-the-art diffu- sion models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.090597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.440178Z digest=sha256:1b726e12b0c79b535bc414035914e452916b379ad6c4947ca98eb17c33ef62dc

Observation 9ecec232-1529-4f0c-8e49-8b75efcf6708 · outbound

This paper cites Taming rectified flow for inversion and editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Taming rectified flow for inversion and editing

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.077209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.444139Z digest=sha256:6db9cb74404185f6141b0afbd299d9e9139869166d1d245a85655ada0f5f586a

Observation 3f931160-5402-467d-98f4-17128d193c34 · outbound

This paper cites Fleet, Radu Soricut, Jason Baldridge, Mo- hammad Norouzi, Peter Anderson, and William Chan.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Fleet, Radu Soricut, Jason Baldridge, Mo- hammad Norouzi, Peter Anderson, and William Chan

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.063153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.448035Z digest=sha256:99785abe0a6e9e80bf8455571870e554ff3a1827be1b3d1a66573d738020017f

Observation a30c6caa-e7a2-4129-b528-9a7ae2f176a8 · outbound

This paper cites Instancediffusion: Instance- level control for image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Instancediffusion: Instance- level control for image generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.049945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.452743Z digest=sha256:5d7c3b28abde94e89b680e6b26d56d84205b52338e542382f5617da81fe37370

Observation 4b74cc81-6479-4344-a7e8-b57a3d5a6336 · outbound

This paper cites Event-customized image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Event-customized image generation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.036191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.457190Z digest=sha256:fed3c8328925a284d04316c8321d39c1cadfc8e8594a87919e85e6db2e9d36e8

Observation 522d2c98-46dc-4308-a0b4-06447d0dd1f5 · outbound

This paper cites Training-free dense-aligned diffusion guidance for modular conditional image synthesis.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Training-free dense-aligned diffusion guidance for modular conditional image synthesis

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:56.023407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.461721Z digest=sha256:d88179845db747c67b9ae94e5b9b12e15fc08225418923d0d26f25ecb50fca3d

Observation 1812adaf-4bc6-43e5-9448-393b37afd95f · outbound

This paper cites an unresolved cited work.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:56.010988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.466342Z digest=sha256:4aac87c2bd2678de73f646a0df29071a79dbb405a3589c0f2d06ce7afd8e6d85

Observation 255e6fe5-90fc-499b-acce-17b6100e686b · outbound

This paper cites Qwen-image technical report,.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Qwen-image technical report,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:55.470098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:55.470098Z digest=sha256:c43cac7833df7661ceb444eef2be906c7b6275a61611c6e9d289f1f0733304ca

Observation eb004c59-db39-4231-9f15-10b846ac1c48 · outbound

This paper cites Omnigen2: Exploration to advanced multimodal generation,.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Omnigen2: Exploration to advanced multimodal generation,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.991151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.475011Z digest=sha256:17121482b2308b621dacfdb1bca02c271abca40642e768be03caf008ca18c9de

Observation 099b8b4e-5c8b-4af2-9252-0fe1a159a5e7 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:55.478869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:55.478869Z digest=sha256:12887694d528457522ce9952039a664f357cb05e09c82c1a2da2952a4d05d75e

Observation 17bc27b2-8711-45a0-a138-558c55fbb0c4 · outbound

This paper cites Dreamomni: Unified image generation and editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Dreamomni: Unified image generation and editing

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.979628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.482977Z digest=sha256:978cc84a79baeb610c0c798f3ad6b9b9dfd8e0a126a572a018773e0e2a68dd44

Observation 9299858d-c3a7-41bb-989c-2e20e2f857d0 · outbound

This paper cites R&b: Region and boundary aware zero-shot grounded text-to-image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation R&b: Region and boundary aware zero-shot grounded text-to-image generation

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.967993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.487289Z digest=sha256:4929ac18c2169aa79dcf8ede3e89a7d3b5fe3c6d51b910881894bb90d7262084

Observation d8152988-beab-4dba-b0e9-7a90ac043f64 · outbound

This paper cites Omnigen: Unified image gener- ation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Omnigen: Unified image gener- ation

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.954994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.491125Z digest=sha256:df93b2d0e10f6953490455cf242d85847fa2f6b44cac1e921663e2248f9e4c03

Observation d604960e-4e1a-477e-884f-f74ae1b07331 · outbound

This paper cites Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Boxdiff: Text-to-image synthesis with training-free box-constrained diffusion

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.943467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.494874Z digest=sha256:620b48a4a824d511bb05ccdc7349dea0b3bc27f9706c78af218e08d256951a3f

Observation 5b039525-bd97-4a5c-aaf5-f85dd4d3d2c8 · outbound

This paper cites Anyrefill: A unified, data- efficient framework for left-prompt-guided vision tasks,.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Anyrefill: A unified, data- efficient framework for left-prompt-guided vision tasks,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.931415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.498993Z digest=sha256:9d9420191c1f3b400b46aee6ef49b75f11a431529329d383fe3426b939ff88f5

Observation 6e82d324-17da-45a4-83b5-8f9e4b8d3fdc · outbound

This paper cites Imagere- ward: Learning and evaluating human preferences for text- to-image generation.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Imagere- ward: Learning and evaluating human preferences for text- to-image generation

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.918919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.502839Z digest=sha256:122c216d0bac7c77bcf490e94a8513d098aed844b0ce63d430e0797a159ea5a4

Observation c41745bb-fb74-49ba-b467-a6dc6ff93970 · outbound

This paper cites Unveil inversion and invariance in flow transformer for versatile image editing.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Unveil inversion and invariance in flow transformer for versatile image editing

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.906896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.506806Z digest=sha256:30258d7bd8f417ad982e379615c39c4a17b372397fc3710968f7f699b4723c5c

Observation c279068a-2cf5-47c7-a559-053d2fc529aa · outbound

This paper cites Inversion-free image editing with natural language.

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation Inversion-free image editing with natural language

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:55.894544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T20:24:55.510895Z digest=sha256:c898cc428cdda0a7ece12ed68cb668cfc37782987f3ca7fcba16990e7bb8da57

Pith citing papers

No inbound Pith citation observations are available.