Pith. sign in

Paper Citation Record · LEDGER

Control and Realism: Best of Both Worlds in Layout-to-Image without Training

As of 23 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2506.15563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15563 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:39:33.114969Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy44
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9964cf79-1324-411e-88c9-09b2ce98d650 · outbound

This paper cites Dreamstyler: Paint by style inversion with text-to-image diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Dreamstyler: Paint by style inversion with text-to-image diffusion models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.595471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.758922Z digest=sha256:192a38181cb4f38787f6ad40ade3b9d160ec90f11d0979e4ebe45e81093d5a35

Observation 4599e3e6-5e86-4cf6-9e7d-4913b2dbdb40 · outbound

This paper cites Multidiffusion: Fusing diffusion paths for controlled image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Multidiffusion: Fusing diffusion paths for controlled image generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.575065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.764413Z digest=sha256:5c958addde33a36d2cf2d90789832f0690558b145846e748c8ee75054dd3214c

Observation 39a28c7a-dcec-455b-b6e3-cbcf24a6dfbc · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:34.547568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.769315Z digest=sha256:ed8de236f0b7eb3d3149b950e243d2663c723974cb2ec7bf1d92279d445683ec

Observation 583bd8dc-48d7-4610-b87a-dcdf4558737f · outbound

This paper cites Artadapter: Text-to-image style transfer using multi-level style encoder and explicit adaptation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Artadapter: Text-to-image style transfer using multi-level style encoder and explicit adaptation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.523856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.774126Z digest=sha256:6dbc4f8fdc2904af9b7c93d217ddef5b6bc161c6ed6d3172b3b58695b9e22dc5

Observation 709064ba-ed8f-4040-a81b-1e9dce09491e · outbound

This paper cites Boundary attention constrained zero-shot layout-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Boundary attention constrained zero-shot layout-to-image generation

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-15T19:39:33.281778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.778673Z digest=sha256:b0c493e692339ce37a22f1a07f684a31197a9315e97333a13187593476584cff

Observation c3b23d48-f0ea-4268-bb0e-30656c47b3ec · outbound

This paper cites Pixart-alpha: Fast training of diffusion transformer for photorealistic text-to-image synthesis.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Pixart-alpha: Fast training of diffusion transformer for photorealistic text-to-image synthesis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.504934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.784874Z digest=sha256:0f30b3deccdde61fd5599fa734fa18611e7c8acc96b98709569ebf3ce7fec324

Observation dd7de1fb-91f8-457a-ad21-c97b69eb5f91 · outbound

This paper cites Training-free layout control with cross-attention guidance.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Training-free layout control with cross-attention guidance

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.477236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.790046Z digest=sha256:9b52a91fa6dcf5e22feae91384a02fde9b0b752e1fdac9f99df76c243c32594b

Observation 2ba8cb0c-a23e-4c76-8f13-00563b3e2bdc · outbound

This paper cites Vp3d: Unleashing 2d visual prompt for text-to-3d generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Vp3d: Unleashing 2d visual prompt for text-to-3d generation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.456850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.796447Z digest=sha256:7a24e7d727ef662c7efcef2a3fb30214a5330cbd207095eed12571ae432ddd12

Observation a691c3f3-4a49-453a-b05c-99e610291666 · outbound

This paper cites Zero-shot spatial layout conditioning for text-to-image diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Zero-shot spatial layout conditioning for text-to-image diffusion models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.433390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.801465Z digest=sha256:30e8ed475d742dadc33002dbe02f33c49d2a60e4c7dfba6206352e65b32d34e7

Observation 0da3a683-4da3-46ae-bfe7-ce4a3d796ec3 · outbound

This paper cites and Nichol, A.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training and Nichol, A

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.409927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.806910Z digest=sha256:887c36f60b7be4b93fd374ff906ca18cef7361290f27f486457016e87dc279a0

Observation ba09fbb3-c620-46ab-9282-a238d37d2354 · outbound

This paper cites J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.388514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.812701Z digest=sha256:5138e882b2c89331e74ebae0a651abc507b9c8d0cde328d5db5eea115a5ec686

Observation 742dc1df-8d03-4aef-89dd-fda53e3e0d2f · outbound

This paper cites Denoising diffusion probabilistic models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Denoising diffusion probabilistic models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.360205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.817868Z digest=sha256:99370678660ff06eeeeedb24770a5d952ff67e568e6c98a8314ea78b86d0e6cb

Observation 0f13fe51-31fe-4f97-bb14-fe0a9f5fb8a2 · outbound

This paper cites J., Norouzi, M., and Salimans, T.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training J., Norouzi, M., and Salimans, T

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.824800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.824800Z digest=sha256:0deda3db2f2263183c512c894da31972bdde6b85a323fa4404bc069d47eaa80d

Observation 19de3453-4595-4711-b4af-5c7bd193dfd0 · outbound

This paper cites Ssmg: Spatial-semantic map guided diffusion model for free-form layout-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Ssmg: Spatial-semantic map guided diffusion model for free-form layout-to-image generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.318045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.830229Z digest=sha256:4b781caa19c43951c05308acfd2e30dbc96f98803d0cfed9385c880162978ba5

Observation f2f6b5d2-5df9-494b-ba88-a1ed5968af42 · outbound

This paper cites C., and Liu, Z.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training C., and Liu, Z

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.298042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.838535Z digest=sha256:42544d79e30b7dd649e31756ec5d6b2eb7d28374e868cb9fb51192a18db224c3

Observation d1a29da9-37ee-4fa6-9e10-4a512a557891 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Imagic: Text-based real image editing with diffusion models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.273840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.843602Z digest=sha256:7d16c9743e9493366120e363377f5e1bb227b25954a455c9b6d926be1a99798c

Observation 6d0765c7-3246-47b1-8d88-5282f958155c · outbound

This paper cites Dense text-to-image generation with attention modulation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Dense text-to-image generation with attention modulation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.252969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.848829Z digest=sha256:3e32bb207044be3680c3ec40399f34da25617de8094ca37ab0c2b4e8f01f579c

Observation 8282940d-9fe9-4ad0-9746-841448a7928c · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Pick-a-pic: An open dataset of user preferences for text-to-image generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.230462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.853259Z digest=sha256:0f5dfa885df4d3ecc0b33a3bd988e7ac5da9554ad4085fe2bacd4574fb1f2586

Observation 9eea1795-8b9e-44b6-a7b8-3f97c5246401 · outbound

This paper cites W., Zhou, Y., Liu, D., Lee, J.-Y., Cai, H., Liu, B., Liu, F., and Uh, Y.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training W., Zhou, Y., Liu, D., Lee, J.-Y., Cai, H., Liu, B., Liu, F., and Uh, Y

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.211206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.857931Z digest=sha256:4e168628b7e34c9eb63fccc4b2b2933c2d23ea2a323c051b4285f30189b43c26

Observation 47da847f-4bbb-42b8-9a02-91e91b7f80d6 · outbound

This paper cites The role of imagenet classes in fr \'e chet inception distance.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training The role of imagenet classes in fr \'e chet inception distance

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.184928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.862467Z digest=sha256:600c4aa23cc0fdb4f980a05aeeb5381a2bee1432d6a6ec3ab4af2627f74a94bf

Observation 0b1de4c6-08a0-4fbc-942f-64ec6cef9fa7 · outbound

This paper cites Instant3d: Instant text-to-3d generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Instant3d: Instant text-to-3d generation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.161247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.867028Z digest=sha256:3173088eb0f06845d56d00511b7c09b010b016b2b926b6bdd7c1fa12a5b806c5

Observation b6a7a721-b7ed-4a06-baa0-0e40d345f945 · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:34.144275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.871302Z digest=sha256:94a4ba531f002c6ce52e76a6340ee66a323f69917433106a771a539d3346ba20

Observation f776caa9-6247-40b6-b73c-3a4e61184b9b · outbound

This paper cites Image synthesis from layout with locality-aware mask adaption.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Image synthesis from layout with locality-aware mask adaption

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.115681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.875415Z digest=sha256:18410dd9a6fc7035ff688a09a89e6f5dcc9e41b181762038360a5ad90378595e

Observation 09b92f7e-8082-49ed-bc82-1c3487f72c6e · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:34.080273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.880067Z digest=sha256:9a1a7e3eb0b9a3f9052c25351fc6f349278e9f9139262cd7e5f76c636b4afeb2

Observation 506ba1ac-3b0e-42ca-ba1f-4daff2098c99 · outbound

This paper cites Training-free composite scene generation for layout-to-image synthesis.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Training-free composite scene generation for layout-to-image synthesis

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.049123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.889072Z digest=sha256:7b8c5c3ceb71445ddc1c2526f42a9c7fd72ba41b8dc98f2b41d54956967d42b3

Observation f356b745-0ea3-40ae-a3b1-a8cbe02dcf27 · outbound

This paper cites Hico: Hierarchical controllable diffusion model for layout-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Hico: Hierarchical controllable diffusion model for layout-to-image generation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.025510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.895304Z digest=sha256:d5fbe60cef51719847a00be4bd5b6c5c26cf69b826885e2c9d3dc9086fbaaf95

Observation 363bfdfa-3e93-4907-a185-15b0d70f7d3e · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Null-text inversion for editing real images using guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:34.000059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.899921Z digest=sha256:5e29749ef14f3a4ce38e752548339d21dd7b9e368396126e79dc898349da7f8d

Observation 8cd3df7d-aa35-4f24-bdaa-c5837a359043 · outbound

This paper cites Multi-Task Learning as a Bargaining Game.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Multi-Task Learning as a Bargaining Game

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.905572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.905572Z digest=sha256:fc997857ac9dcce9ff37014875335c1d2f837adc5eb9bfbb75bc24c532582939

Observation f452c7cc-bed6-4f1f-8265-ff86a379950f · outbound

This paper cites Zero-shot image-to-image translation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Zero-shot image-to-image translation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.958726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.911525Z digest=sha256:6c3670ec49ad79bce0ebf91d57871bd8c823c11b54ee3043c9dfb2ddd2db926b

Observation d08d67a9-aa44-44e0-bff8-62a158f043f7 · outbound

This paper cites and Xie, S.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training and Xie, S

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.919553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.919553Z digest=sha256:f85eb6bb49aa59b78c7901fbed1e31021d73751d73c52e946ff258bfbcb0b3b7

Observation 2dd468d8-c2c7-49c1-81b8-f0e427795e3a · outbound

This paper cites Grounded text-to-image synthesis with attention refocusing.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Grounded text-to-image synthesis with attention refocusing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.927054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.925459Z digest=sha256:d358cf21016a728cb5ea0acface080232bc98eee98df06c734d3f6f042fa57aa

Observation ba6f1867-8db0-41d2-b909-816aa3cdba4d · outbound

This paper cites A., Wang, L., Cervantes, C.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training A., Wang, L., Cervantes, C

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.899579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.931276Z digest=sha256:ffc2d3b6142955cee6657e244ffe83426f74a3a26f7c64cfdb45d5879f298e5b

Observation c205fd3e-f196-4ef5-8a97-26358e2ce691 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.943305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.943305Z digest=sha256:3c271646b7fe01fbc99151c4628d7dbf205707871e591870a830d4f83a8dc648

Observation 60bfd674-10e2-4fb3-b81f-158efe77b3ab · outbound

This paper cites Hierarchical spatio-temporal decoupling for text-to-video generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Hierarchical spatio-temporal decoupling for text-to-video generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.872648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.950249Z digest=sha256:b2c0bc7785d5149467df56cef820f1840b8c1f78fa372e5be09bca12c280eb09

Observation e39618b1-b766-429e-84d5-ce6ac49b52d1 · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.956795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.956795Z digest=sha256:1e3d974d1d09ba72e0b67d189f0dc244c657d91ba1dac0c91f79e6b51ee984d4

Observation 8157c530-605e-4caf-8d6e-c8c1efaabcaa · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training High-resolution image synthesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.961678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.961678Z digest=sha256:ba76cddbcd242e4b046ab16d65bb971ff51807242e8200822b4d10f38b9f17da

Observation b4b32b84-0992-49ea-85f7-91cced8dd8d6 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training U-net: Convolutional networks for biomedical image segmentation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.966909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.966909Z digest=sha256:eda27f169eb54e057197b6d66f23acb02bdce505e1f21aa07240e2a11ed41d5d

Observation 4f887243-9669-4e31-aeac-0cd40e0da428 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.787466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.971523Z digest=sha256:2c45e5873ed3b5d7729355b5d681ffb470a2dfab02f63f5802e2e481d815128a

Observation 0ce2bdee-b3ea-4a9f-82d2-d05e6e6009bb · outbound

This paper cites L., Ghasemipour, K., Gontijo Lopes, R., Karagol Ayan, B., Salimans, T., et al.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training L., Ghasemipour, K., Gontijo Lopes, R., Karagol Ayan, B., Salimans, T., et al

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.767525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.976557Z digest=sha256:c0d7314b154b926f87ef58b6322a96a56aa7b8d946e14258cce2d780e79c37c2

Observation 10d30582-dd04-47e4-ae55-d186736975db · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:32.982045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:32.982045Z digest=sha256:a2b066c05d3cf0033575aa01795d01bdad7546e45e3514799e7e134f3666b676

Observation 3bfb80a7-55f7-42d6-91bc-ff7f3bd4ce54 · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.745037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.987429Z digest=sha256:ae508150ebae657ecd7668829667c5f3aadc06f9c364f6f7b5ce590040d5e91e

Observation 7ab556a0-960d-48ec-8586-dcfe7ad7d250 · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:33.727001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.992787Z digest=sha256:002dfb428d5850892728fe1e7f96c3a53163311c254a2b81ca877b3d9aeb3e80

Observation 55928e86-8014-43fe-ae75-4e6d396e66d6 · outbound

This paper cites High-fidelity guided image synthesis with latent diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training High-fidelity guided image synthesis with latent diffusion models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.694975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:32.999536Z digest=sha256:84a2e62ce598af99a05536f5cb691806b29200f18338698e4dd8a07c706e5700

Observation 55d0a7ff-f0aa-4b58-9028-cafc8be26cea · outbound

This paper cites Styledrop: Text-to-image synthesis of any style.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Styledrop: Text-to-image synthesis of any style

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.670575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.005728Z digest=sha256:c5f0027533380a868737e36b12831b0c65ce55d3fcee87183b104ad3a060d0f6

Observation 8b390fff-bb87-421a-9df5-e37a25c078ae · outbound

This paper cites Denoising diffusion implicit models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Denoising diffusion implicit models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:33.010133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:33.010133Z digest=sha256:1f8a0c9d59176531303a743697974d9424b66d0b1eb306c84ef325a2e7ff9f5b

Observation e8a1b145-0a96-4900-810a-ec7f06e4de46 · outbound

This paper cites and Ermon, S.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training and Ermon, S

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:33.016572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:33.016572Z digest=sha256:d46a09e243b0837c1ad6860b2a892fea40b4af2d9ff3cd18f56bec9a9c186ae6

Observation 7a82103d-cb35-4cb7-ae6b-bea554cbd578 · outbound

This paper cites P., Kumar, A., Ermon, S., and Poole, B.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training P., Kumar, A., Ermon, S., and Poole, B

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:33.021573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:33.021573Z digest=sha256:e60f51fc4c3f4d230beee31b975a43b898ef9de71eda9cc1517f801ef74c78c7

Observation d43fa42f-f361-4661-8334-5552b07ef95e · outbound

This paper cites Z., and Poggi, M.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Z., and Poggi, M

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.600871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.026679Z digest=sha256:bb15f938ac6902ed8d1f05259e61da919ee009d7b0d28c8955e464b07fd7ad0d

Observation 2d198d4c-5446-4bfb-b4bf-b36476e2e859 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Plug-and-play diffusion features for text-driven image-to-image translation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.584668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.032813Z digest=sha256:0c8f453f221bb8c15a4409967f7fffbc26942bd4fc31de8c611a1ab3792b32b2

Observation dcd19ba3-2c48-4bd4-8ad8-3875aa318116 · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:33.567308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.036882Z digest=sha256:a92e2b87f3cd20e111a582aaa156f6dedf6661345781e9598ecbc9846d336da5

Observation 5ab45eec-f0d4-4b2a-92db-08729bf4ee3e · outbound

This paper cites S., Girdhar, R., and Misra, I.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training S., Girdhar, R., and Misra, I

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.548715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.042630Z digest=sha256:d885c85f8cf135c43c037ce4c8a2ac2ad49f720f3ff3a2ccbd8cf7fa47d1b75d

Observation 0ab26e7d-c41b-468a-9679-4fd34a0c7670 · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.521946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.048053Z digest=sha256:4272316ad079d11731bdc2ea213041cb86eeef4e563db533e0a5ccedcaac0dc3

Observation 5757c3e6-7608-4f68-8e3a-c2dd8bc05091 · outbound

This paper cites Z., Ge, Y., Wang, X., Lei, S.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Z., Ge, Y., Wang, X., Lei, S

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.503173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.053717Z digest=sha256:5e25a231d733d2fd16dd47fa44765c18b41ce6baf4839fd427b7632283916e5c

Observation cd394c21-902f-4ad4-a3ec-a53980f3cd6a · outbound

This paper cites IFAdapter: Instance Feature Control for Grounded Text-to-Image Generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training IFAdapter: Instance Feature Control for Grounded Text-to-Image Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:33.057727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:33.057727Z digest=sha256:b443c6fb20af328edd261bb3fddeca977f2ef570fb51c825d15591dc3af5c14f

Observation 3c992d49-261c-4f31-929e-525c3ce3b79b · outbound

This paper cites R&b: Region and boundary aware zero-shot grounded text-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training R&b: Region and boundary aware zero-shot grounded text-to-image generation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.477217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.063722Z digest=sha256:04b74a3dac4908f5a59ace0d465d8fd1d85dfd54a327cc39c874df2370461bd8

Observation 81872ef9-8658-4f69-bbef-84e6d2891668 · outbound

This paper cites an unresolved cited work.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:33.453791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.071425Z digest=sha256:1a8e63f86c58877b2c78c35f3db72e84a3882afb106e20f505cd03ce05e8d4bf

Observation 74bb8536-e27a-442f-aeb0-a80abb8d13a1 · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.437589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.078541Z digest=sha256:92c62ced1bb83fca9a2acb941da1236fd837132ad5ef65f43f1c9d27645c8a82

Observation fb12e919-e639-491a-b341-6d4522128fe6 · outbound

This paper cites Inversion-free image editing with language-guided diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Inversion-free image editing with language-guided diffusion models

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.413121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.084248Z digest=sha256:3cc6bb092ac13906ef94f3cc406bb2a2a940655081677b9797f1bc59c98b4a22

Observation fbdfc0e2-b3eb-4b04-b5d4-8ded009edbc8 · outbound

This paper cites Freestyle layout-to-image synthesis.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Freestyle layout-to-image synthesis

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.394558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.089038Z digest=sha256:8639274fd79d387d05e6602dc5ebfebaac3fa89ab96b8a7b8a0d8f0b665dd13d

Observation cdaa147f-3269-4b1c-843f-acb4f41395c3 · outbound

This paper cites Reco: Region-controlled text-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Reco: Region-controlled text-to-image generation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.374615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.094550Z digest=sha256:9b4f5bf97137284ca76f24566ac87f772587f59cf46f697d2add09fbf3a8f790

Observation 81279142-8d1f-4fb8-ab91-994fbe7aa14f · outbound

This paper cites Towards consistent video editing with text-to-image diffusion models.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Towards consistent video editing with text-to-image diffusion models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.354759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.101265Z digest=sha256:0652fed70082585af6b20d27f0c8f084286b702099a7df8e73dd9c0ed8ae7259

Observation 4d8e3ad9-f770-4178-baa3-77be846a3405 · outbound

This paper cites Layoutdiffusion: Controllable diffusion model for layout-to-image generation.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Layoutdiffusion: Controllable diffusion model for layout-to-image generation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.337710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.105923Z digest=sha256:3fd9acf6584afd5dc878c1c5f6c564bdaa4c7a39dac3180fc1a75809b2a5a8bb

Observation d6cc944e-2f0b-4a17-81f2-832cd5acafcf · outbound

This paper cites Migc: Multi-instance generation controller for text-to-image synthesis.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training Migc: Multi-instance generation controller for text-to-image synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:33.319198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T19:39:33.110963Z digest=sha256:f542b2588f808e14ce458b9f45a52658a2044d9e86f41fd6cb2b17f81c652d1a

Observation 161595ce-bc65-49c8-bc7e-8f81cec4b469 · outbound

This paper cites write newline.

Control and Realism: Best of Both Worlds in Layout-to-Image without Training write newline

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:33.114969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:39:33.114969Z digest=sha256:6c1126127920db00f329aaff19064817a6314625844b173575b202db0c1234c3

Pith citing papers

No inbound Pith citation observations are available.