Pith. sign in

Paper Citation Record · LEDGER

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models

As of 17 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2509.04446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04446 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:32:15.454530Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T19:51:31.657802Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T19:53:11.597613Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ec10e816-9bec-4557-833f-bc87b516bc56 · outbound

This paper cites GPT-4 Technical Report.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.245097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.245097Z digest=sha256:30b57893802ed502297572c6b517598bee99cd71d1333cd087bee5f050ccad45

Observation 75cfc74d-c150-493a-8165-c7ef7faf0426 · outbound

This paper cites ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.250967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.250967Z digest=sha256:8b2765a36fd53738bb11e3f8c2e96320733ebfaa828e913e0a1547747f11bf1c

Observation e99ba86b-632f-454a-9b4a-2a9cb832e62a · outbound

This paper cites Blended Latent Diffusion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Blended Latent Diffusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.255493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.255493Z digest=sha256:3ce19490749b1e1c0ddb47aa096c263a04cc5b8ab0c385b75de245b14ed14a44

Observation 20b114b9-1580-4fe2-9611-e43ddc54dd4e · outbound

This paper cites The Chosen One: Consistent Characters in Text-to-Image Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models The Chosen One: Consistent Characters in Text-to-Image Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.259614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.259614Z digest=sha256:cc317a496bdd70b55376d27ec6e6e3cdfe03367ff9a65b9d75f842fc6e267d45

Observation f56c5461-407e-4037-9d84-c9e1b26ee601 · outbound

This paper cites SEGA: Instructing Text-to-Image Models using Semantic Guidance.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models SEGA: Instructing Text-to-Image Models using Semantic Guidance

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.264293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.264293Z digest=sha256:80a7f3fa759ff21b5e91f1b449044ab6d6c5988aa22ea8fc96bd4b24be1390a1

Observation 567bd695-be3f-485a-92bc-78e9891a4310 · outbound

This paper cites Ledits++: Limitless image editing using text- to-image models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Ledits++: Limitless image editing using text- to-image models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.095550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.268891Z digest=sha256:a59b2d6fe459758ec6769fb36e293f14ddc76330ef556c32a449c19ed3b6aae1

Observation 4bdf3bb4-eb11-4b9a-a585-9812b8da4bd2 · outbound

This paper cites InstructPix2Pix: Learning to Follow Image Editing Instructions.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models InstructPix2Pix: Learning to Follow Image Editing Instructions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.273719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.273719Z digest=sha256:38483c5cb36194f303c5450f70d6e61997e6f74d21281e7f290dc4f31f44bb2d

Observation 05f043fc-a80d-404a-820f-de9b06d98443 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Emerg- ing properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.279215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.279215Z digest=sha256:fb6f1caf1c03b2b10d380dd4a6af8db2f0deea98d53993357484cc85eab2a02f

Observation a08540ac-7be7-4f99-a965-535f544bbce4 · outbound

This paper cites AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.284324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.284324Z digest=sha256:c55bf9bd5bc3158977454853ac71976792108ea75eb4cdf14847a8ee0ccd6889

Observation 72e2eba5-85a0-4c0e-9dac-c3c9333c742d · outbound

This paper cites TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.289174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.289174Z digest=sha256:c4ff77e2255e2e76600135e5a037fe64be7af5ade4d2ed50e95fd886c617be6a

Observation 023ed9d8-31c2-431f-afb8-441390605b70 · outbound

This paper cites Yolo-world: Real-time open- vocabulary object detection.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Yolo-world: Real-time open- vocabulary object detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.065920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.294108Z digest=sha256:7c0b80ae22301eafe3f984c0ca4bb482dc4ae735285f9b58518ff581e13ae70f

Observation abd3b556-d7d3-4824-b760-16f7c879e7c2 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.299028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.299028Z digest=sha256:5c99e39fe164b9ddad79b142ff72a09a902a061f05f2f9fd88bcc47a415565be

Observation cdfca3c4-af59-468a-aff9-9057c7e60484 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.303626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.303626Z digest=sha256:b94e5cf691cc3cb544ff0947de7bec64c8df1386be07bed13c1492634d422430

Observation 1e24d965-133d-4d7d-b415-3e2ceb5a864b · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.310852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.310852Z digest=sha256:9da8f7d7cc90bfeafd81ea363978dd37063cf0db088511a16f10b69bcd73aa94

Observation 7f1d5e92-0182-4b6c-bc51-96102be9b408 · outbound

This paper cites Prompt-to-prompt image editing with cross attention control.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Prompt-to-prompt image editing with cross attention control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.316405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.316405Z digest=sha256:4b800e23bec1788a4abc82af4b0d542887ffbbfb8409a034b3f078751e595dd7

Observation 3740d32d-9f22-4d50-a3b2-f79a0eaa0f06 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Denoising diffu- sion probabilistic models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.322765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.322765Z digest=sha256:09844189b489486b1bb55a94d13e27e535b67fe73cea4bc56f366482063c03a3

Observation 91c1462f-b309-44bf-880a-6e537600a9b3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.328415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.328415Z digest=sha256:551593717cf1efffc6a80484a96cc6513c4b427b8f95c7d24529918befa97a8f

Observation 5ae3d5ec-1f06-4366-be0e-93eb42d6e43f · outbound

This paper cites Zero-shot generation of coherent storybook from plain text story using diffusion models, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Zero-shot generation of coherent storybook from plain text story using diffusion models, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.018633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.334093Z digest=sha256:4ce9972062d3776fd89387ee10dd28a45eba4d231debe7e83af5dc4de6358af3

Observation 65fdaa25-bca4-4476-ae2b-64723c054408 · outbound

This paper cites Rehg, and Pinar Yanardag.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Rehg, and Pinar Yanardag

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.004941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.338889Z digest=sha256:c12e86ed9c97b5ae1c700996d8a638e1ac9a768bc0da31f86fed4c76d15d37e3

Observation 104fb183-4600-42f1-a24a-38400b5e14fc · outbound

This paper cites an unresolved cited work.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:15.991864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.345203Z digest=sha256:593fa41dab97b948c8fce038e3b43c71b5a3a5e7610d932d5385905420d35972

Observation c35c054c-e6b4-4508-80a6-aaacac436d56 · outbound

This paper cites Intelligent grimm - open- ended visual storytelling via latent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Intelligent grimm - open- ended visual storytelling via latent diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.979050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.349572Z digest=sha256:dddd83bb2df171dc7416cf1763accc2c95c1b262c7721851709115bddbce1571

Observation 0bec2a1c-7fba-4523-ae6b-a34d0af0ccfd · outbound

This paper cites Compositional Visual Generation with Composable Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Compositional Visual Generation with Composable Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.354327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.354327Z digest=sha256:df3d1fc707098f1b23881f0563761d4934e3966c0e5848c97af707397c3b3059

Observation dfc7d8d8-0c02-4956-bdc2-44da5466c05a · outbound

This paper cites One-prompt-one-story: Free-lunch consistent text-to-image generation using a single prompt,.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models One-prompt-one-story: Free-lunch consistent text-to-image generation using a single prompt,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.964746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.360081Z digest=sha256:1b117c7b0450a6aaebb4cb5f6cdbaf070fcdb40790f1be3bc50734119174ba1c

Observation 0d48eccd-937b-49a4-99b4-a75eac916e2a · outbound

This paper cites Storydall-e: Adapting pretrained text-to-image transformers for story continuation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Storydall-e: Adapting pretrained text-to-image transformers for story continuation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.364085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.364085Z digest=sha256:7f904f05d7d810f64e8863639965a620b91a4be52ef16a6665d2cfd8850fa52f

Observation a20ec3eb-bb09-4089-ac12-fece52d407bc · outbound

This paper cites Synthesizing coherent story with auto-regressive la- tent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Synthesizing coherent story with auto-regressive la- tent diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.368858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.368858Z digest=sha256:7b86f4e155db27410d2e5910cae99c801cf358620ea802a61c32d834d8c47328

Observation 6fd444aa-4d76-48ed-bafa-c32787b56865 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.928794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.372979Z digest=sha256:f5c8356d82f429c913724996da853523f6f43d6ac37a0a0b997b4cadf9487cc1

Observation 7f431f05-1ad7-48e9-b8a1-deeba02e6abb · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Learning transferable visual models from natural language supervi- sion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.376979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.376979Z digest=sha256:1bfda77e6b2084771725d2ad12bc9dcc4f1692544aa0c099329175c1a74b73e0

Observation af38eb90-2aa1-4dc0-86ee-1560d61d0040 · outbound

This paper cites Make-a-story: Visual memory conditioned consistent story generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Make-a-story: Visual memory conditioned consistent story generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.380905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.380905Z digest=sha256:6da37ac4e7b4f999ced5f2dfbd65607d66c96623f27b37a284ab0fc584957345

Observation d06b27e6-3103-4909-a357-87d18bb466b5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.385321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.385321Z digest=sha256:e6e5ae47f552d82c4672da30854ce5bbf3ae9f87bbfa8919f61eff9818288805

Observation 8772e8a5-3c82-4e16-abba-4c9bacd3e796 · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models U- net: Convolutional networks for biomedical image segmen- tation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.388970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.388970Z digest=sha256:3654f302c92efebad3d1e4d1e744207cf9a96ce2e41de91d1e8969b94aee55d4

Observation 31cb85f4-7843-4490-8a8c-8b02bffac400 · outbound

This paper cites StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.393267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.393267Z digest=sha256:d5938770f844a8c792706efb163aafc01147bfcb221e155a56dcf1663c4529da

Observation 47cd6ce4-141e-45a4-a9d4-5c3bd58edf33 · outbound

This paper cites Training-free con- sistent text-to-image generation, 2024.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Training-free con- sistent text-to-image generation, 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.868777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.397679Z digest=sha256:400228b47213fe942ec7f6772a74f6b7eae4efd83d9f821fef1d6038e4ed863e

Observation ebdea003-d86f-4b7d-982f-f65ec5672f68 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Plug-and-play diffusion features for text-driven image-to-image translation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.855287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.401911Z digest=sha256:5183cf01cda365d075549f1070c5b2255149b7e02485287f45464692635f699c

Observation 466afb1c-8997-48ae-a3fa-d403378039d2 · outbound

This paper cites Unitune: Text-driven image editing by fine tuning a diffusion model on a single image.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Unitune: Text-driven image editing by fine tuning a diffusion model on a single image

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.841121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.406601Z digest=sha256:cff472b88682ce305e6a7f75e01c5fe64a54beb76ebf84de495bde151569f88c

Observation fb5c0268-3b07-4b19-9a57-ad596bd543a4 · outbound

This paper cites AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.411330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.411330Z digest=sha256:7d19f88b1b4b8c934069b1e91fd480516db8742fc6a3ca683daa7c281620960e

Observation 7f7efe33-3ce6-4ada-84d5-a809d3f62ccd · outbound

This paper cites Nerfiller: Completing scenes via generative 3d inpainting, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Nerfiller: Completing scenes via generative 3d inpainting, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.824288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.416686Z digest=sha256:359db917e80e0c1250c7ce9be199c3dac42977f9b0abfd50aa9e4d72e3a1ea98

Observation 5a27e0b5-1a4f-48d1-b6e2-71a169312b34 · outbound

This paper cites Efficientsam: Leveraged masked image pre- training for efficient segment anything, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Efficientsam: Leveraged masked image pre- training for efficient segment anything, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.799095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.422393Z digest=sha256:4979dbe7e344cc2eae3d1f2d369177374b6303ddc7d801c0de399107ab204790

Observation d6ca206d-e27c-4a3d-9ecc-e3e62a9f4729 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.429255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.429255Z digest=sha256:0788b434b2024c7f0efec85ab6d7bc8d8a1f5cf18587b657897780bbbdd93e84

Observation f74521c4-473b-44d3-a342-98a6c9f39be7 · outbound

This paper cites Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.780780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.433717Z digest=sha256:f55dfd42d92f7171f0818785735fbb777669ae4dbccc1286d8804530199694cc

Observation 49114b78-55e0-410e-9fca-66092e9fe2fe · outbound

This paper cites Adding conditional control to text-to-image diffusion models, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Adding conditional control to text-to-image diffusion models, 2023

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.438012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.438012Z digest=sha256:a766de5f15375699afa99e73e516c02e446158c96b1969340077f522594bcb52

Observation bd67c397-40f9-4619-b1ae-87ea982011ac · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.442490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.442490Z digest=sha256:45b9797f3a79116e8bc604842ad228e06d565577ceb1025746c3c89fa138d651

Observation e37a885a-6bc3-474c-b000-8133b8e551e3 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models The unreasonable effectiveness of deep features as a perceptual metric

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.446611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.446611Z digest=sha256:33fbb2ba668dd55345bae99c78ba7018ce8d22b7cc9c25933b3bcc72b41f7966

Observation 450a8d1f-0422-441c-9e18-783561a3fae1 · outbound

This paper cites StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.450281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.450281Z digest=sha256:e69468af121bb540613cdead3bca035d18ab93a5f83934e3d5e827d2497ae5fc

Observation 38b223ef-d41a-4ad1-9634-bdaa3dc64c06 · outbound

This paper cites Main Characters.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Main Characters

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.739911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.454530Z digest=sha256:983227c29ac862ffcb1aeb5fa347cae3aaff6a88ba87e32208ddcc50c090573d

Pith citing papers

Observation f7a7c12c-febf-4eba-9be0-d0a904557c90 · inbound

ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop cites this paper.

ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:53:11.599109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T19:51:31.657802Z digest=sha256:978e4d2e01600432aae6c4d2006110fb441a561c11348b60e3f00c347910dc37