Pith. sign in

Paper Citation Record · LEDGER

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models

As of 17 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2509.04446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04446 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:32:15.454530Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T19:51:31.657802Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T19:53:11.597613Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ec10e816-9bec-4557-833f-bc87b516bc56 · outbound

This paper cites GPT-4 Technical Report.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.245097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.245097Z digest=sha256:d320e9448f0fa7ee7746974b3cb29dfdce3980ce99f8e35c1db2e918201744d5

Observation 75cfc74d-c150-493a-8165-c7ef7faf0426 · outbound

This paper cites ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.250967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.250967Z digest=sha256:bea2b2284d960e70e4fac4c2abc80a95dce447c1a5dc4458480adcd75c15ddcc

Observation e99ba86b-632f-454a-9b4a-2a9cb832e62a · outbound

This paper cites Blended Latent Diffusion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Blended Latent Diffusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.255493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.255493Z digest=sha256:fcbf4b9d2a654a1bce70cd0dee8867a9bcd70d4f837f31d22f7d60022b281bb2

Observation 20b114b9-1580-4fe2-9611-e43ddc54dd4e · outbound

This paper cites The Chosen One: Consistent Characters in Text-to-Image Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models The Chosen One: Consistent Characters in Text-to-Image Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.259614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.259614Z digest=sha256:4b64d543534db0a79032e952aa6bf241f8632c735955cea799d3f1c22c8ff5f3

Observation f56c5461-407e-4037-9d84-c9e1b26ee601 · outbound

This paper cites SEGA: Instructing Text-to-Image Models using Semantic Guidance.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models SEGA: Instructing Text-to-Image Models using Semantic Guidance

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.264293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.264293Z digest=sha256:149eb27c0b9c91d122838e65824e4e4185cc236a1ba9393d055fad71e1d1ce72

Observation 567bd695-be3f-485a-92bc-78e9891a4310 · outbound

This paper cites Ledits++: Limitless image editing using text- to-image models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Ledits++: Limitless image editing using text- to-image models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.095550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.268891Z digest=sha256:19f15a85bcbd4ca2dadcf2921ad2959725ecb71dc11cfde04018578f580ddf4f

Observation 4bdf3bb4-eb11-4b9a-a585-9812b8da4bd2 · outbound

This paper cites InstructPix2Pix: Learning to Follow Image Editing Instructions.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models InstructPix2Pix: Learning to Follow Image Editing Instructions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.273719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.273719Z digest=sha256:da043c3a023ef97b41d71879e62cd4b14b53e3131971fbc96b8734812bf826ab

Observation 05f043fc-a80d-404a-820f-de9b06d98443 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Emerg- ing properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.279215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.279215Z digest=sha256:c376dab43439e1261a930cc9f77c644e90eb3ad28abc7de2cedd1b9f4edaab4c

Observation a08540ac-7be7-4f99-a965-535f544bbce4 · outbound

This paper cites AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.284324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.284324Z digest=sha256:cedc085c3fe428463ce911915d9fbc96cf9b7e271a0342adcc9ee957dced9605

Observation 72e2eba5-85a0-4c0e-9dac-c3c9333c742d · outbound

This paper cites TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.289174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.289174Z digest=sha256:ecc78e306d23f9e7adde04883b708896b9636f8b1eb552a719f2133eda9adc96

Observation 023ed9d8-31c2-431f-afb8-441390605b70 · outbound

This paper cites Yolo-world: Real-time open- vocabulary object detection.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Yolo-world: Real-time open- vocabulary object detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.065920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.294108Z digest=sha256:1a64b2c7b0188bf190d3182346950c1778045a4f9b0ab75cfe016c70754f3b84

Observation abd3b556-d7d3-4824-b760-16f7c879e7c2 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.299028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.299028Z digest=sha256:b0f9912e71b80ad61d5c5457694708d27d21de342232920f723f667af0b6a139

Observation cdfca3c4-af59-468a-aff9-9057c7e60484 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.303626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.303626Z digest=sha256:3c266521e6894a861c90a70696be8617ddb9ccc678bb15b1ab96baf6cfc33071

Observation 1e24d965-133d-4d7d-b415-3e2ceb5a864b · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.310852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.310852Z digest=sha256:d2d2cb2fa79a30b4aa34852b08000360af808f7b244e74993c11d219e545e988

Observation 7f1d5e92-0182-4b6c-bc51-96102be9b408 · outbound

This paper cites Prompt-to-prompt image editing with cross attention control.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Prompt-to-prompt image editing with cross attention control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.316405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.316405Z digest=sha256:80b76bd473e713f561bb0d97a8ee4aab0be08bbac44fe594424410bc906e484c

Observation 3740d32d-9f22-4d50-a3b2-f79a0eaa0f06 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Denoising diffu- sion probabilistic models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.322765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.322765Z digest=sha256:e0dbf418a602894c1c3696b6a1859ee408d823ddbfa1090a6c440e9df09285cc

Observation 91c1462f-b309-44bf-880a-6e537600a9b3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.328415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.328415Z digest=sha256:9add89a38a6a37290899d1c6ae5e1991900a5c8fbc62204eccf8592143b2fdad

Observation 5ae3d5ec-1f06-4366-be0e-93eb42d6e43f · outbound

This paper cites Zero-shot generation of coherent storybook from plain text story using diffusion models, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Zero-shot generation of coherent storybook from plain text story using diffusion models, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.018633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.334093Z digest=sha256:ef9d0aa70f4a97522d93d6474129eb39fad19343398b2e2a4b7a79c85135257c

Observation 65fdaa25-bca4-4476-ae2b-64723c054408 · outbound

This paper cites Rehg, and Pinar Yanardag.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Rehg, and Pinar Yanardag

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:16.004941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.338889Z digest=sha256:a2dfe80a5e82723c1eff6a335560bc1a09e5372292ee80d4383ac9a9d77a2d07

Observation 104fb183-4600-42f1-a24a-38400b5e14fc · outbound

This paper cites an unresolved cited work.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:15.991864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.345203Z digest=sha256:0ca0a123ea13afada84786cacd0b4a2d2ff2f3e7dbe25e50c6b0e674516fc096

Observation c35c054c-e6b4-4508-80a6-aaacac436d56 · outbound

This paper cites Intelligent grimm - open- ended visual storytelling via latent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Intelligent grimm - open- ended visual storytelling via latent diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.979050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.349572Z digest=sha256:de216bb012c523fc1f98f3dbbde0f0fbcab41fda0c3dad508272bf6ba5cdaae1

Observation 0bec2a1c-7fba-4523-ae6b-a34d0af0ccfd · outbound

This paper cites Compositional Visual Generation with Composable Diffusion Models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Compositional Visual Generation with Composable Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.354327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.354327Z digest=sha256:9af701c806e214342f96ee616e2769e832af9abca0ae75d5c5fadd07a6655f90

Observation dfc7d8d8-0c02-4956-bdc2-44da5466c05a · outbound

This paper cites One-prompt-one-story: Free-lunch consistent text-to-image generation using a single prompt,.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models One-prompt-one-story: Free-lunch consistent text-to-image generation using a single prompt,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.964746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.360081Z digest=sha256:34201a16535c088c78f8f2992c5c605f1c580a673daf1c8146a089b223e129c2

Observation 0d48eccd-937b-49a4-99b4-a75eac916e2a · outbound

This paper cites Storydall-e: Adapting pretrained text-to-image transformers for story continuation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Storydall-e: Adapting pretrained text-to-image transformers for story continuation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.364085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.364085Z digest=sha256:168b50533acd209db520a66293b861b20e26725a33c931ec3292c2d8d43b636d

Observation a20ec3eb-bb09-4089-ac12-fece52d407bc · outbound

This paper cites Synthesizing coherent story with auto-regressive la- tent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Synthesizing coherent story with auto-regressive la- tent diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.368858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.368858Z digest=sha256:2c53d3ccb13694ecfef3a0ec1290ec0bbed89732229473987571a51090dd5bea

Observation 6fd444aa-4d76-48ed-bafa-c32787b56865 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Sdxl: Improving latent diffusion models for high-resolution image synthesis, 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.928794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.372979Z digest=sha256:f90339091184d73309f7f5b92960e294fa29472906b8e267b6287d47073cfa36

Observation 7f431f05-1ad7-48e9-b8a1-deeba02e6abb · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Learning transferable visual models from natural language supervi- sion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.376979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.376979Z digest=sha256:2edddb1f99fc01c1a26705dfd229ef3ffc782a16b00a15a0ead8863bf3e80a60

Observation af38eb90-2aa1-4dc0-86ee-1560d61d0040 · outbound

This paper cites Make-a-story: Visual memory conditioned consistent story generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Make-a-story: Visual memory conditioned consistent story generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.380905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.380905Z digest=sha256:fb4d9de08cc1b35974d9286ef272b8319fa484f5b76c01f7145cc78843bd6075

Observation d06b27e6-3103-4909-a357-87d18bb466b5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.385321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.385321Z digest=sha256:34eb45f436c7ef0b8c0a5169278590a9373883c60822b6ac0ade1e7c716878f9

Observation 8772e8a5-3c82-4e16-abba-4c9bacd3e796 · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models U- net: Convolutional networks for biomedical image segmen- tation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.388970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.388970Z digest=sha256:efa85f151fc3977d7113dbb7dcabd4dd9f32892fe157f9ae6910986067e1bef3

Observation 31cb85f4-7843-4490-8a8c-8b02bffac400 · outbound

This paper cites StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.393267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.393267Z digest=sha256:68b5043aeaae10492e6d07d90d8126e3ccda8804c39cd5483d32f269fb4a3bd4

Observation 47cd6ce4-141e-45a4-a9d4-5c3bd58edf33 · outbound

This paper cites Training-free con- sistent text-to-image generation, 2024.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Training-free con- sistent text-to-image generation, 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.868777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.397679Z digest=sha256:8f7367282512aaafd2cf637cc03c69566b157921e5b55a8d5c39d0258451c9c2

Observation ebdea003-d86f-4b7d-982f-f65ec5672f68 · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Plug-and-play diffusion features for text-driven image-to-image translation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.855287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.401911Z digest=sha256:8505c0b9216f69415b9b2f3fd9a162376dbcc4f94351f0ae812e9dbef9414a25

Observation 466afb1c-8997-48ae-a3fa-d403378039d2 · outbound

This paper cites Unitune: Text-driven image editing by fine tuning a diffusion model on a single image.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Unitune: Text-driven image editing by fine tuning a diffusion model on a single image

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.841121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.406601Z digest=sha256:0a16c5dd91f6eb00d8d5552f4a35b3f215a9c10843fb8a8b8a15cbc99d55784a

Observation fb5c0268-3b07-4b19-9a57-ad596bd543a4 · outbound

This paper cites AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.411330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.411330Z digest=sha256:090507cd2ee1f526a61ac96352d3e3d73349fd819b1b6535737878fb9aed6aa9

Observation 7f7efe33-3ce6-4ada-84d5-a809d3f62ccd · outbound

This paper cites Nerfiller: Completing scenes via generative 3d inpainting, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Nerfiller: Completing scenes via generative 3d inpainting, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.824288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.416686Z digest=sha256:6943699fef7b8efd8e726243fbebc73800584b9e4e7c50850921a99cbdd47dbb

Observation 5a27e0b5-1a4f-48d1-b6e2-71a169312b34 · outbound

This paper cites Efficientsam: Leveraged masked image pre- training for efficient segment anything, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Efficientsam: Leveraged masked image pre- training for efficient segment anything, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.799095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.422393Z digest=sha256:3d3d9916699d59857de04c8c2e4855847b17a5f4903c0ed383a798a2a55d0ca0

Observation d6ca206d-e27c-4a3d-9ecc-e3e62a9f4729 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.429255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.429255Z digest=sha256:8d655b978324a084380705535da441ecd2bdfc5a380ca0eb81478fa09d156868

Observation f74521c4-473b-44d3-a342-98a6c9f39be7 · outbound

This paper cites Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Ip- adapter: Text compatible image prompt adapter for text-to- image diffusion models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.780780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.433717Z digest=sha256:92607a90abc235c1fcc3c02134aeb975d9f69243d5cc25a2ee145c31ccade5ee

Observation 49114b78-55e0-410e-9fca-66092e9fe2fe · outbound

This paper cites Adding conditional control to text-to-image diffusion models, 2023.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Adding conditional control to text-to-image diffusion models, 2023

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.438012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.438012Z digest=sha256:9664a05da8c1e5e42c1ff35b2fac1849e0a21806f521174b84cb2fe5227ff8dc

Observation bd67c397-40f9-4619-b1ae-87ea982011ac · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Adding conditional control to text-to-image diffusion models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.442490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.442490Z digest=sha256:d2f6a0365772ce467dfb892ad1606f53e3fbe2a3aa81cdffce4be7117244c426

Observation e37a885a-6bc3-474c-b000-8133b8e551e3 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models The unreasonable effectiveness of deep features as a perceptual metric

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.446611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.446611Z digest=sha256:79af65cfc4c8f515e03d67f114fc1ca0cf6d0091a48cfc9f39f7809abcb66e77

Observation 450a8d1f-0422-441c-9e18-783561a3fae1 · outbound

This paper cites StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:15.450281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:32:15.450281Z digest=sha256:beff148097c08145abc11f98ff03df4b64ae8657049bb73b7313e8e3945ad682

Observation 38b223ef-d41a-4ad1-9634-bdaa3dc64c06 · outbound

This paper cites Main Characters.

Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models Main Characters

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:15.739911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:32:15.454530Z digest=sha256:688e94520170a25827e59b6b602700b7cc8eb7e4da092000021e44ddf0ad5ed0

Pith citing papers

Observation f7a7c12c-febf-4eba-9be0-d0a904557c90 · inbound

ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop cites this paper.

ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:53:11.599109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T19:51:31.657802Z digest=sha256:d10dec3814be79a161a1218ad908deb65cb603af592e307cba2dff852ac54fa6