Pith. sign in

Paper Citation Record · LEDGER

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization

As of 23 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2508.03735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.03735 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:47:00.786871Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:38:37.640445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-14T04:38:39.565150Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact6
  • verified fuzzy20
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 018a1c90-d249-4928-94ef-fd26d2c6f89c · outbound

This paper cites an unresolved cited work.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:47:01.972732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.509179Z digest=sha256:2ad31aee2b3ee7204ebcc59bfcc8aef65eab47de55b609da2456702bb4cbf168

Observation a092c6e8-e2f2-4182-b4ef-60315905d343 · outbound

This paper cites Kandinsky 3.0 technical report, 2024.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Kandinsky 3.0 technical report, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.930235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.515460Z digest=sha256:3b8ea016806fa50c4ff91b3fb63eaef24e39ef6c04178de32f54eafbb4acca17

Observation e6b0c73e-a936-4eb4-acd8-a936012f0ab3 · outbound

This paper cites The chosen one: Consistent characters in text-to- image diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization The chosen one: Consistent characters in text-to- image diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.911168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.521730Z digest=sha256:b729c18bb146acdcf8bdb434dce9a488d2a7b894d9a8ced6ba4deb9b65799e96

Observation 07287fd2-36b5-461f-bb83-b7d6b51313a9 · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.530198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.530198Z digest=sha256:19eea23abbd7f28d78fe57efbfb62c92efa05eff34fa36d8ba2f542529d66eab

Observation 14428dd0-dc76-4f88-b120-b1fcddd30d81 · outbound

This paper cites Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and edit- ing.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and edit- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.891538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.536152Z digest=sha256:328db37a6f6c74537cdf03da652b64bfd1bc145cc87c759d8214c14119b02711

Observation 22e57c8c-9606-492e-9227-c891e2fe8fd9 · outbound

This paper cites DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.541286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.541286Z digest=sha256:e968b3b40443f6dccc7dfb8e24a44778e72ba6b870e0f86b6d24e10b543b278e

Observation 25999249-f595-49b7-89a0-68e4bd432bcb · outbound

This paper cites Taming Transformers for High-Resolution Image Synthesis.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Taming Transformers for High-Resolution Image Synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.547306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.547306Z digest=sha256:fc077851db405e2e07fca2f145dd815a8dc721dba3c73b158575af07559290db

Observation aec35b91-984b-4222-b953-b25f49f1e8ac · outbound

This paper cites Improved Visual Story Generation with Adaptive Context Modeling.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Improved Visual Story Generation with Adaptive Context Modeling

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:01.404398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.552932Z digest=sha256:2784f8cb00b8d770b178486bf05d9f240e59d6bef9615569b46e8c25af4556ee

Observation adce62b9-8e2a-4e76-ad60-b0d681ffd213 · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.560554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.560554Z digest=sha256:5518389f62f9f9f15d0e0c84f0a7b35f81d3ddd65aaa83d629d6654af35fdf0e

Observation 6571d272-d148-4d71-a8e7-fd11c2d026c3 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.572791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.572791Z digest=sha256:3300b0b85d17e5efbc00d0034c96bb743996dac553f727966403d3b1887fd61c

Observation b73d4ce0-5bfe-4ef4-9fab-81c3abb8e661 · outbound

This paper cites Encoder-based Domain Tuning for Fast Personalization of Text-to-Image Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Encoder-based Domain Tuning for Fast Personalization of Text-to-Image Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.577992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.577992Z digest=sha256:6b2c2a4c5657630de19c28d3400c7a5915bcc101ad7b978afc8995bf7d0d034c

Observation 9ade7593-8cb9-4394-b308-dce86cbb59de · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.584961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.584961Z digest=sha256:2337831f50af62b65a429e32fda2a8fa048e9f4a2802c9c36288cea7bcb1be92

Observation 79c0c6d2-a1c0-47e1-8b69-b24a56af1539 · outbound

This paper cites Mix-of-Show: Decentralized Low-Rank Adaptation for Multi-Concept Customization of Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Mix-of-Show: Decentralized Low-Rank Adaptation for Multi-Concept Customization of Diffusion Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.589744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.589744Z digest=sha256:f91c05252e6c9ae106e1f80136455d633ce003742577b434b815d7dad9e56d02

Observation a2e2c82a-bfd1-4f2d-96ec-257236ae6a7b · outbound

This paper cites DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.594729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.594729Z digest=sha256:48e223c32f841719d97a01818bcea7e67c3451d8ea5bdef172dfc914ba70b59b

Observation c7bff57d-9a73-403c-af65-7851b189d290 · outbound

This paper cites AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.599692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.599692Z digest=sha256:dc0b06eb18eacdf18b96247a810ab5427cc643962ff30fafe36a65c6ad2b5026

Observation 1645d4bc-b7a3-4586-979b-f79f8d788a03 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Denoising Diffusion Probabilistic Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.604372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.604372Z digest=sha256:045b9c23db052052b2e1b899539038484a3e4abff7b8d48358d383c81b00b17c

Observation cccccd20-6f7f-4c21-8642-fd32fcaffd38 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization LoRA: Low-Rank Adaptation of Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.608886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.608886Z digest=sha256:64226269e44b6738c9a7ddb467d3125bd87f5f8ba41c4cd35e7bd8042a2b48f4

Observation 9dd73ac3-cf52-419b-98c5-bc0a72023aad · outbound

This paper cites ReVersion: Diffusion-Based Relation Inversion from Images.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization ReVersion: Diffusion-Based Relation Inversion from Images

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.613396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.613396Z digest=sha256:3805caf145d626da2d3c99d8f40d665c31a26ccbc029b63a6432a6409b2436d8

Observation 8ab4e140-2421-4a4d-9470-9f7527e7e5e6 · outbound

This paper cites Zero-shot Generation of Coherent Storybook from Plain Text Story using Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Zero-shot Generation of Coherent Storybook from Plain Text Story using Diffusion Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:01.191150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.617823Z digest=sha256:5b825ef4ef69b20064086beb10c850bad3a5ff98c40d4551e09310e322ab6102

Observation 8045024f-401d-4538-8902-da1d7a8e78a5 · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.868230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.622363Z digest=sha256:aa5dd8500007f2b765ca97220d090c7fc372e285aa2c8ed335590a3da5d9b7a7

Observation 3f25451f-be29-4d99-851b-4c9275a5110b · outbound

This paper cites Multi-Concept Customization of Text-to-Image Diffusion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Multi-Concept Customization of Text-to-Image Diffusion

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.845910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.626882Z digest=sha256:6f8f2916fabc30d0b0ad67e007520d13f44dd22ac324768d3e3292be457b7f13

Observation 9132284e-050c-475a-96b0-468086712568 · outbound

This paper cites BLIP-Diffusion: Pre-trained Subject Representation for Controllable Text-to-Image Generation and Editing.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization BLIP-Diffusion: Pre-trained Subject Representation for Controllable Text-to-Image Generation and Editing

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.631610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.631610Z digest=sha256:05c401649ab436f6b4403bb198b0be9afaf89778d29cdb4a89249c9a5f587a0f

Observation e7251611-9c29-4688-998e-b213aec37d65 · outbound

This paper cites Intelligent grimm-open-ended visual storytelling via latent diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Intelligent grimm-open-ended visual storytelling via latent diffusion models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.824993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.636587Z digest=sha256:2264dd8a6da76f3b76d956197fad57b04ec9ad83165fc7b1ea0d9de7411b0455

Observation 264ccd52-ab94-4a42-8e16-690223903349 · outbound

This paper cites Intelligent grimm - open-ended visual story- telling via latent diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Intelligent grimm - open-ended visual story- telling via latent diffusion models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.802044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.641246Z digest=sha256:daca635ea24b7e1abfc0301f1faba0d1506cebb2052bcc99af8a992dfd481591

Observation 2f0714e6-472f-49c8-a2d5-b60d2b50ed1a · outbound

This paper cites Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.776334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.645517Z digest=sha256:c10d36856f276aa24953ef1a0f3a524f3e736d219420a47efa6e9e6ef3257b4c

Observation f27955d2-1ff9-4e2d-a38a-b1021b68719d · outbound

This paper cites One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.650209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.650209Z digest=sha256:1a4f1fd45c240537046fdc3294cd0a869b700db371b463aee77682cf79e3a807

Observation 81b192b7-dec2-4647-af9d-92ae36fed566 · outbound

This paper cites Summary of chatgpt-related research and per- spective towards the future of large language models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Summary of chatgpt-related research and per- spective towards the future of large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.753452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.655132Z digest=sha256:15aa3c96baaa744efb797e120d317c451d2f376dc3bc7dacc616b671e7d0e46e

Observation fe898dac-3265-46f3-ace2-620d0d3f4aa3 · outbound

This paper cites A threshold selection method from gray- level histograms.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization A threshold selection method from gray- level histograms

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.730940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.659585Z digest=sha256:c5e26e7dbb2aa1c43fa446c79509465593b58aae9cc785585dcf380fcf5b5f4d

Observation da5d112d-a56f-41f2-b7ce-5f8d647ad88a · outbound

This paper cites Synthesizing coherent story with auto-regressive la- tent diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Synthesizing coherent story with auto-regressive la- tent diffusion models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.708130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.664058Z digest=sha256:b2040c1aca7e40081a8a6f617075bf8ba31db9026f912a34066b987bc9d4dea0

Observation a6560ec4-8a5c-48e8-a5a2-1eceac642c53 · outbound

This paper cites Localizing object-level shape variations with text-to-image diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Localizing object-level shape variations with text-to-image diffusion models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.681051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.668418Z digest=sha256:d7b587c67293c6fda235d956557b44bacdf094031ea479fd709b04747d1ac63d

Observation f0e30f7a-20a3-49c9-a3f4-f93d9e897865 · outbound

This paper cites Orthogonal adaptation for modular customization of diffusion models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Orthogonal adaptation for modular customization of diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.660916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.672722Z digest=sha256:3f637829715e0efba3e612c37a88a0e6f2939466e207d7e1bb79143e1767412d

Observation 8fc822b2-2a81-418f-890d-e800765249d4 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.677299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.677299Z digest=sha256:499fbbac42a8315794680e8cf4781287256a8d9ceaf16549851054e9943c2ac3

Observation fcbbc347-f342-4fb1-a131-003d70b4b2d6 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Learning transferable visual models from natural language supervi- sion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.643135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.681785Z digest=sha256:75126aa651ccab548e2c167e335b99a0431e388ae7b94b2fcf926f7e5ae9f2ca

Observation 13ecdec0-3ffb-495c-99f8-0504733fb826 · outbound

This paper cites ConceptLab: Creative Concept Generation using VLM-Guided Diffusion Prior Constraints.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization ConceptLab: Creative Concept Generation using VLM-Guided Diffusion Prior Constraints

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:01.098750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.686373Z digest=sha256:755a44b0834587f93a51a86391863c099f67d042e70616100bb5a9ef056df20c

Observation b6356c3a-2c27-4333-8da4-7fa65de6a571 · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization High-Resolution Image Synthesis with Latent Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.691048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.691048Z digest=sha256:4dbdffc78fa6f71f898488aee2ba771b8fe1b6db62641648d1cf9ad66c551862

Observation 4d62dc61-9338-45a2-883d-7edcf1017cfe · outbound

This paper cites U-Net: Convolutional Networks for Biomedical Image Segmentation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization U-Net: Convolutional Networks for Biomedical Image Segmentation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.695807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.695807Z digest=sha256:3dfe0a63db3a78fad38de2aa499d79a4b0f08e65d7b8e0959e6850634ca13fb4

Observation c0f78d9a-a2e8-422b-8f13-af83bf97a927 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.622221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.700353Z digest=sha256:785f60efd402023522fe23a312832644f16d7275559578ee50b08b0ac9c0da5f

Observation 04d00cac-74fe-43b7-9df6-2dfc3d49f82c · outbound

This paper cites Instant- booth: Personalized text-to-image generation without test- time finetuning.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Instant- booth: Personalized text-to-image generation without test- time finetuning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.601868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.704766Z digest=sha256:ceee5d2163bb7ca773626d3af55cc48136fa15100b2edf2a2e0c6daef7b34b41

Observation 963e731c-b1f4-4e35-b139-45dd52c818fe · outbound

This paper cites Make-A-Storyboard: A General Framework for Storyboard with Disentangled and Merged Control.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Make-A-Storyboard: A General Framework for Storyboard with Disentangled and Merged Control

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:01.036647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.709158Z digest=sha256:663a987d68a48bcc4a1fcc098be54b374fe5c3233fd9fb2338d43b723456b7a9

Observation 866eb49c-9464-4ecc-8513-26b1eb262209 · outbound

This paper cites Emergent Correspondence from Image Diffusion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Emergent Correspondence from Image Diffusion

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.714088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.714088Z digest=sha256:d555f8dd6041f9c33f5bacf1de2b12b803b7568b5dc743f5978e859629b36caf

Observation dd87f771-1aa3-43f6-be44-ec0f7206468c · outbound

This paper cites Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.718928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.718928Z digest=sha256:fcea024924a599903cba26596b48efbca50d80fd7fbcd384e6e9289894f5c554

Observation e5e32d00-8f2b-43c0-a51f-58dea7038eb0 · outbound

This paper cites Training-free consis- tent text-to-image generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Training-free consis- tent text-to-image generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.724877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.724877Z digest=sha256:0d3e6dd12fc8349f7abc5744776d2eec2d8fe4b283717f2e16fa0ec8dd97cef2

Observation 29a66fdf-c594-44a5-bebc-debb565025e9 · outbound

This paper cites OneActor: Consistent Character Generation via Cluster-Conditioned Guidance.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization OneActor: Consistent Character Generation via Cluster-Conditioned Guidance

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.729381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.729381Z digest=sha256:ec32974411ede3fc0c484b857cd8a34623b5b6f60e2c7c0c78a9294a5797da06

Observation 3dfeb8c4-b9e4-43c9-ab35-240b706e871b · outbound

This paper cites SpotActor: Training-Free Layout-Controlled Consistent Image Generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization SpotActor: Training-Free Layout-Controlled Consistent Image Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.733948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.733948Z digest=sha256:3e421a00457bc3f2f8e6888bb3f988e8705614e120af01d88512fb1df0f1cfd5

Observation 0ae00d22-5031-43d9-94e4-0f6e74db6597 · outbound

This paper cites CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:00.936339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.739136Z digest=sha256:9672a6c8fea17e6082a36512b1e44cd5b52ba9f77b4ffd613b317726bcac4939

Observation f37fa062-1a6c-4f4c-88c0-2aab1af6fd86 · outbound

This paper cites Elite: Encoding visual concepts into textual embeddings for customized text-to-image gener- ation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Elite: Encoding visual concepts into textual embeddings for customized text-to-image gener- ation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.567203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.744144Z digest=sha256:0288fde8c610a1efcbe7ffb47fb030fdd9da447c51b43e0a59049635c75c03d3

Observation bdeafe01-89f5-4e18-acd9-b30695a6cbd9 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.749222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.749222Z digest=sha256:2aa0715d8001274d6d418195857d499ca30b87fdd325f7e110b2a05c85a299c1

Observation f698b1de-2d72-4ff7-af37-a1589c6f4d32 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.754074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.754074Z digest=sha256:048775c5d86340f73dbefd692725ae66140c1e4c356f983f08ff19167d25ff57

Observation cd4285ba-af4d-4ae7-875e-5b3e471e9232 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.758937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.758937Z digest=sha256:4efe02fe2d1bcfc21d7ccad231223534df58177177e3bba48c5957a028c89255

Observation 8c04c3e6-40ab-41dc-9dc1-ebaf43e7b606 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization The unreasonable effectiveness of deep features as a perceptual metric

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.534001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.763414Z digest=sha256:455d0270ef1b36a3c0ac1b491681062d7fc4d531e196f6f9f4a9efeaddd3048d

Observation b7f0222d-27df-4943-8c72-d96e33326e93 · outbound

This paper cites Inversion-based Style Transfer with Diffusion Models.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Inversion-based Style Transfer with Diffusion Models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.510829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.768283Z digest=sha256:8d75636e137cbccc16b74b63475d7e401c5aea86503aa89596016b1e8b7c910d

Observation a53f593e-5185-4b55-950f-900fe7ebcae4 · outbound

This paper cites ContextualStory: Consistent Visual Storytelling with Spatially-Enhanced and Storyline Context.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization ContextualStory: Consistent Visual Storytelling with Spatially-Enhanced and Storyline Context

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.773062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.773062Z digest=sha256:428116d9046e1da8bf1f0897635b686f57f6ca62c3716d9d5d79de8fc3315743

Observation 4a11a7ff-9c4c-42d3-a04a-eb7f79ee8c6a · outbound

This paper cites Storydiffusion: Consis- tent self-attention for long-range image and video genera- tion.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization Storydiffusion: Consis- tent self-attention for long-range image and video genera- tion

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:47:01.492983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.777743Z digest=sha256:9fe5dddb957de0c8d5759dbd9e88cbc3b4c33f6b2e6e70d31594cd7a7fba5545

Observation 1b3a3a08-608c-4ac7-ab13-0da71aa48377 · outbound

This paper cites StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T10:47:00.781995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:47:00.781995Z digest=sha256:1808a235ec874a094cd5248de38dd05a706e271d7b628d6d0925a6cd37b6922a

Observation 5a94c37b-7a28-40e7-816c-932cdd1478f0 · outbound

This paper cites DomainStudio: Fine-Tuning Diffusion Models for Domain-Driven Image Generation using Limited Data.

StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization DomainStudio: Fine-Tuning Diffusion Models for Domain-Driven Image Generation using Limited Data

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:47:00.838402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T10:47:00.786871Z digest=sha256:5ed861edd5f469d697c99b735e5613b35c2e7b8a043cd6028282619ea5d42367

Pith citing papers

Observation 82acd1da-56b1-429c-916a-bb4bb211287d · inbound

InstructionCrafter: Generating Consistent and High-Fidelity Visual Instructions cites this paper.

InstructionCrafter: Generating Consistent and High-Fidelity Visual Instructions StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:38:39.637125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-14T04:38:37.640445Z digest=sha256:f2c2870965524fd3730f7dc434b7bc069cddf2b2a3e466801ed9a883fe3718fd