Pith. sign in

Paper Citation Record · LEDGER

ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2302.13848.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.13848 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:43:51.844097Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T08:15:33.740930Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 59638623-0b99-4598-87f9-53b98ec51a93 · inbound

InstantID: Zero-shot Identity-Preserving Generation in Seconds cites this paper.

InstantID: Zero-shot Identity-Preserving Generation in Seconds ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T21:02:41.479856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T21:02:41.399183Z digest=sha256:cf0e95b97b92837ec0764ec9a294f9e2b6cae6e8ba1293751f25def9d4904b0b

Observation 6cdd4fc8-b4e4-4bd1-b7d3-b37a6a950138 · inbound

AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks cites this paper.

AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T14:03:28.547482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:03:28.547482Z digest=sha256:061fa28e8137ecac220c66e2642107a77eb8802153aac71218f5ea4f070716cd

Observation 6dce1375-f810-4626-8d3e-b4e101d19de0 · inbound

Imagine and Seek: Improving Composed Image Retrieval with an Imagined Proxy cites this paper.

Imagine and Seek: Improving Composed Image Retrieval with an Imagined Proxy ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T14:02:55.883966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:02:55.883966Z digest=sha256:a4d7d22859fb895041def224accaa93bc8859427232d8c7bcba71d0a698ec4c4

Observation 8522b030-20fc-49af-9d69-6cb3556527ae · inbound

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion cites this paper.

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T11:36:17.380384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:36:17.380384Z digest=sha256:eb266d3d79ed79372a77b312fad15d9ce0b2f11af863256340abec51316f5136

Observation ea9ec3d0-297b-4b4d-aa43-3003e885faa9 · inbound

SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation cites this paper.

SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:37:44.687479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T08:37:13.529654Z digest=sha256:38baed6c9da0581c5724fc1c6067e0c5bad8dc207342e8037133270e2d2966de

Observation 9ab75395-0db3-43fa-88b7-226aaedf1e72 · inbound

DIVE: Taming DINO for Subject-Driven Video Editing cites this paper.

DIVE: Taming DINO for Subject-Driven Video Editing ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T22:33:37.157503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:33:37.157503Z digest=sha256:68c025ee7f949af5dd7fa9ae4a5b0b8b980d11fe864a460339b671c83abc253e

Observation db47e8a1-1982-4929-bca5-aa81b5b3b573 · inbound

SUGAR: Subject-Driven Video Customization in a Zero-Shot Manner cites this paper.

SUGAR: Subject-Driven Video Customization in a Zero-Shot Manner ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T15:56:59.857766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:56:59.857766Z digest=sha256:4e046fb0bfae7b7129b0e384c18a0247e5e7b8bdf2ba5e8f7ea7b3c5f91dca34

Observation e46424b5-4ece-438d-b9b1-15237378cba7 · inbound

Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization cites this paper.

Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:43:51.844097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:43:51.844097Z digest=sha256:85eb6f474e208f74561c864ca3b6e46e0859b625a99e96a07b87b4179607477f

Observation e0b415a3-d8e1-4072-bfb2-0e3918fbb44d · inbound

CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up cites this paper.

CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T10:54:30.991426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:54:30.991426Z digest=sha256:107f7c23024d41c55b0229fcfb84516b33f0b0de84d31858b0a2585198f160a8

Observation 773761ff-eac1-4634-9cdc-673cc12e259c · inbound

Dense-Face: Personalized Face Generation Model via Dense Annotation Prediction cites this paper.

Dense-Face: Personalized Face Generation Model via Dense Annotation Prediction ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T05:04:28.119025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:04:28.119025Z digest=sha256:2df674106ebea08ecbdddba67cc0ef5681104eeb11488d56eaae11a259c4ca6e

Observation 09ec6af3-1827-49b0-b942-f86982ab3923 · inbound

VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control cites this paper.

VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:14:40.681243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:14:40.681243Z digest=sha256:af4ec6b86ed77c9a46a1eb1ed93b4748219c5e6e19226bff0a10ab5a7835ab5b

Observation eb7becec-6434-4168-85d5-f61ee20bd04b · inbound

SceneBooth: Diffusion-based Framework for Subject-preserved Text-to-Image Generation cites this paper.

SceneBooth: Diffusion-based Framework for Subject-preserved Text-to-Image Generation ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:07.792817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:07.792817Z digest=sha256:17d4412a22bfd3f21000ea314a4a96b0553e3a94167fe7e03ff02bc37757e913

Observation 6dd289fe-89c7-4668-b203-073125092ef1 · inbound

FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors cites this paper.

FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:55.669450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:55.669450Z digest=sha256:9fd89468e903ce92e9dcacfa935937eb84238b50c2d49fdba3e81840d5c098ba

Observation 4ad4c8e7-a019-47a5-acc0-8a2054b57920 · inbound

AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation cites this paper.

AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:21.605154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:21.605154Z digest=sha256:36dfedd463eaa6cb5737769d4aa6deaf4339930bf7a29ed1000df8953d0d0637

Observation 43066a6f-3058-4fa9-8fab-dc8d4c3bc853 · inbound

ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions cites this paper.

ComposeAnyone: Controllable Layout-to-Human Generation with Decoupled Multimodal Conditions ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T17:30:47.961855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:30:47.961855Z digest=sha256:4998ea00901bf3357d208bbdba5d13e13272adb22ea8cd1db3247ef49e1bd01d

Observation 648dc1a1-1a9c-4e34-9a3f-2583a7f3dbe5 · inbound

Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis cites this paper.

Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:53:28.602302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:53:28.602302Z digest=sha256:406272b555492c6f7f70018919f29f7ebb50041af41fddb750e15a5d89e33940

Observation 78dd465b-dc7f-43a5-8e14-ee443a0b0e8b · inbound

Parallel Rescaling: Rebalancing Consistency Guidance for Personalized Diffusion Models cites this paper.

Parallel Rescaling: Rebalancing Consistency Guidance for Personalized Diffusion Models ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:22.824038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:22.824038Z digest=sha256:aa8dfd20b19e8e1909bbe992bd5b7dcaeb0be343663e5127a75b5c69ce0407c2

Observation 02bda76d-e424-4691-969f-c0f4595b5dd1 · inbound

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation cites this paper.

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:02:10.375261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T07:59:49.398271Z digest=sha256:999728cff23aed4c01b9f2791f056cd553be1767d8c090976391d8f07d4e9a11

Observation b423358a-2ed6-4ab7-b5d2-9d8f889729f1 · inbound

UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries cites this paper.

UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:15:33.744173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T08:14:28.540629Z digest=sha256:26a6e0947941aa627dd23b3d17c793241aba1c37c028a9d69ff5fe57c60e71e9

Observation 0e8bbf22-ba05-4b4d-9ada-d4ee2c5d2f5a · inbound

Story2Board: A Training-Free Approach for Expressive Storyboard Generation cites this paper.

Story2Board: A Training-Free Approach for Expressive Storyboard Generation ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:09.318239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:44:09.318239Z digest=sha256:54a979a04ae2e1e79f44c3a91f5908f3776e184c9f6ecd294124653e73f20f6a

Observation a602dc1d-904b-4d58-b7d0-4b2ba2872fe5 · inbound

LooseRoPE: Content-aware Attention Manipulation for Semantic Harmonization cites this paper.

LooseRoPE: Content-aware Attention Manipulation for Semantic Harmonization ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:08:04.432693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T16:06:20.797660Z digest=sha256:aa8a31aa7486baf8ec49641bfc8075e46f5d8a19ed960dc6c4ec2638a602e60d

Observation b0b03caf-83d0-48ad-bdef-57cb6a70bb11 · inbound

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models cites this paper.

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.851674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T05:47:49.309008Z digest=sha256:aa72b6319643577990719ba518504f29ce4ba87bbf8f3e5ed60a1cedecda9089