Pith. sign in

Paper Citation Record · LEDGER

VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2412.20800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20800 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:35.795204Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:08:36.784582Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 031d3ef1-bce8-4299-934a-9c699feb9392 · inbound

Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model cites this paper.

Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:27:36.347287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T08:27:36.242416Z digest=sha256:7c18fa052599be8ac62b05bc0b0cd40fc4d2ef845c82cca34e74491cfe954304

Observation 66558f9c-122f-4364-96d7-1705e3e425ee · inbound

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model cites this paper.

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:35.795204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:43:35.795204Z digest=sha256:2e52b7f9cec103df9be556362a1e66deac4fac8d8731c3a322c4c2fa8f6aa275

Observation aa3ec40c-2d58-4253-bce6-f92b281dfccb · inbound

USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning cites this paper.

USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T16:07:24.712500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:07:24.712500Z digest=sha256:27ffd92cb9c783394d3a292f76fa4b01c3055f774961e7de513048da80f93a00

Observation 231858dd-4c04-402d-9551-0a4a7fdc36d3 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.898019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:52ae4afbf34673dca5b8acbbb97943518d77f8e5f6a8b4f637bfc595b4bdecb5

Observation 7f1577d7-0629-488d-bc36-be05546242f9 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.615292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:11bd5f4b5923880fcee991452be2442b146eb0f9b9d8a37c85479b4611ad4cde

Observation c918efdd-0c08-4a17-879a-48199593ed99 · inbound

Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models cites this paper.

Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:08:36.786079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T16:03:53.914362Z digest=sha256:56d02dae49d63d80f777f136c05e141c68274b19b71eeb83164f61a317bf226e

Observation 77166575-9b0b-43f8-89ff-3aad018eef03 · inbound

Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models cites this paper.

Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T08:29:25.498449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:29:25.498449Z digest=sha256:6af5a71060e9f407d278ee3b547f14ea6da16f4626095242450d43dcad1a3fe4