Pith. sign in

Paper Citation Record · LEDGER

Identity-Preserving Text-to-Video Generation by Frequency Decomposition

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2411.17440.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17440 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:59:36.479084Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.422443Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 157cdf74-9c98-48e7-9520-9f6312c91447 · inbound

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute cites this paper.

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.418911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T17:51:41.947939Z digest=sha256:1d5216c9e46bead3c0f7fe0ed3d3ddd9af0eff0ec02923acf2f6eba029067251

Observation a0f306ac-9e16-44a7-9521-3256679dafe1 · inbound

ImgEdit: A Unified Image Editing Dataset and Benchmark cites this paper.

ImgEdit: A Unified Image Editing Dataset and Benchmark Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:17:45.442169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T18:17:45.123690Z digest=sha256:1e51819bac425ad8b919e6781cc6f326fd6db37c78600e69b0ef713d65192583

Observation ab3fa46f-dd86-49fc-afe4-7e1a6d1947b2 · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:36.479084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:36.479084Z digest=sha256:708f348b8e11d5e753736e4d1aa5aab71f8d4b5f22f701596ae473ce75bcebe5

Observation 46a1ad32-721b-48e2-a3d0-d42d1fac1549 · inbound

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation cites this paper.

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:34:27.070137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T17:34:26.951644Z digest=sha256:7b2e6a6ac7ea1bed40accce8b1d6cd6033f94ce8bf51421ceb7edd5f26990bf5

Observation aaadd1b3-ad36-4399-88df-ec3aae29cbb2 · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:42.872478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:42.872478Z digest=sha256:b5ec88c268c9cb4208de3b73aba821f662634af3f2a3efbf67d1aeb9a2fd8ad2

Observation fef7acec-22d0-4247-bda7-769ee6ae4dbd · inbound

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting cites this paper.

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T10:43:26.982019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:43:26.982019Z digest=sha256:a7c956251a17663037db4136115e0b4344ad727cd43a0c4bfbc26f4247ece8b9

Observation c3761732-b039-47c0-a26f-1e59f3b149d6 · inbound

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement cites this paper.

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:33.052194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:33.052194Z digest=sha256:1888004c9bfd4549672cb8bd0426540fb027bde310b5cc869fde5791e40533e9

Observation 241d36e5-0297-4afd-97b0-cd103d00fd43 · inbound

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers cites this paper.

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:39.711058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:27:39.711058Z digest=sha256:f161992cb4219855ecd3df00c08758e4129e4f675ace69c9f02f36e89a3c4513

Observation 75be7455-bf0e-4288-ad58-2f692377874b · inbound

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation cites this paper.

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:02:10.514105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T07:59:49.398271Z digest=sha256:acfd6f5e29c89e8b0be870aede651d6372348bd7f3daf69ed5f7727c89e54312

Observation 9a45fe12-994c-4630-b9ce-94e3317f18ec · inbound

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation cites this paper.

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T19:19:32.838413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:19:32.838413Z digest=sha256:5cf5c273e88d092f8c28db13f6cdcb661cf42040618c11989db53d3f47e1f4c4

Observation 368fc191-8da6-4ce5-9ed0-67e9b7be3445 · inbound

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea cites this paper.

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T04:44:59.127465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:44:59.127465Z digest=sha256:bbeb39a444adf5ea323658a50a3bb9adb7a07d01f30708f587e908f814f6caf0

Observation 1d4434e8-b090-4277-b226-97d0e06fddcc · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 196

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.584551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:a56a37fcfcdeec9f37460a346fa797a9521ea3f2e514117a0c63c52043e43f3e

Observation 0975307d-bb4c-460d-ae3a-0e738b08c3fc · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:03.483661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:3ffc8f8df0ff89343f0ed97a55ec010e98376c895ee907f9718f7fd60e930d59

Observation 609aa3f1-5b4b-42f1-a784-2f903f394879 · inbound

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation cites this paper.

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:55:10.368181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T06:53:13.152677Z digest=sha256:9c458b5e00191782289f9fd7e24fce5f840ae357bba62f217a466fc0816509eb

Observation beaaa26e-a4a8-4af6-88e3-1f09932a983f · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.795138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:a3b6de7b62f50a850df0274c161cee098b89868a09eec2848fe7570ae63ff4c8

Observation 22a7e0d9-e635-464f-ba6d-a50f53319213 · inbound

Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy cites this paper.

Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.308647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:41:06.600948Z digest=sha256:429c2cda0669f96abac72be7b579487fc47a9997991054b4b72db7de2a828252

Observation 5a9005f6-bf64-40aa-a5a7-0e02b26ee519 · inbound

A Comprehensive Ecosystem for Open-Domain Customized Video Generation cites this paper.

A Comprehensive Ecosystem for Open-Domain Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.902889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:06:33.572182Z digest=sha256:2e9a55bea049beb536b0b65f62d726dfe36c13494e6df4b92443e2f72a78781b

Observation 2819a73a-0c3c-428b-8e6b-a5cfa6e6ec2d · inbound

Customizing Video Portraits via Identity-ActionDecoupling cites this paper.

Customizing Video Portraits via Identity-ActionDecoupling Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.423833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T10:56:11.553110Z digest=sha256:9d764c8ec6999acd7d8c77d510d4a631f661af0ab646d2e3b0a5f0f5c47ef14e