Pith. sign in

Paper Citation Record · LEDGER

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation

As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2608.05648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05648 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:54:58.686040Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14876000-a03e-4650-b884-99611cdddd6e · outbound

This paper cites Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.639683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.639683Z digest=sha256:711ebcd0bf8e300d2b084880fba3d3a7a2a1bdb1c3dfbb880bda7107a3a7a0bb

Observation 694d1c61-ed22-4361-b70e-37d4d2dc49a7 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.647292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.647292Z digest=sha256:1108f1fc56500fb08efc64053e9c9cb6b9045e9761f4c6930dd5f30706874887

Observation d25c41a4-1df8-4587-af7b-1bf1dcea5886 · outbound

This paper cites LTX-2: Efficient Joint Audio-Visual Foundation Model.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation LTX-2: Efficient Joint Audio-Visual Foundation Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.650815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.650815Z digest=sha256:d17cafb51a0d03cd1a9e5564c88b5636f1c84190f3950726094524748940bc2d

Observation f6fa9955-a18d-487f-8b09-5d0b3c84cff8 · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.658009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.658009Z digest=sha256:d2b0a9ad7a25c20634a0e3f746e89d8568859223c83324de51c3a0757c76892c

Observation 91ea9fcc-b054-465a-9e74-1cafd7393850 · outbound

This paper cites AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.661523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.661523Z digest=sha256:76e578d75725b9df95afd08e4f2890e4e742d4629c21beac6412a0c424b5bf19

Observation 9438765d-d80d-42f6-bb78-657360bcfe7b · outbound

This paper cites One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.667551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.667551Z digest=sha256:1a6a8e49c1ee14b6c7deb9d17fddc93742442b66fd1c884f5558818ad76b2c70

Observation 99330b26-b962-428f-9d6e-ed4b9a17c12a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.670639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.670639Z digest=sha256:90760e891e79092ed08ffc7c0a97f02e7db6c5741b52f6b5149cc12108685307

Observation 75a0b93e-7885-4c3b-be6b-3beb7251ee16 · outbound

This paper cites Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.677211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.677211Z digest=sha256:e61b26dd3e5b7ab629207eb42a452482e0bac1b473bb0c31a51a7f5b38730d18

Observation f01d21ae-a35f-41aa-a0cd-21ca06c4ba1f · outbound

This paper cites SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.680041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.680041Z digest=sha256:b6d81aa29c85b49a195ff023384b78c3ee29c01a3dbe5a6115ca62c33551281d

Observation 5bc91c86-05b6-4636-90b4-a3ee629a747e · outbound

This paper cites SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.683114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.683114Z digest=sha256:ef2d9f422ca32abdf5dbeda415552861000e1ff24e384dd2bdeb2bf89f7527d9

Observation ea3bcdc1-ed91-4ee5-aa03-278c26125255 · outbound

This paper cites Taming Teacher Forcing for Masked Autoregressive Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Taming Teacher Forcing for Masked Autoregressive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.686040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.686040Z digest=sha256:f455a6091de4c715feec9606d02e2f86c87ec23aa315749407ab3593b3e3e2d5

Observation 9eb543e8-d88d-4823-bb44-91125cb2dd70 · outbound

This paper cites Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.664692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.664692Z digest=sha256:c34339d17c9c18642810084bb9cba0e30acde8ed97798aebcda9b781e213fb92

Observation e7b16efa-ece7-4586-a970-745ac168715d · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.673754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.673754Z digest=sha256:b2fb723035bd9449bb11d19c1056a82520a01d125647cdd450531bc30d55dfcc

Observation 74bb284c-54b8-47aa-99d2-c80300e66aec · outbound

This paper cites SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.643717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.643717Z digest=sha256:cc4518a5d238aa7e019e3c0a2c0d1fcdc8865fdc1bc2af155ec573d4637fd476

Observation 7b68b13c-0de6-45ee-807e-dc1472fd2363 · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.654243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.654243Z digest=sha256:c144254ad539eea9462473898d982464ecf3ad24c62140d0b7e14e12cea740d5

Pith citing papers

No inbound Pith citation observations are available.