Pith. sign in

Paper Citation Record · LEDGER

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation

As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2608.05648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05648 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:54:58.686040Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14876000-a03e-4650-b884-99611cdddd6e · outbound

This paper cites Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.639683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.639683Z digest=sha256:b099ed66cda50e93c97d6ba30eeed41caeef56a417ca8633639268a866be8cb4

Observation 694d1c61-ed22-4361-b70e-37d4d2dc49a7 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.647292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.647292Z digest=sha256:90073e02993fbcf780117d9dc1c6f55d034f16086dd168e7fea1a192968a2998

Observation d25c41a4-1df8-4587-af7b-1bf1dcea5886 · outbound

This paper cites LTX-2: Efficient Joint Audio-Visual Foundation Model.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation LTX-2: Efficient Joint Audio-Visual Foundation Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.650815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.650815Z digest=sha256:6fe0d1d22625c5c202b411232ed928242a1f12c5757c4eafcaebf1c7a1cfe045

Observation f6fa9955-a18d-487f-8b09-5d0b3c84cff8 · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.658009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.658009Z digest=sha256:a127a1840e0d177526e60eff5aa1c76e26563b920a534b7a13b5c921347d6aa8

Observation 91ea9fcc-b054-465a-9e74-1cafd7393850 · outbound

This paper cites AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.661523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.661523Z digest=sha256:fa9ec9a0d8a69f0fb272f9b2f4bd6785f661c3bbca44547548bdabf985c44a69

Observation 9438765d-d80d-42f6-bb78-657360bcfe7b · outbound

This paper cites One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.667551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.667551Z digest=sha256:58de8fe2447b4f32e6c05740dad9eff62032df6ab330a64db1bebdba5f7f246d

Observation 99330b26-b962-428f-9d6e-ed4b9a17c12a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.670639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.670639Z digest=sha256:4416d479fef782f0f5533f5c7a106d178096e059da41b1cc44d3ddde02c22ae0

Observation 75a0b93e-7885-4c3b-be6b-3beb7251ee16 · outbound

This paper cites Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.677211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.677211Z digest=sha256:9fec1dda8b0c729d4fd452de5daaa5f578c1fcf983258f056f357e71590f95da

Observation f01d21ae-a35f-41aa-a0cd-21ca06c4ba1f · outbound

This paper cites SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.680041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.680041Z digest=sha256:a49196d82b0d23dfff8ef3fa5a102ed4d719ee880c80ebf57a7f4be7768cbc54

Observation 5bc91c86-05b6-4636-90b4-a3ee629a747e · outbound

This paper cites SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.683114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.683114Z digest=sha256:f742b6e70e39d83e8e4811beab9d22539aae93c3d510709c6992da6d7c520436

Observation ea3bcdc1-ed91-4ee5-aa03-278c26125255 · outbound

This paper cites Taming Teacher Forcing for Masked Autoregressive Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Taming Teacher Forcing for Masked Autoregressive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.686040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.686040Z digest=sha256:a30b27869c0bc7fcc83c3f65903f31459d4063df1acf07226f67778c1fe9144b

Observation 9eb543e8-d88d-4823-bb44-91125cb2dd70 · outbound

This paper cites Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.664692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.664692Z digest=sha256:7a231b575d96404959f9db061ec26974fbe8391747c6f97ef89fb616dec59f3c

Observation e7b16efa-ece7-4586-a970-745ac168715d · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.673754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.673754Z digest=sha256:a941c7be83fbc2dce223e8f04414fb5b6aeab0bb56592f570eee7e09c58f1b48

Observation 74bb284c-54b8-47aa-99d2-c80300e66aec · outbound

This paper cites SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.643717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.643717Z digest=sha256:10eb57cad364d49bcfb14ec9d677d1ab2a5776be1db9376a1ab3ab491398c101

Observation 7b68b13c-0de6-45ee-807e-dc1472fd2363 · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.654243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.654243Z digest=sha256:1bbbd989563a267be06e5666f908713bfc1ce37e053578a5e73c64aaae1fa82a

Pith citing papers

No inbound Pith citation observations are available.