Pith. sign in

Paper Citation Record · LEDGER

Identity-Preserving Video Dubbing Using Motion Warping

As of 21 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 1 inbound Pith citation observation for arXiv:2501.04586.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04586 v2

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:33:20.047608Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T01:07:49.064263Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T15:45:48.564618Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f6b97bc2-2d69-4782-9e45-790c8109709b · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild,.

Identity-Preserving Video Dubbing Using Motion Warping A lip sync expert is all you need for speech to lip generation in the wild,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.442649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.936940Z digest=sha256:e2c3a1cb84beed284ea1617e57f52cd37defb4a4934a88248de984ec4c01c9b1

Observation 98ed4075-273b-43e4-8a9c-5a2054b8a440 · outbound

This paper cites Identity- preserving talking face generation with landmark and appearance priors,.

Identity-Preserving Video Dubbing Using Motion Warping Identity- preserving talking face generation with landmark and appearance priors,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.430322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.941263Z digest=sha256:beee16a3f9e54d3578840924b1ed664f8a20c7fe0f29ef4abef4944286dc2fb6

Observation a0a86678-78d2-46be-a0b9-b0c66d232278 · outbound

This paper cites Makelttalk: speaker-aware talking-head animation,.

Identity-Preserving Video Dubbing Using Motion Warping Makelttalk: speaker-aware talking-head animation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.417064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.945130Z digest=sha256:cae54c2708c3ab2b3d68551c3279e2ae231b10b7702a64b545f8d4b2301be8e7

Observation 321b9ccf-fe5a-44e9-b522-4b6835cea730 · outbound

This paper cites Towards realistic visual dubbing with heteroge- neous sources,.

Identity-Preserving Video Dubbing Using Motion Warping Towards realistic visual dubbing with heteroge- neous sources,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.405427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.948962Z digest=sha256:5b67c474bfdf3a1b51fb2d778ddf1ae7b5fe148dc446657cb63cd7bc71a22512

Observation e0546987-95f9-4999-beb6-1d608424ac57 · outbound

This paper cites Towards automatic face-to-face translation,.

Identity-Preserving Video Dubbing Using Motion Warping Towards automatic face-to-face translation,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.391706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.952855Z digest=sha256:a47a37dc54eb5a8204c9236d389069d8fa7f256d187fb8b5a53407110f9635be

Observation f039edde-b4ea-4ec1-8d39-268d5ffa2f38 · outbound

This paper cites Semantic-aware implicit neural audio-driven video portrait generation,.

Identity-Preserving Video Dubbing Using Motion Warping Semantic-aware implicit neural audio-driven video portrait generation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.378501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.956621Z digest=sha256:df25009e20657b6a149fa4169acaabac35e188605d2c1c4e106adb493dd52b42

Observation 9a086f91-0dcc-4cc3-9e75-38e27c1044db · outbound

This paper cites Learning dynamic facial radiance fields for few-shot talking head synthesis,.

Identity-Preserving Video Dubbing Using Motion Warping Learning dynamic facial radiance fields for few-shot talking head synthesis,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.367292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.960811Z digest=sha256:6124c097f8e2871e4bb68e5add747d977de1058a3d13895549a765ad0d0c560d

Observation 2fa95ff0-e4f6-4bbd-a099-81151c5b5c67 · outbound

This paper cites Ad-nerf: Audio driven neural radiance fields for talking head synthesis,.

Identity-Preserving Video Dubbing Using Motion Warping Ad-nerf: Audio driven neural radiance fields for talking head synthesis,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.355607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.964405Z digest=sha256:a85e952ac84f3d6277a3931653068da8e1348a8227fb8ee5942c4c709546cc50

Observation 6de5c8ec-6312-4955-a102-3d588bba4056 · outbound

This paper cites Expressive talking head generation with granular audio-visual control,.

Identity-Preserving Video Dubbing Using Motion Warping Expressive talking head generation with granular audio-visual control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.342269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.968023Z digest=sha256:43cdca55b4a2eee4616be0ab08beb71897908fffa6b20a5ee2d617f333aa0ec2

Observation 2af13967-d630-4524-9de8-8a88f3b51861 · outbound

This paper cites Vfhq: A high-quality dataset and benchmark for video face super-resolution,.

Identity-Preserving Video Dubbing Using Motion Warping Vfhq: A high-quality dataset and benchmark for video face super-resolution,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.327670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.971610Z digest=sha256:197a51a93cbd78d824396638e3aae363b58cb6c08ddb9a870f798946fadedfe5

Observation b9148167-636c-4398-9b32-84782ea83571 · outbound

This paper cites Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset,.

Identity-Preserving Video Dubbing Using Motion Warping Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.313593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.975293Z digest=sha256:03c4e13b8b9a965eab25f75a9a672858117b149b16574b689d9b20215f72497e

Observation 6d78fc0b-0f1a-4a77-90bf-0803da937fd8 · outbound

This paper cites A morphable model for the synthesis of 3d faces,.

Identity-Preserving Video Dubbing Using Motion Warping A morphable model for the synthesis of 3d faces,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.299923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.978797Z digest=sha256:05fd2862b26af5fa322b4cebfb7fc336b441564f0075b27088dcddaa103e0dce

Observation 1f020e12-d836-4225-9492-3e1509f42fa9 · outbound

This paper cites Styletalk: One-shot talking head generation with controllable speaking styles,.

Identity-Preserving Video Dubbing Using Motion Warping Styletalk: One-shot talking head generation with controllable speaking styles,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.286215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.982087Z digest=sha256:a19ca3aad6607dc8c525b28e4efb49234de61a2b4495b89ddec379f7507276a5

Observation e50816ea-6815-4186-8edd-469687d262ce · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis,.

Identity-Preserving Video Dubbing Using Motion Warping Nerf: Representing scenes as neural radiance fields for view synthesis,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:19.985252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:19.985252Z digest=sha256:8ccc87d072de0b9b9052811839860616f5b0b60a3d027902312993bac1bc18d1

Observation 43a5bcc9-3997-461f-9c45-a4e422a1664a · outbound

This paper cites Attention is all you need,.

Identity-Preserving Video Dubbing Using Motion Warping Attention is all you need,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:19.988266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:19.988266Z digest=sha256:80e2220900fbf2340d95ff8b2ecb9a4f47f32f88fe44cdad50c86683e94f3499

Observation 2528ce05-625b-41e9-ab33-92ba21ac7815 · outbound

This paper cites Dinet: Deformation inpainting network for realistic face visually dubbing on high resolution video,.

Identity-Preserving Video Dubbing Using Motion Warping Dinet: Deformation inpainting network for realistic face visually dubbing on high resolution video,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.259490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.991250Z digest=sha256:ecaa109a611c07e04803ef50647e25d52be49f87a5f8398ad3e039e76213ae14

Observation a6288e1a-b07a-4d0f-90ab-ce19f5a93ad7 · outbound

This paper cites Adaptive affine transformation: A simple and effective operation for spatial misaligned image generation,.

Identity-Preserving Video Dubbing Using Motion Warping Adaptive affine transformation: A simple and effective operation for spatial misaligned image generation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.248355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.994059Z digest=sha256:7f86241276c5e00e6bc2da50c1df68824533d2c1b997fa4dc950c15facf03f77

Observation 68a38553-035e-49de-b206-eb9c5de09a40 · outbound

This paper cites Pirenderer: Controllable portrait image generation via semantic neural rendering,.

Identity-Preserving Video Dubbing Using Motion Warping Pirenderer: Controllable portrait image generation via semantic neural rendering,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.236230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:19.996963Z digest=sha256:0996a38f7c6cb28aa7948bb90b3d183be492fa56d4e67159d9974d48f68a6c27

Observation e034d0e2-be52-495c-a857-98a479d1898c · outbound

This paper cites Arbitrary style transfer in real-time with adaptive instance normalization,.

Identity-Preserving Video Dubbing Using Motion Warping Arbitrary style transfer in real-time with adaptive instance normalization,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:19.999961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:19.999961Z digest=sha256:25625d22a377d863c8d160df98b32d0cf6cf2a29880304efe211e52a3a827b91

Observation 3b6fc4d0-de6b-419f-921c-cdf418a739e4 · outbound

This paper cites First order motion model for image animation,.

Identity-Preserving Video Dubbing Using Motion Warping First order motion model for image animation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.216187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:20.002853Z digest=sha256:0295283b9f37268e21a9d353137ddbb46aabbb3ec5442c8696546666ac6d8b84

Observation 916c99e5-c375-4aeb-aee5-a96ca41b4956 · outbound

This paper cites MediaPipe: A Framework for Building Perception Pipelines.

Identity-Preserving Video Dubbing Using Motion Warping MediaPipe: A Framework for Building Perception Pipelines

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.006197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.006197Z digest=sha256:3ccfb8cc2af12a80191ebf2685e3ffe6e72878729ec4f25b7ea024126ead151a

Observation 2da28093-0eda-4ad2-9ab5-c8d50dc12217 · outbound

This paper cites Semantic image synthesis with spatially-adaptive normalization,.

Identity-Preserving Video Dubbing Using Motion Warping Semantic image synthesis with spatially-adaptive normalization,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.202034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:20.009779Z digest=sha256:118547f7140ad462c9a4a95c1fe865e20389a3cce140d0070ec38a05558c92c0

Observation b0f986ce-dcd5-4269-81e7-5b8195dd8939 · outbound

This paper cites Perceptual losses for real-time style transfer and super-resolution,.

Identity-Preserving Video Dubbing Using Motion Warping Perceptual losses for real-time style transfer and super-resolution,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.013297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.013297Z digest=sha256:05c16926472f03ac25d9b4c00ed17352a95dbd09f93cedf83821b983f990ebc6

Observation 6d9a882d-f5e5-4c0d-a5f7-66ecc26d6638 · outbound

This paper cites Least squares generative adversarial networks,.

Identity-Preserving Video Dubbing Using Motion Warping Least squares generative adversarial networks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.181914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:20.016815Z digest=sha256:1fd333cd08a76c837f169b2db9e40adcb1bebc1fb7966cea63edabce3f399381

Observation 7da0b854-af17-4118-a6ca-85ca81bd8044 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

Identity-Preserving Video Dubbing Using Motion Warping Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.020274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.020274Z digest=sha256:a67d4ecbc8e3d463855d4bad3bf6e34a66870ce2d9883bd2f18901abe2a2caa8

Observation 63eaae43-c52f-475f-ab8f-29971d333160 · outbound

This paper cites DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models.

Identity-Preserving Video Dubbing Using Motion Warping DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.024188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.024188Z digest=sha256:dc9e670f808c855fd94967f6cf13a79c97ea873affada3bbf2dfafed14b3085d

Observation 4fafed32-fa17-4481-810f-3c079d939cc3 · outbound

This paper cites Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion.

Identity-Preserving Video Dubbing Using Motion Warping Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.028036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.028036Z digest=sha256:5239349d560223a0640a610e56a31eae99524b0d70bf67116d50ecaca9462248

Observation c7467d18-e775-41e5-b2e1-4cb2b6c62ecc · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation,.

Identity-Preserving Video Dubbing Using Motion Warping Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.169093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:20.031977Z digest=sha256:b3f354251115a9b434f733cb0e346cdfa5ecd87dce4bb8a5a98eaf672f2424c2

Observation ba43f0de-9237-4519-b331-337ed749f485 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Identity-Preserving Video Dubbing Using Motion Warping Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.035895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.035895Z digest=sha256:2192fba864ea43153b94298d3f5d1d6d219b1acd72799a7900de75c3f01adafd

Observation e7c0954f-f1e7-4092-9389-10f6f9747d3e · outbound

This paper cites AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation.

Identity-Preserving Video Dubbing Using Motion Warping AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.039736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.039736Z digest=sha256:4f7f1cf97c0b9ee627c512576ef52ca87c35ee8ce3d29eab93bbce4f9d7493b2

Observation 026c6ca1-1821-4c9e-9810-af851643ce9f · outbound

This paper cites Image quality assessment: Form error visibility to structural similarity,.

Identity-Preserving Video Dubbing Using Motion Warping Image quality assessment: Form error visibility to structural similarity,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:33:20.147658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T21:33:20.043859Z digest=sha256:13bcb27d1a65c432d36688512a48833fdf2cf6602e12a18cb63d927dd26ecad3

Observation 00a3884e-3a5c-41cc-ab7d-e0e59a5a92c8 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric,.

Identity-Preserving Video Dubbing Using Motion Warping The unreasonable effectiveness of deep features as a perceptual metric,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.047608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.047608Z digest=sha256:64f50396d6049f744560a0fb44b4d4e24053813be0e8ca355675c1e8e7fdf656

Pith citing papers

Observation 9837e775-7815-4679-b8e8-0d6b09b77fa7 · inbound

KM-Speaker: Keypoint-Based Style Control for High-Quality Speech-Driven 3D Facial Animation and Dialogue Localization cites this paper.

KM-Speaker: Keypoint-Based Style Control for High-Quality Speech-Driven 3D Facial Animation and Dialogue Localization Identity-Preserving Video Dubbing Using Motion Warping

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:48.566525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T01:07:49.064263Z digest=sha256:62a98e1d917aea2672026d1999b5119c5635af84a9b74ea9f6a19559ab461229