Pith. sign in

Paper Citation Record · LEDGER

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

As of 9 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2509.20128.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.20128 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:18:02.076576Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:18:02.076576Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-18T14:21:28.416903Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy25
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eb3ed498-689b-41fc-b7e1-efa123530228 · outbound

This paper cites KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:21:28.419042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:aa3fc01c8fb55ace9918623ab69b0cf2145ce171ef1c92289d96f40d77e9f30b

Observation 524236a2-c7a1-4ce4-893b-6c9f7f13bd94 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.870577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:891221fc48c1b835a593065828ed919681437efc9e5030ae03c69779375f85b5

Observation 14298c15-b4c2-4b14-ae3c-182f77e4ca19 · outbound

This paper cites Dataset We train and evaluate our model on two benchmarks.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Dataset We train and evaluate our model on two benchmarks

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-05-18T14:22:40.862010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:d8e906ade7a684e323962a7a51254dac67be8c5d3af0df8924bef1006e90aa91

Observation d6425d6b-ceee-4d6f-8f3c-03fd823e1784 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.865213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:91372bd44021e02ef42b8db55b2882a694d6c4621be35712a46403eafc421c7e

Observation 1c9ed7f2-7038-438f-9ee1-122cb4a34e54 · outbound

This paper cites an unresolved cited work.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-18T14:22:40.823338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:b92a27646aec89ff23b1bfcb9dd3da347a43274260055fbb277ceeb588106cf0

Observation fa464dfa-eaa9-437d-adbc-9de5b15d5c7f · outbound

This paper cites Facediffuser: Speech- driven 3d facial animation synthesis using diffusion.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Facediffuser: Speech- driven 3d facial animation synthesis using diffusion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.837800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:a71a5b88e396d4cad671b2aef3a61cdfd013b6da00b4309948872fe4ebf51f36

Observation 1223e6a9-01d4-4471-8fd8-0e2ba71ed2f5 · outbound

This paper cites Difftalk: Crafting diffusion models for generalized audio- driven portraits animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Difftalk: Crafting diffusion models for generalized audio- driven portraits animation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.809585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:ceca3e84f255dd1d5cbbe36cacc00fede7beebbf238364c018c67008fa883f9d

Observation 56562c9d-a4e3-488b-9d83-2ceeb5410941 · outbound

This paper cites DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:21:28.415049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:3e34b16e82fa589a00ac656e83a70dc7e0c558b60cce8a153dc1b4bcd78ad03c

Observation 7d31c9a5-3cca-4099-9c66-8b439395e30b · outbound

This paper cites Emotivetalk: Ex- pressive talking head generation through audio information de- coupling and emotional video diffusion.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Emotivetalk: Ex- pressive talking head generation through audio information de- coupling and emotional video diffusion

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.805335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f0efa86f2fb9ef4b50ac83f2066fa62002449cd1fdac40a97a72b3c346112f42

Observation eefb76c7-2aeb-43b9-8a5b-05ef2b929d1f · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.799341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:d3c489c8bb29f680cd5e5712a6b1baf289b5c4572fe35c0e51561d435efcb9fe

Observation 4c23ca70-4119-46f6-b3f5-e1dbc91596bf · outbound

This paper cites Synctalk: The devil is in the synchro- nization for talking head synthesis.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Synctalk: The devil is in the synchro- nization for talking head synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.924511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:b0ca4ad387d902c47d190866d1434d13dd656ea4a844f44e2def48d96f715f60

Observation 4075752b-5b85-4d09-bc8e-aba6af630474 · outbound

This paper cites Prosodytalker: 3d visual speech animation via prosody de- composition.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Prosodytalker: 3d visual speech animation via prosody de- composition

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.920557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:feaf251599d9b56a01f0e5274f6aeb904c86a0aa54f777608aad8e1129fb9a80

Observation 41225913-770f-4596-b402-a9bcc0d97999 · outbound

This paper cites Keyface: Expressive audio-driven facial animation for long sequences via keyframe interpolation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Keyface: Expressive audio-driven facial animation for long sequences via keyframe interpolation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.916558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:ff39a2be6338e8c0fbefac1b3b672a67a389829608e7a7b3342b6890c109137e

Observation a7a3b5e2-9c9c-4751-a504-daf40d7dae21 · outbound

This paper cites Speak: Speech-driven pose and emotion- adjustable talking head generation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Speak: Speech-driven pose and emotion- adjustable talking head generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.913070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:8c3f19ce1f7da6c19f6d222e1031a6fd1b924c1c2589b396e788ac5483884f0e

Observation 29bd2907-6bf2-4142-8af7-a179eb30449d · outbound

This paper cites Fd2talk: Towards gener- alized talking head generation with facial decoupled diffusion model.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Fd2talk: Towards gener- alized talking head generation with facial decoupled diffusion model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.909664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:71b1b02825068971dbb39fcd6dea40ac3bfa3e3e6cfb0b71f2a18f1b44aed64a

Observation e66ee7f3-c55a-4a14-8053-eb40964e37bd · outbound

This paper cites Learning an animatable detailed 3d face model from in-the-wild images.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Learning an animatable detailed 3d face model from in-the-wild images

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.905080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:c5db8e750cab205a02b8373381dd34f94365c371cd9089d1d7476238118c0af0

Observation da8c9033-b841-4664-b959-90a603d48439 · outbound

This paper cites Disco- head: audio-and-video-driven talking head generation by dis- entangled control of head pose and facial expressions.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Disco- head: audio-and-video-driven talking head generation by dis- entangled control of head pose and facial expressions

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.900229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:d7d6c3de0b1a2854759170e20f53cb0ee118b54c13301bae3fda87e3e6cbe8bc

Observation 8f62fb29-327d-45b3-98ca-ec67218cc8bd · outbound

This paper cites Nerf-3dtalker: Neural radiance field with 3d prior aided audio disentanglement for talking head syn- thesis.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Nerf-3dtalker: Neural radiance field with 3d prior aided audio disentanglement for talking head syn- thesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.896595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:7722000b61c394dd4a856c6115ee3d54127d1371c2550c6349d262307f0700de

Observation 8ab9707f-82e6-449a-82c1-25603f0328b2 · outbound

This paper cites wav2vec: Unsupervised pre-training for speech recognition.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation wav2vec: Unsupervised pre-training for speech recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.892830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:daedaed5eabea741df195caa2777294bae5a6225301b2bfad05bff49ef22d941

Observation dc6a7aa7-8566-4c2b-9638-1c258428a924 · outbound

This paper cites Spsinger: Multi-singer singing voice synthesis with short reference prompt.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Spsinger: Multi-singer singing voice synthesis with short reference prompt

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.887342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:60d6c983486c9d10f8e9ee89acac595eed63755ccf4b5b7ec9d7c6080865baf1

Observation c2560861-549d-47fe-b38d-d0a98e250f28 · outbound

This paper cites Prosody-Adaptable Audio Codecs for Zero-Shot V oice Conversion via In-Context Learn- ing.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Prosody-Adaptable Audio Codecs for Zero-Shot V oice Conversion via In-Context Learn- ing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.883688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:30343ba13836c7e125d7b7d9b686798909ffd76a8be2e187398b626765b046d0

Observation 26b28b96-2fd3-445b-bf14-db9b433fe9bc · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.879932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:34f9c298f53ac8d67468f94fc3bae514b8d3f726e3f0c9824d89f1f0bb2f28c8

Observation 164b0662-aa5c-47ac-9d5a-6e6e0e64ecd4 · outbound

This paper cites Improved parallel wavegan vocoder with per- ceptually weighted spectrogram loss.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Improved parallel wavegan vocoder with per- ceptually weighted spectrogram loss

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.876106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:26d8b541566ce5188fe0576f6e1dadc260d87fe2a2b7876ae248482dcd64cdb6

Observation 6dc7dfbd-83c9-44a3-80c8-a88b3a5dfee1 · outbound

This paper cites Flow-guided one- shot talking face generation with a high-resolution audio-visual dataset.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Flow-guided one- shot talking face generation with a high-resolution audio-visual dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.871754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:7aed635b7f6241cdcd9da9c303a0379e959ac91c3d949d3a170d518f5e71b40e

Observation 19b246d7-0020-49fd-8cf5-9eb74c6546ec · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation V oxceleb: A large-scale speaker identification dataset

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.864861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:4fbbbe202dae6cd65712aa3caba66ffecaf967b682cdeb8596409410e5dfca84

Observation 4baf7e6b-bcc6-4e7e-b8d3-eabe5609d273 · outbound

This paper cites Hallo2: Long-duration and high- resolution audio-driven portrait image animation.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Hallo2: Long-duration and high- resolution audio-driven portrait image animation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.859854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:e5e3b04c89b5cb37239f075dd41805c9b1c3b251ce6623c28e9827719683d873

Observation 026f25cf-ed6c-4621-b2cc-8d4f843ab51b · outbound

This paper cites Meshtalk: 3d face animation from speech using cross- modality disentanglement.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Meshtalk: 3d face animation from speech using cross- modality disentanglement

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.852453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:506f169b99db76b09e9c521be615fc703d1ead4f24a9cb54c664db7d0a0b0ca4

Observation daca2d65-86e7-47af-9fb2-7760c3dbc0d6 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation A lip sync expert is all you need for speech to lip generation in the wild

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.848920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:b0977fb26e3aeb0dac4b397423f8921c013cd89838d5b312696d5d80534f12cd

Observation ec6bad9b-95f9-479c-afbb-ac19f0b4c697 · outbound

This paper cites Fine-grained head pose estimation without keypoints.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Fine-grained head pose estimation without keypoints

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.844160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:f3c8cdf30ac048ba3641003ee126a7b060a3d47e8aa50c9b92751cd81fe144bc

Observation 21fcfb77-80ee-4ee8-93a2-68dfb07fc2cf · outbound

This paper cites Bailando: 3d dance generation by actor- critic gpt with choreographic memory.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Bailando: 3d dance generation by actor- critic gpt with choreographic memory

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.839699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:fbd76ebe3f0e6aedb085728d46b7eafbdcb0f02fa4888089792cb377bd92f12b

Observation 0241366b-e474-4b36-be55-2c3d835f8e63 · outbound

This paper cites Robust speech recognition via large-scale weak supervision.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation Robust speech recognition via large-scale weak supervision

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T14:22:40.836070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:2b522e911d9c3abd7aead64bfae0b2c13544c824e8914541ac9f1df8f7a48859

Pith citing papers

Observation eb3ed498-689b-41fc-b7e1-efa123530228 · inbound

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation cites this paper.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:21:28.419042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:aa3fc01c8fb55ace9918623ab69b0cf2145ce171ef1c92289d96f40d77e9f30b