Pith. sign in

Paper Citation Record · LEDGER

DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.05712.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.05712 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:09.795467Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T14:21:28.412464Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 474495de-3784-412f-b3f0-820278826618 · inbound

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations cites this paper.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:09.795467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:09.795467Z digest=sha256:4ebc9bef2baf681018d613ea64003bec7f9eeb9294e12609c82a51d06bf7dd71

Observation 9c2e43db-8b8a-431b-b31b-810aef76b4fc · inbound

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation cites this paper.

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:38:20.507706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:38:20.507706Z digest=sha256:d8606735a177fa47fe0fabeedfcf91ca139ba909a16327b08f9e1e097ce1b7b1

Observation 56562c9d-a4e3-488b-9d83-2ceeb5410941 · inbound

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation cites this paper.

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:21:28.415049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T14:18:02.076576Z digest=sha256:a267b2db3e3d06f86f42ff79ff6dd58c1b7a6d30da1d7ea688df475680f8a4fc

Observation 876d8baf-384e-4ae5-9810-27f4eadaca8f · inbound

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes cites this paper.

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:48:38.179260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T22:46:51.954117Z digest=sha256:69aff74f3dc099ec14e853bbab8ad3b6d4a8719f8b777e8abe57a04a8ac876fb

Observation c6d9f09c-45dc-49af-8b6b-1fcd988ad88b · inbound

ETHead: Generating Expressive 3D Facial Animation and Head Movement from Speech cites this paper.

ETHead: Generating Expressive 3D Facial Animation and Head Movement from Speech DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T00:18:38.203449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:18:38.203449Z digest=sha256:ed5bee6962426fdda26411bd485e30a1388489e31cacf85d5d3569e27f82dcdc