Pith. sign in

Paper Citation Record · LEDGER

TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2304.00334.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.00334 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:50.626113Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T22:48:38.173879Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c502d7a-fa97-41c4-8b13-25c3d36acaa5 · inbound

Exploring Timeline Control for Facial Motion Generation cites this paper.

Exploring Timeline Control for Facial Motion Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.626113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.626113Z digest=sha256:f2676733c3ccf2983cda1322b28a967d8b73d1b1fb5dd3911b4519d8b070d068

Observation 64cca734-c5d4-4cfc-b158-db51f42a451a · inbound

MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding cites this paper.

MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:56.996444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:56.996444Z digest=sha256:6231f94595e6a5887a1a868bbc1ebbbb7e0652337aad19c4b55753933c38378e

Observation 05d2aa18-834e-4629-82bc-df357b077bd1 · inbound

CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation cites this paper.

CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T19:32:44.266588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T19:32:44.266588Z digest=sha256:be6e6de35fa155d33e80c0a7a6ef018bf45b72da474922742e5fe4f411c709f6

Observation d108b1ca-21c1-43c9-af87-5e2c40af0373 · inbound

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis cites this paper.

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T19:07:40.503235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:07:40.503235Z digest=sha256:3a85f10e1e5f606338423cd31fd5f1a9ca69907360455bface8721da48bc395e

Observation aafaaea3-8f20-4361-a88c-38dd72fe881e · inbound

Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation cites this paper.

Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T11:47:50.527329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:47:50.527329Z digest=sha256:743827778af1a5afeecf085b4a5a71e901733ac7bf829497a4107d434afe46c4

Observation 3a40e739-f33a-4a78-9729-aed7e8e10021 · inbound

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes cites this paper.

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:48:38.175829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T22:46:51.954117Z digest=sha256:f3b4b4032c84033358513e8730aa5925de58f54f2a555e5c2ca425c2f0029dd4

Observation 0dd9bd23-9571-421e-a853-2f245cd8f8f2 · inbound

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence cites this paper.

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:11.245024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:27:41.839123Z digest=sha256:fc1390ec007b582a3099081bd37e8510eadb1c93ad4eabff0876ea7780c8e109

Observation 73213a83-b957-4825-bc89-a3b3c87e2f69 · inbound

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions cites this paper.

EmoteGPT: 3D Human Facial Expressions from Natural Language Descriptions TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T07:49:23.168452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T07:49:23.168452Z digest=sha256:c0c4c5521fe3e6785243757b84da40411adf6dcbb83e6328c350602924fff19f