Pith. sign in

Paper Citation Record · LEDGER

Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2107.09293.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2107.09293 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T21:11:25.336787Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T21:46:15.581615Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4990ae01-00a6-440d-94d2-4aee3b75ffcb · inbound

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model cites this paper.

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T21:11:25.336787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:11:25.336787Z digest=sha256:5ea73585c0c1dad0d8b7945223269a2231d65458eddafa34e76ffe0199b8acd5

Observation 49565245-a006-4168-b27b-68b18c85b1c1 · inbound

NTIRE 2025 XGC Quality Assessment Challenge: Methods and Results cites this paper.

NTIRE 2025 XGC Quality Assessment Challenge: Methods and Results Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T11:20:51.176183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:20:51.176183Z digest=sha256:15ffe309a826e438a5cb776f7cea75753500ce7da2b6e826b306cd14d1ba7c45

Observation 5f3739da-5e95-40cc-ad27-7280efcb2de9 · inbound

SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting cites this paper.

SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:44.518622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:44.518622Z digest=sha256:3d54240d3d63f0fd30f3ecd349ee988a90b5a1807eb52e3d494d13d71ebd327d

Observation b549b02b-96c3-43dc-8314-61caf0adf92d · inbound

Who is a Better Talker: Subjective and Objective Quality Assessment for AI-Generated Talking Heads cites this paper.

Who is a Better Talker: Subjective and Objective Quality Assessment for AI-Generated Talking Heads Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:19.120381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:19.120381Z digest=sha256:d5e7bba3c066f11940ae252f05e981a1e9a66e26ccd62736c0e3e6ed19cacb7e

Observation 1b9799a9-e6e2-4e32-af47-421865ccbb43 · inbound

AUHead: Realistic Emotional Talking Head Generation via Action Units Control cites this paper.

AUHead: Realistic Emotional Talking Head Generation via Action Units Control Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:50:40.246667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T05:49:15.734418Z digest=sha256:5267a722f08dd3607a98a4d60abd950e65ab1c40531bb74bd0efafe7a3764b57

Observation 707c910a-ce2d-4375-ba78-2e5ec50a3b58 · inbound

CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation cites this paper.

CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:41:04.425900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T03:14:45.834520Z digest=sha256:2a930cf410dd6067ba9c0a8b84fd59a14168ac9bde01b9f100235627c5e3f99c

Observation 98ed7b23-2b3a-4cd5-9d47-00028f0b7d70 · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:46:15.583006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:b048c4b57a089e068cb382b1574b88c8e3e6178db17e9163545a2811e93237b0

Observation dbf346d2-dc65-4f91-9949-97b7be03f722 · inbound

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation cites this paper.

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-04T17:26:50.150881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:26:50.150881Z digest=sha256:3546c1991648a6ed4e1f6a10b1ab9be2c4bd4b9a0f1c224565eb6d95d2d64420