Pith. sign in

Paper Citation Record · LEDGER

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

As of 21 July 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2505.22647.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22647 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-21T06:31:05.380196+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T07:00:53.496569Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.572116Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e141d963-46eb-4e0b-9dd4-a4da10d6e3bc · inbound

AUHead: Realistic Emotional Talking Head Generation via Action Units Control cites this paper.

AUHead: Realistic Emotional Talking Head Generation via Action Units Control Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:50:40.255880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-16T05:49:15.734418Z digest=sha256:38a4c8169fffa9573b55b2943388f095bc85fa463e37663fc7fc555f28bfec03

Observation 8a522e79-6f7f-40c3-9027-d8ac000648e5 · inbound

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation cites this paper.

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:06:05.818834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-10T15:09:02.727887Z digest=sha256:7d7253f6cd320eda05c72b8866d7b27b63a34267385978dc48c32a7e550b404b

Observation f203a03a-a1d4-41f3-8e1a-4be122b0e434 · inbound

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation cites this paper.

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:15:10.341482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-10T11:13:27.689539Z digest=sha256:15af0bcc981ba33465c03734025838d7fa064ce6b17c5a4948831e68baa0655c

Observation e91fb8d3-1b09-4292-9d6b-8aa21c96918c · inbound

PresentAgent-2: Towards Generalist Multimodal Presentation Agents cites this paper.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.441607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:c5a36b9b1aad9b80ee8d5034856145065c23d91eba6688c3aff10ef08418e2f5

Observation 4465b051-f896-42db-a793-18df0d46d9ca · inbound

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation cites this paper.

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.724087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-06-29T22:36:53.138354Z digest=sha256:f81c40006d399c197fa16b2ec99a432e139f8576b947a719d5d32591e13842ec

Observation 4373d6ea-fa3d-45f8-8968-bf36cdb2981b · inbound

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars cites this paper.

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:44.573639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-06-26T09:09:06.925645Z digest=sha256:4e9b8e72d771317db5580a504533fe343412b8dffd9b13ece837bfc29d3ec36a

Observation 5b418ee3-d410-4682-877f-e5aac23dcbe5 · inbound

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars cites this paper.

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:05:29.129385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-07-01T07:00:53.496569Z digest=sha256:1f54d2e57c2cb03a495d50b6ded2032a535c7d93a07d5553be3534f0095e21bc

Observation 46353343-5273-4f76-8a0e-0cb7f3a567fe · inbound

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation cites this paper.

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:35:40.311712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-21T06:31:05.380196+00:00.

source=pdf_text observed=2026-07-01T06:26:20.283349Z digest=sha256:3a95cf09dc3bd88847b863739303a67e96868d6d95cc9e22b4d83bfe0ed37b60