Pith. sign in

Paper Citation Record · LEDGER

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2505.18096.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18096 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:13.746751Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3b796bc-c835-4589-85c3-27c5b74dad10 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations.Advances in neural infor- mation processing systems, 33:12449–12460, 2020.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations wav2vec 2.0: A framework for self-supervised learning of speech representations.Advances in neural infor- mation processing systems, 33:12449–12460, 2020

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:24.059388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:06.664873Z digest=sha256:373824eebdfef1f35024b9797bdc576518ebf9b24de9595c2e4ea277063423f9

Observation 06a19676-da99-4faa-b7bf-e69fd9de279f · outbound

This paper cites High-fidelity fa- cial avatar reconstruction from monocular video with gen- erative priors.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations High-fidelity fa- cial avatar reconstruction from monocular video with gen- erative priors

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.759935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:06.727929Z digest=sha256:5f7d104b95d4ba2a882daba5bf5c2ec1a912af818db5d481a6781a42ef274a0e

Observation ff12b923-45aa-4306-a924-592a3dcc8f5a · outbound

This paper cites Pyannote.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Pyannote

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.359982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:06.854167Z digest=sha256:8aaed912442d46e0e5b777a051d4469287814be1bb8901806959451a4008afff

Observation 84ff27c2-43a8-4706-a46d-1e6bacefe58d · outbound

This paper cites Expressive speech-driven facial animation.ACM Transactions on Graphics (TOG), 24(4):1283–1302, 2005.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Expressive speech-driven facial animation.ACM Transactions on Graphics (TOG), 24(4):1283–1302, 2005

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:23.050517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:06.984124Z digest=sha256:5cb25a4a77ff0f1a91c823eb2628f7128c6f7944d4fef21df262c34cbea409d7

Observation 3f2b911b-466e-4456-ba0a-cab1070360c2 · outbound

This paper cites Human conversation as a system framework: Designing embodied conversational agents.Em- bodied conversational agents, pages 29–63, 2000.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Human conversation as a system framework: Designing embodied conversational agents.Em- bodied conversational agents, pages 29–63, 2000

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.655777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.100300Z digest=sha256:3cdf905abf35380caf2108f73b6779a76c5bd66211da24c2e5629b91f52bedd6

Observation c4643fe1-6037-4362-b040-2837c21949fe · outbound

This paper cites Cafe-talk: Generating 3d talking face animation with mul- timodal coarse-and fine-grained control.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Cafe-talk: Generating 3d talking face animation with mul- timodal coarse-and fine-grained control

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.372229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.224197Z digest=sha256:7ed28962d85bdaa4f2634ff56eb07f5cb28b2c05450136387e503b66c3b82356

Observation fd2b9128-deaf-4bc0-a755-0d6898416b8a · outbound

This paper cites Capture, learning, and synthe- sis of 3d speaking styles.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Capture, learning, and synthe- sis of 3d speaking styles

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:22.019144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.362385Z digest=sha256:783608a41375a5db6e38df9488077b327ae57364e90a34f717a31016c7e88306

Observation 8d7d3b56-896f-4f2f-8a5a-3977234923ad · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:21.672545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.466214Z digest=sha256:f86e19494711a657ba7cecd2e5834c2cf1d71fa53d024a60952052743032173f

Observation dfd6207e-f5b7-464a-90d4-e2c7fbae2d00 · outbound

This paper cites UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:07.573603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:07.573603Z digest=sha256:ec72e2c73381f227c55c69cb9f45c515909b79e595f4151e6e48f3bd86a7cf94

Observation 4827c36e-ac0c-4dd1-970f-f3ce33a5c38a · outbound

This paper cites Faceformer: Speech-driven 3d facial anima- tion with transformers.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Faceformer: Speech-driven 3d facial anima- tion with transformers

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.535978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.722156Z digest=sha256:f7b01ab04ee1c1e36e9fc0afb65f52364c17636469d1a8bfa9bc8e2ec8e6ebd7

Observation 7cc3a6c8-ad41-4a75-8f5a-08ec483ea04b · outbound

This paper cites A 3-d audio-visual corpus of af- fective communication.IEEE Transactions on Multimedia, 12(6):591–598, 2010.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations A 3-d audio-visual corpus of af- fective communication.IEEE Transactions on Multimedia, 12(6):591–598, 2010

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.366870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:07.853427Z digest=sha256:5c1b6cc293295117a874a3e59a746526cf2c4f8a542a939ff45869a56c7ddcd3

Observation 55544b14-4360-47ec-b385-6c9b16d8e246 · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Affective Faces for Goal-Driven Dyadic Communication

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:07.961032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:07.961032Z digest=sha256:5f06b629d222f7e52fc311809ce6ceff6258a1a38a4d57bb88e827185d65a9c1

Observation eddc3004-6431-47f9-8f8d-d31ca1d0e223 · outbound

This paper cites From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.118474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.118474Z digest=sha256:dcb88dda44dadef90974361c71fe0fab8fc76073124b242aeea4ef82b2ed0287

Observation ca303e15-0c16-4bb4-be8e-e2eac657cb9a · outbound

This paper cites Long short-term memory.Neural Computation MIT-Press, 1997.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Long short-term memory.Neural Computation MIT-Press, 1997

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:21.231329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:08.246824Z digest=sha256:113493b577498e94e31a0d06b28e9398ee6be0cf5e1d3d1585e5bd4bacf58496

Observation 3f0ddd31-29ba-466f-b13e-f4d3ebfcf144 · outbound

This paper cites Toward rnn based micro non-verbal behavior generation for virtual listener agents.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Toward rnn based micro non-verbal behavior generation for virtual listener agents

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.979997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:08.357873Z digest=sha256:794c2548b0d193e4b811955567af2ffa0cfea1290a8fb4b61fc2ea99c5f68a9c

Observation 6309504b-f066-41d6-953b-f129ecc8e25b · outbound

This paper cites Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.474671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.474671Z digest=sha256:9003027403edd75fa744a57c1199e2b395b634441f2c4a6468a7f51d5c359faf

Observation 7ef5a1c9-c907-4aac-9f80-de3e4a631326 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Adam: A Method for Stochastic Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:08.615584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:08.615584Z digest=sha256:3c88185aa0b89756bad212d2d7650edc4aed1622a1ed997d161251c9250f31aa

Observation a180a451-e273-4a16-9e5e-ae34426f373f · outbound

This paper cites Iianet: An intra-and inter-modality attention network for audio- visual speech separation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Iianet: An intra-and inter-modality attention network for audio- visual speech separation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.711792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:08.753831Z digest=sha256:a0e62a5a4b5c512e5163b22c67ced7d2565ac3e849f32908ab680b5b179c1dfe

Observation 29ff24c6-5222-46b6-941d-14e1c6653dd4 · outbound

This paper cites Learning a model of facial shape and expression from 4d scans.ACM Trans.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Learning a model of facial shape and expression from 4d scans.ACM Trans

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.498754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:08.849891Z digest=sha256:2b82bb79c449129a8a50186f52a8f97fb92637bb903d1d8d375e7f122296fc7f

Observation 4d1a5a81-b6b6-4ed5-894c-c1e9831dc938 · outbound

This paper cites One-shot high-fidelity talking- head synthesis with deformable neural radiance field.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations One-shot high-fidelity talking- head synthesis with deformable neural radiance field

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:20.205414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:08.936420Z digest=sha256:eb2d47f0d576399c81b6ef0db13af8aa993265ffd9e29c55c75ae1dcaf8ba431

Observation 7b4f4828-477d-4500-9c81-0c5ef5543cc2 · outbound

This paper cites Proactive con- versational agents in the post-chatgpt world.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Proactive con- versational agents in the post-chatgpt world

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.990434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:09.113832Z digest=sha256:b2151f50ecfdf5e99495c537174936a99b82b54b0beb6ce7da8573c38a1c0727

Observation d54a3abc-70b4-4dbc-b34d-9f4fd95a92ca · outbound

This paper cites Mfr-net: Multi-faceted responsive listening head generation via denoising diffusion model.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Mfr-net: Multi-faceted responsive listening head generation via denoising diffusion model

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.774491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:09.273424Z digest=sha256:650fe6c00a3b46bf7aa349c0841a924a7d94b1f279b46aad30cbd285b56fbe27

Observation 7eaf8bea-ed03-437c-af30-c8df55ba20ab · outbound

This paper cites Customlistener: Text-guided responsive inter- action for user-friendly listening head generation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Customlistener: Text-guided responsive inter- action for user-friendly listening head generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.530943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:09.391277Z digest=sha256:5c3dabfbe6a49f95e4638048ba8226489443e9615661b006532b83ff6f9caff8

Observation 5644a605-6430-4064-a6ec-56219208bce4 · outbound

This paper cites MediaPipe: A Framework for Building Perception Pipelines.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations MediaPipe: A Framework for Building Perception Pipelines

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:09.516177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:09.516177Z digest=sha256:eb62f407fb49528517efefbf5ff49101c4bb048b78a1f94373f1cdd69ced60cb

Observation 4abc0771-f79b-42ec-9f2a-145e8594ce56 · outbound

This paper cites ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:14.078463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:09.681648Z digest=sha256:929e3602eb9c08592fff1131fcc72c4add20fe5bd7eb6ff4f3df675ff2ae7590

Observation 474495de-3784-412f-b3f0-820278826618 · outbound

This paper cites DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:09.795467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:09.795467Z digest=sha256:4ebc9bef2baf681018d613ea64003bec7f9eeb9294e12609c82a51d06bf7dd71

Observation 323b7c63-c06b-4cf7-af42-8cf5fb3c9973 · outbound

This paper cites Learning to listen: Modeling non-deterministic dyadic facial motion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Learning to listen: Modeling non-deterministic dyadic facial motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.277225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:09.900396Z digest=sha256:891c44e8ae14497f1a6e704d13eaa194db1632f5b9ad42f4572ac7a0a4e41d66

Observation 3edcf984-b1f2-4670-b137-3871b492d646 · outbound

This paper cites Can language models learn to listen? InProceedings of the IEEE/CVF In- ternational Conference on Computer Vision, pages 10083– 10093, 2023.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Can language models learn to listen? InProceedings of the IEEE/CVF In- ternational Conference on Computer Vision, pages 10083– 10093, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:19.076045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.012426Z digest=sha256:a922434c853274198b544dc606a2c904ab51df995a1df16250ebe14ee6a0cbb9

Observation 6149e777-ba56-4bfb-ab1c-234408708823 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.821438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.117617Z digest=sha256:a243b5d995f0176be87393170451ffd93a55192ae9b08bdafd21ab90f2f9e433

Observation 2d384c6f-2910-4821-bb5d-474f55be4d64 · outbound

This paper cites Real-time 3d talking head from a synthetic viseme dataset.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Real-time 3d talking head from a synthetic viseme dataset

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.577419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.259878Z digest=sha256:1cb2de5e81cf240431f23bb42f780456eeed6ac7574b5f29f2abeebbb3416459

Observation 298d4bb5-b1bd-42a8-bfd0-31f5d2d268f1 · outbound

This paper cites ScanTalk: 3D Talking Heads from Unregistered Scans.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations ScanTalk: 3D Talking Heads from Unregistered Scans

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:10.356811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:10.356811Z digest=sha256:a7f0d426c9addf408005a51655c4d18d1ae4298376567b109d5b19c03cbec8da

Observation 22857179-683d-4e48-a559-6bea681d0ce2 · outbound

This paper cites Dpe: Dis- entanglement of pose and expression for general video por- trait editing.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Dpe: Dis- entanglement of pose and expression for general video por- trait editing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:10.470371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:10.470371Z digest=sha256:b26a73f75b56ad2ade25125b44d998c6f66f6fc1abbb559ff66ba4756bb530e4

Observation b777e672-7a0f-4494-b473-a645593ed547 · outbound

This paper cites Selftalk: A self- supervised commutative training diagram to comprehend 3d talking faces.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Selftalk: A self- supervised commutative training diagram to comprehend 3d talking faces

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.342972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.596544Z digest=sha256:b1e5ba49cc0a9486c6b01b161e1943c2ca73052a7bef715c03134bad791f787f

Observation cf1bd416-72da-4c94-bef3-b8f3d1f3b6c9 · outbound

This paper cites Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:18.136046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.717041Z digest=sha256:97f2f2bb5ced6556f6707faa53f2f9633ed2878d663abc9d1c5356ea92293220

Observation 3b82443b-c7d1-49ac-a94f-fbc492e180dd · outbound

This paper cites Synctalk: The devil is in the synchronization for talking head synthesis.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Synctalk: The devil is in the synchronization for talking head synthesis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.923327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:10.847276Z digest=sha256:b52917c93708d58e6499a7e4213db213721b144e9ac801305c1713697c4b2efc

Observation ca293a36-ff56-41e0-bb22-1ebf757d3c72 · outbound

This paper cites Meshtalk: 3d face an- imation from speech using cross-modality disentanglement.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Meshtalk: 3d face an- imation from speech using cross-modality disentanglement

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.685594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:11.024346Z digest=sha256:f91a210a62fed6208703aafa77ab67e9804104a379685b7f7a4368a23c4c6f3b

Observation 744d021a-c70d-45b0-8101-8011d94f188f · outbound

This paper cites Emotional listener portrait: Neural lis- tener head generation with emotion.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Emotional listener portrait: Neural lis- tener head generation with emotion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.459200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:11.164248Z digest=sha256:a59365a5bd578f2df9af7157e219d4ece1c7321d81045bb6bc49123de6ff58d4

Observation 422e804e-b1db-41f4-84f9-84e6df6d5335 · outbound

This paper cites React2023: The first multiple appropriate facial reaction generation chal- lenge.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations React2023: The first multiple appropriate facial reaction generation chal- lenge

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.268369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:11.336635Z digest=sha256:04981ce03f61ca71bc5501d28fc9fc01690bd2bb77f7fefe61d534a1b143e83f

Observation 6cd5d85d-b83a-4509-8cbc-15c47c3a9567 · outbound

This paper cites TransNet V2: An effective deep network architecture for fast shot transition detection.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations TransNet V2: An effective deep network architecture for fast shot transition detection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:11.450440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:11.450440Z digest=sha256:4be61840b72214bd6193d0e761c170401e45798eac9dd4c29c2c2c09d2afc9cb

Observation 8c9f52ef-28d9-45f5-9f83-a0ceb0336a92 · outbound

This paper cites Laughtalk: Expressive 3d talking head generation with laughter.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Laughtalk: Expressive 3d talking head generation with laughter

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:17.031219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:11.606460Z digest=sha256:4e86a68e208ac3e2a38286edfb4499c06cdac2608c3e37492a146e5876f5eaea

Observation 03f9e27e-2e52-44d7-b817-9560636c1dfe · outbound

This paper cites Imitator: Personalized speech-driven 3d facial animation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Imitator: Personalized speech-driven 3d facial animation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.845207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:11.704385Z digest=sha256:4f59e9a2363c35f41c2300ba67ad17a4449cd288d84329e4f0edf549213b5bb5

Observation c45ad741-e225-4f69-90c9-ea95eaa12a2b · outbound

This paper cites Dyadic Interaction Modeling for Social Behavior Generation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Dyadic Interaction Modeling for Social Behavior Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:11.891223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:11.891223Z digest=sha256:c90025a551c10f2dc7ee6d62e8e9bf34918203fba855f68bdee5f38ffc9deb4d

Observation 5cc2cdf7-7a8a-4c89-a0db-6bd81d8d7aae · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems, 2017.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Attention is all you need.Advances in Neural Information Processing Systems, 2017

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.620401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.029845Z digest=sha256:2dca13d665bd72184c029e988f031e673fcb1e7c036aa2c7a43a031e10b6f167

Observation 91285c40-8cba-4b34-8140-50857e423498 · outbound

This paper cites VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.156492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.156492Z digest=sha256:ea6bce174fa55dcba67206a427b485cabe46dc0234482e3d6f656a0a3597946c

Observation f0df8f66-6cb5-4eb3-8895-e9c2949a2219 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.290540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.290540Z digest=sha256:1f06dcdc124b8e8d2e4c9ccb8b0e998ebe6cb0385f0fef8c1a5f13d324490b8f

Observation 563c1a0b-a1bb-428e-946d-541f3395724e · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.436078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.406648Z digest=sha256:86356f6a3555709c0d88210229b47cdee6c910b4a4f42ded1d71df03adf3e116

Observation c93a91c0-ced5-4d0f-b9c0-7e464d4e2201 · outbound

This paper cites Nofa: Nerf-based one-shot facial avatar recon- struction.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Nofa: Nerf-based one-shot facial avatar recon- struction

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.283323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.535992Z digest=sha256:9e2b6a6988416a5871b78ea83e93e08c8138d8d4cc4ce72c339441fb686c202a

Observation 2e539a54-fe2a-433d-aafa-2864bcae76bc · outbound

This paper cites Human-computer interaction system: A survey of talking-head generation.Electronics, 12(1):218, 2023.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Human-computer interaction system: A survey of talking-head generation.Electronics, 12(1):218, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:16.057653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.656839Z digest=sha256:a420a6113a5fdc7cfebcf4452df4185cd4aae6e00455d3840194acd10c79cff3

Observation bfb7cbfd-d1cd-4672-8cd5-d9f5676e63dd · outbound

This paper cites Responsive listening head generation: a benchmark dataset and baseline.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Responsive listening head generation: a benchmark dataset and baseline

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.836708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.761642Z digest=sha256:a23d09a8046f81d6953021760990403679b21fa3ccfb7279a7d24d5b36bb3f0e

Observation e6428501-3063-4c70-be21-f5a5aa78b528 · outbound

This paper cites Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:12.873023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:12.873023Z digest=sha256:e028b37015fddf44c096af8b90c076b3dc51aa4b6a4a80cfea6b6a3aa3c33e20

Observation 22127174-c4f6-4c6c-8004-ca10ebe3e35d · outbound

This paper cites Visemenet: Audio- driven animator-centric speech animation.ACM Transac- tions on Graphics (TOG), 37(4):1–10, 2018.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Visemenet: Audio- driven animator-centric speech animation.ACM Transac- tions on Graphics (TOG), 37(4):1–10, 2018

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.631924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:12.982202Z digest=sha256:2b9ffc958d4c6443869c014a2730cd9d28d45e11a85aca226a01b6155565abf6

Observation e4c46f00-d496-4697-8ba5-f8be58c03463 · outbound

This paper cites Network Architecture In this section, we provide comprehensive implementation details of our DualTalk framework.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Network Architecture In this section, we provide comprehensive implementation details of our DualTalk framework

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.443570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.092431Z digest=sha256:6ab7776665743c2fa8507690d43e823a0991f104869749792b12aa755cc017b9

Observation 9660e38e-2e40-4514-b2d1-b2349330c556 · outbound

This paper cites Here, we provide detailed in- formation about our data collection, processing procedures, and dataset statistics.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Here, we provide detailed in- formation about our data collection, processing procedures, and dataset statistics

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:14.955563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.359279Z digest=sha256:e8e0aa2d93475f16b2433415261533e4a5034c07f1ef80bceb4c848645329144

Observation 0b79ff21-5b4d-459d-8fac-6cb437bf2369 · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:14.715905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.474746Z digest=sha256:3632c2d8a3bec5df8c3f49867f802d3f6b0b617f49d9c008d39aa4adf5d8349e

Observation 0fa3e84a-56c3-4942-a4a8-81a4a20ccb57 · outbound

This paper cites an unresolved cited work.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:39:14.567364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.629376Z digest=sha256:a44854ee107f97a83af1d9396313e3b793c68cc76133768a56dc17d270bb9122

Observation 0bef61ec-2aa8-4f8d-9050-f704ba02b849 · outbound

This paper cites While DualTalk ex- cels in creating synchronized and natural two-speaker con- versations, it cannot yet handle multi-party interactions, which are common in real-world applications.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations While DualTalk ex- cels in creating synchronized and natural two-speaker con- versations, it cannot yet handle multi-party interactions, which are common in real-world applications

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:14.318967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.746751Z digest=sha256:3641340934a0e9b320f3615d040db9e7bbed31e7e06aaf55a19b93cde365728b

Observation 483660ca-40fb-420d-be30-2d4a6a2692a7 · outbound

This paper cites The decoder follows a similar structure but includes additional cross- attention layers to integrate information from both speakers.

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations The decoder follows a similar structure but includes additional cross- attention layers to integrate information from both speakers

Reference 512

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:15.192741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:39:13.227145Z digest=sha256:47789fc32b25f4ea42c2545795c149b7bd4adceae679bd55fdaa1ae2ed8c9c51

Pith citing papers

No inbound Pith citation observations are available.