Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:54:07.692944Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2505.23290.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:54:07.692944Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-15T08:44:49.087457Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T08:45:19.106137Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fb9c18dc-9ec7-48ec-964e-53082ce003f5 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Facetalk: Audio-driven motion diffusion for neu- ral parametric head models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 880fe63a-bdeb-4098-ac64-810dc5327e99 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Gesturediffuclip: Gesture diffusion model with clip latents
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d4a71ea-d624-4e29-a067-de88585b9d1d · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation wav2vec 2.0: A framework for self-supervised learning of speech representations
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 977188a8-b5d4-4ae9-bacd-4406c897afa0 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Expressive speech-driven facial animation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75aeae6e-a4ff-41c2-9523-564d9d338a9d · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Talking head generation with audio and speech related facial action units
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34bf5d04-6b3c-4cac-89c6-f140472b8119 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Peters, Michael J
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2db9501e-b2c4-40c2-9685-533be715a0f3 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b861fe77-0988-4cfc-84c1-3d4d68035ae6 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Emotional speech- driven animation with content-emotion disentanglement
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f361a3e-6cdf-4c5b-996e-ae2be75dc88c · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation BERT: pre-training of deep bidirectional trans- formers for language understanding
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4e84034-5fcc-4be4-88bd-f59e57c348cf · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Cross modal audio search and retrieval with joint embeddings based on text and audio
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3c3ab6f-bd83-452c-8259-e903f434c39a · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Unitalker: Scaling up audio-driven 3d facial animation through A unified model
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4275f2ce-7cbc-487b-91c7-c2bb3693ddf8 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Faceformer: Speech-driven 3d facial anima- tion with transformers
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f35f5ab2-2349-4993-b5cb-30f6c1d2c79e · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Joint audio-text model for expressive speech- driven 3d facial animation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7358179a-b5e2-466d-9ac3-41e30f57c03f · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Mimic: Speaking style disentanglement for speech-driven 3d facial animation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65055018-5466-4ab9-9fed-4aa94ed05350 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Deep Speech: Scaling up end-to-end speech recognition
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e964826-975d-4e95-b6c5-7fe2b55c6323 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Facexhubert: Text-less speech-driven e (x) pressive 3d facial animation synthesis using self-supervised speech representation learn- ing
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f512baf2-847a-4f41-9713-c95797b3b3b8 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Phonemic similarity metrics to compare pronunciation methods
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a46928a1-0165-48dc-adc8-8f7941280d8e · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eecda33e-2527-47ba-9aae-186f8dbaabf7 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Kingma and Jimmy Ba
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fec146e-fe99-423b-97d8-873b698ca677 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Mask-fpan: Semi-supervised face parsing in the wild with de-occlusion and uv gan
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c087cc0c-1812-41a4-8c06-d83571ade16d · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Omg: Towards open-vocabulary motion generation via mixture of controllers
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2732a15f-42c3-42d0-afa1-c77dc26619d7 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b09fbcc8-af9a-42df-9bea-3a8ed2adcf6c · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Convofusion: Multi-modal conversational diffu- sion for co-speech gesture synthesis
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a7ae7cd-79d9-47c0-8b7a-f82bf9c98c67 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Librispeech: An ASR corpus based on public domain audio books
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e573cb65-3a46-4b16-99f9-2d03cfab0138 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61c77f8b-5e1c-4a0b-be0f-fd65413d8565 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8266a45f-18e7-4439-856a-1e54ffee48e6 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Meshtalk: 3d face an- imation from speech using cross-modality disentanglement
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd2f1206-42e4-4598-9c91-74258ffc8680 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Expressive 3d facial animation generation based on local-to-global latent diffusion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b98f197b-bb53-44d7-9993-02bbaadae7f7 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Talkingstyle: Personalized speech-driven 3d facial animation with style preservation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06772a52-8deb-4a45-bc51-303597147eef · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Facediffuser: Speech-driven 3d facial animation synthesis using diffusion
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 791ea625-63bf-4195-96fd-ad4b096e68fd · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Sun and Li Deng
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e96ea9f-ce27-47dd-ba78-6515fd3d71dd · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Diff- posetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df5437b6-5604-4c32-abf7-e877c91d6c41 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Taylor, Taehwan Kim, Yisong Yue, Moshe Mahler, James Krahe, Anastasio Garcia Rodriguez, Jessica K
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a486deaa-0431-46a5-947c-fd63a0f7ca75 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation 3DiFACE: Diffusion-based Speech-driven 3D Facial Animation and Editing
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d4d6f0a-9126-4816-8bc8-bcc575928f07 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Imitator: Personalized speech-driven 3d facial animation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea41a362-d842-4fd4-98df-b3789ecea87e · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Neural voice puppetry: Audio-driven facial reenactment
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 060bec6f-0892-4851-8333-5a8919ee5369 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd18af0b-fee2-41f1-a369-5dba1b63a086 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation End-to-end speech-driven realistic facial animation with temporal gans
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d9de22c-4ae4-47f7-9b79-9f859a04a1f4 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation One- shot talking face generation from single-speaker audio-visual correlation learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44cbe3cd-bdcc-4584-b2da-c1c17694c446 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Codetalker: Speech-driven 3d facial animation with discrete motion prior
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37acc315-2d75-487b-ad54-50a0676394f1 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Feng, Stacy Marsella, and Ari Shapiro
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ff9156e-1d75-40b2-886e-aac50714b777 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation You only speak once to see
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a57b2918-b078-455e-96e5-8d9bd0cce1d0 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Mining audio, text and visual information for talking face generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 250eb052-f068-48e3-ae08-fc2ee0989261 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Motiondif- fuse: Text-driven human motion generation with diffusion model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d34f567-4d6b-45b9-b940-0482e0c8f529 · outbound
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation Media2face: Co-speech facial animation gener- ation with multi-modality guidance
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e50333f4-7b0c-4c81-a95b-b4b2d291f5ec · inbound
RAM: Recover Any 3D Human Motion in-the-Wild Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.