Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.409083Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2507.19225.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.409083Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.313393Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T18:00:25.553973Z
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2c27026a-919e-41fc-8f9f-5703037ab91b · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 83e6ac62-8308-42da-96c8-27038d8d0981 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Early works, such as Chen et al
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 22efcff7-672d-4d90-b23a-2671ebb36325 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The overall pipeline is illustrated in Figure 2
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87414495-88b1-4b7b-967a-b5bb12f9c997 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Experimental Settings We conduct our experiments using LRS2 [25] and HDTF [26] datasets
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7e6a5e38-357c-483e-b70b-5f3f05211d08 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Unlike prior works assuming a fixed face-to-voice mapping, we model it as a probability distribution problem, capturing natural voice variability
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9ee54083-5cf3-4e22-a893-571590461611 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The authors also acknowledge CSC-IT Center for Science, Finland, for pro- viding computational resources
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 722461cd-aaf0-4b4a-9b9f-61a6be60a579 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Faces that speak: Jointly syn- thesising talking face and speech from text,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 528e5442-e7fe-46db-9ce9-340f80441f27 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de75e7c6-801e-40fe-b406-d591e6d1d665 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 56528005-6330-4514-844e-c815f4f267aa · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 148e966c-0294-4380-b1ef-f250356590f6 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6948bfb9-8c77-415c-8015-912c09564061 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Edtalk: Efficient disentangle- ment for emotional talking head synthesis,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d7ff594d-5dbd-4e7e-8794-aa9fdee2c6c6 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text2video: Text- driven talking-head video synthesis with personalized phoneme- pose dictionary,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 67c19a2b-da6e-467c-9d6d-4c26437f3c53 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Hierarchical cross- modal talking face generation with dynamic pixel-wise loss,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bc2c698a-8c6d-40c9-8b93-1ca433a53ebb · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text-to-Video: a Two-stage Framework for Zero-shot Identity-agnostic Talking-head Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84a5a3ca-24a1-4903-a9da-8eab5a86c72c · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Ada-TTA: Towards Adaptive High-Quality Text-to-Talking Avatar Synthesis
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6f5660d3-f3e6-4c6f-a1f7-5f6c7198310e · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Uniflg: Unified facial land- mark generator from text or speech,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 184ab4eb-47c4-46cc-bc0a-bd461d4cabc9 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The results are presented in Table 2
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 46042e3d-83f1-4f93-8279-30356db2fe67 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text-driven talk- ing face synthesis by reprogramming audio-driven models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 20c4eb9e-8899-47ae-9cc1-30283cd3ef84 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Generating talking face with controllable eye movements by disentangled blinking feature,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 09c8d03e-9910-49b2-92a7-c4a79a389e9e · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation FT2TF: First-Person Statement Text-To-Talking Face Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8782ce8f-2e4e-45fe-878e-4b902f3e400b · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A lip sync expert is all you need for speech to lip generation in the wild,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ffc3daf-e65c-4dc4-98aa-2c2bd46ebec3 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation GAIA: Zero-shot talking avatar gener- ation,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ca890290-afe7-4e63-a4e0-3a71afb04cd9 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Por- traittalk: Towards customizable one-shot audio-to-talking face generation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00b847a5-6355-4a1a-8901-76a6684e56db · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d8ac069-c4e6-4f8e-88fa-5f9b1cfe73e3 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Imaginary voice: Face- styled diffusion model for text-to-speech,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18510ab0-9738-42ca-967c-bf720c881921 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Fvtts: Face based voice synthesis for text-to-speech,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a5a81537-2d25-4ce3-83ce-33b928c947ee · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A kernel two-sample test,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e4c43a4-0897-4631-ad91-db557d188078 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d27484d8-f90a-420b-b452-ce83f1e295f8 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation wav2vec 2.0: A framework for self-supervised learning of speech representa- tions,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8a58124e-3080-420f-b763-787bc99da613 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A low-complexity permutation alignment method for frequency-domain blind source separation,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ef4bcd4f-944b-4239-b044-f38775ac6371 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Lip read- ing sentences in the wild,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f19735e6-1e20-4659-95aa-7e1a785bf69b · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Flow-guided one-shot talk- ing face generation with a high-resolution audio-visual dataset,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6b5e534e-7341-447d-a3ce-b64e0398625b · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation On estimation of a probability density function and mode,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d65e1152-961a-4841-97ab-07921088c082 · outbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Adam: A Method for Stochastic Optimization
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c27026a-919e-41fc-8f9f-5703037ab91b · inbound
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.