Pith. sign in

Paper Citation Record · LEDGER

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation

As of 21 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2507.19225.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19225 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.409083Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.313393Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T18:00:25.553973Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy19
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2c27026a-919e-41fc-8f9f-5703037ab91b · outbound

This paper cites Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:00:25.556673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.313393Z digest=sha256:6dfb97fc628f4af717f1a9a3f146ee83147f7e02c35d2d2254bdfd32cf4ded0a

Observation 83e6ac62-8308-42da-96c8-27038d8d0981 · outbound

This paper cites Early works, such as Chen et al.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Early works, such as Chen et al

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.717577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.317314Z digest=sha256:1e3af63184aa419df0f4e9d806f3a2f7e2318170d67f94a47a08b48c69e9e379

Observation 22efcff7-672d-4d90-b23a-2671ebb36325 · outbound

This paper cites The overall pipeline is illustrated in Figure 2.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The overall pipeline is illustrated in Figure 2

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.710970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.321243Z digest=sha256:2393307b091bef3f90e78d244f9da0d0981be56c255d6c679cc2e1b6969cebab

Observation 87414495-88b1-4b7b-967a-b5bb12f9c997 · outbound

This paper cites Experimental Settings We conduct our experiments using LRS2 [25] and HDTF [26] datasets.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Experimental Settings We conduct our experiments using LRS2 [25] and HDTF [26] datasets

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.704335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.325105Z digest=sha256:880af9689ae67b3f80d552fb175f20e26403b70c31099c5d80d6ad82ae287f0d

Observation 7e6a5e38-357c-483e-b70b-5f3f05211d08 · outbound

This paper cites Unlike prior works assuming a fixed face-to-voice mapping, we model it as a probability distribution problem, capturing natural voice variability.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Unlike prior works assuming a fixed face-to-voice mapping, we model it as a probability distribution problem, capturing natural voice variability

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.691237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.330920Z digest=sha256:463c516fc19d4b3749002e76a707f5ab8ec1c43a656f45f5d14fdd37c9b0e6d9

Observation 9ee54083-5cf3-4e22-a893-571590461611 · outbound

This paper cites The authors also acknowledge CSC-IT Center for Science, Finland, for pro- viding computational resources.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The authors also acknowledge CSC-IT Center for Science, Finland, for pro- viding computational resources

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.683902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.333729Z digest=sha256:51e92aa3e50f862c8cc53bc1ebad01a400c111ace216e6322aaaa229929e7719

Observation 722461cd-aaf0-4b4a-9b9f-61a6be60a579 · outbound

This paper cites Faces that speak: Jointly syn- thesising talking face and speech from text,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Faces that speak: Jointly syn- thesising talking face and speech from text,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.654871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.352417Z digest=sha256:1f5310ddf59ce9a12aad1951c18ae6ea310775f75b699a8ef329b89145535936

Observation 528e5442-e7fe-46db-9ce9-340f80441f27 · outbound

This paper cites Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.336103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.336103Z digest=sha256:183795950fc2aded1763953eea953722e3cf80f680c1c6286a79a0e738530198

Observation de75e7c6-801e-40fe-b406-d591e6d1d665 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.676872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.338993Z digest=sha256:ec63b753f4f22dbc4ec67cb18d54e9f69670bad1e1e29ecccdeaab8c4c032bef

Observation 56528005-6330-4514-844e-c815f4f267aa · outbound

This paper cites DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.341577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.341577Z digest=sha256:93248f9241e268abbce2340c8fc7fac386c9aead850f6eab5b70be795f6a596b

Observation 148e966c-0294-4380-b1ef-f250356590f6 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.344468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.344468Z digest=sha256:c646c7be30bb168e9459ebec8d9584a6280db0753579921e8f7c06a928ea9ec5

Observation 6948bfb9-8c77-415c-8015-912c09564061 · outbound

This paper cites Edtalk: Efficient disentangle- ment for emotional talking head synthesis,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Edtalk: Efficient disentangle- ment for emotional talking head synthesis,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.669822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.347286Z digest=sha256:66de9c19d22aa2d1ae428adf2c88c5aa0c6c136cf73c01b2388661f62cef234b

Observation d7ff594d-5dbd-4e7e-8794-aa9fdee2c6c6 · outbound

This paper cites Text2video: Text- driven talking-head video synthesis with personalized phoneme- pose dictionary,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text2video: Text- driven talking-head video synthesis with personalized phoneme- pose dictionary,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.663148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.350020Z digest=sha256:362ad812f5220c530ca699e4bb724d7e44102716f113bef54e03d762a2bfd523

Observation 67c19a2b-da6e-467c-9d6d-4c26437f3c53 · outbound

This paper cites Hierarchical cross- modal talking face generation with dynamic pixel-wise loss,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Hierarchical cross- modal talking face generation with dynamic pixel-wise loss,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.626120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.370700Z digest=sha256:9285bc94c5e2af495a52922524591f5ea9f3f6bcbcf58e6496ac173629b1cc3c

Observation bc2c698a-8c6d-40c9-8b93-1ca433a53ebb · outbound

This paper cites Text-to-Video: a Two-stage Framework for Zero-shot Identity-agnostic Talking-head Generation.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text-to-Video: a Two-stage Framework for Zero-shot Identity-agnostic Talking-head Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.354763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.354763Z digest=sha256:e1ffb1735b8ea6662ccf047f5ae583304a002e2b3398205f249f79b9ee1f8956

Observation 84a5a3ca-24a1-4903-a9da-8eab5a86c72c · outbound

This paper cites Ada-TTA: Towards Adaptive High-Quality Text-to-Talking Avatar Synthesis.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Ada-TTA: Towards Adaptive High-Quality Text-to-Talking Avatar Synthesis

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:00:25.524927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.357330Z digest=sha256:91e0911345ea1f63b4bcf961f8752a1ef31b2f294552e7455f7d9434d44dfdcb

Observation 6f5660d3-f3e6-4c6f-a1f7-5f6c7198310e · outbound

This paper cites Uniflg: Unified facial land- mark generator from text or speech,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Uniflg: Unified facial land- mark generator from text or speech,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.647710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.360567Z digest=sha256:e011d7fbbe6c6a9524c980d98c4656fdc51d2035f18644e6c393509e3ef76ccc

Observation 184ab4eb-47c4-46cc-bc0a-bd461d4cabc9 · outbound

This paper cites The results are presented in Table 2.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation The results are presented in Table 2

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T18:00:25.697841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.327800Z digest=sha256:d9260552ce77534f69559db0712dd7e7a1dc1acec3f581795ad3d67432c9bbf8

Observation 46042e3d-83f1-4f93-8279-30356db2fe67 · outbound

This paper cites Text-driven talk- ing face synthesis by reprogramming audio-driven models,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Text-driven talk- ing face synthesis by reprogramming audio-driven models,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.640584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.362816Z digest=sha256:90477ba3d8cd30e22a487cd0fc35deaa379cea079660f4bfea8e6d81dc9dfc65

Observation 20c4eb9e-8899-47ae-9cc1-30283cd3ef84 · outbound

This paper cites Generating talking face with controllable eye movements by disentangled blinking feature,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Generating talking face with controllable eye movements by disentangled blinking feature,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.633446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.365391Z digest=sha256:f24ea17d3c89a08a79adc11da27acef5135468f00b29bb960c5f5107a93a59a8

Observation 09c8d03e-9910-49b2-92a7-c4a79a389e9e · outbound

This paper cites FT2TF: First-Person Statement Text-To-Talking Face Generation.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation FT2TF: First-Person Statement Text-To-Talking Face Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.367836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.367836Z digest=sha256:423f0e4f2e7f96ad2e5c10086908c305c5be97e4b5eda3a23dd5add79ac28996

Observation 8782ce8f-2e4e-45fe-878e-4b902f3e400b · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A lip sync expert is all you need for speech to lip generation in the wild,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.373254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.373254Z digest=sha256:b51ef743ac441c06d5c4e773cade37822ddf57a3932f94998994050b58fb0603

Observation 6ffc3daf-e65c-4dc4-98aa-2c2bd46ebec3 · outbound

This paper cites GAIA: Zero-shot talking avatar gener- ation,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation GAIA: Zero-shot talking avatar gener- ation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.615387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.376531Z digest=sha256:3b47aa2e8feaa1f5c643fd25fdbb271e633e0b77a4f299adf2a7ab16728506d5

Observation ca890290-afe7-4e63-a4e0-3a71afb04cd9 · outbound

This paper cites Por- traittalk: Towards customizable one-shot audio-to-talking face generation,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Por- traittalk: Towards customizable one-shot audio-to-talking face generation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.379263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.379263Z digest=sha256:550b5a503618fcb2d4d39cf8f26c2e3f27400bef44e7cc25283f6fa366e3ba7a

Observation 00b847a5-6355-4a1a-8901-76a6684e56db · outbound

This paper cites CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.382659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.382659Z digest=sha256:b5a6b6ab60c4aa87055c64316f559337be3c045c221439c7b86e0d64514e8307

Observation 5d8ac069-c4e6-4f8e-88fa-5f9b1cfe73e3 · outbound

This paper cites Imaginary voice: Face- styled diffusion model for text-to-speech,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Imaginary voice: Face- styled diffusion model for text-to-speech,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.385283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.385283Z digest=sha256:f92b8cd4be2baf3ceed532e061e80d27646da586988401dca32b9d7c0a760e06

Observation 18510ab0-9738-42ca-967c-bf720c881921 · outbound

This paper cites Fvtts: Face based voice synthesis for text-to-speech,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Fvtts: Face based voice synthesis for text-to-speech,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.605245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.387864Z digest=sha256:53b2944b4f4c0c9fb8a90cc449fd17787374123d802f105fb93775fce9ab71ad

Observation a5a81537-2d25-4ce3-83ce-33b928c947ee · outbound

This paper cites A kernel two-sample test,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A kernel two-sample test,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.390797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.390797Z digest=sha256:a698c72ebbcd0b5a45f0ad9b41d61a222b404450139e616d173bc9cda2484668

Observation 0e4c43a4-0897-4631-ad91-db557d188078 · outbound

This paper cites CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.393215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.393215Z digest=sha256:d17e9ed716300be290e6895560b8b27a407ae312f46f340b5fc75f8bd6a003e3

Observation d27484d8-f90a-420b-b452-ce83f1e295f8 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representa- tions,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation wav2vec 2.0: A framework for self-supervised learning of speech representa- tions,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.593588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.396518Z digest=sha256:41731311e4fff126241ee70e7728858d37798d7929938a3a05006e9965519795

Observation 8a58124e-3080-420f-b763-787bc99da613 · outbound

This paper cites A low-complexity permutation alignment method for frequency-domain blind source separation,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation A low-complexity permutation alignment method for frequency-domain blind source separation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.586667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.398894Z digest=sha256:2647c292daafd3c1b14f14c437f35f3bf817ee6186ba380d11fa59c772bff5fd

Observation ef4bcd4f-944b-4239-b044-f38775ac6371 · outbound

This paper cites Lip read- ing sentences in the wild,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Lip read- ing sentences in the wild,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.577289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.401287Z digest=sha256:211e48157591047d85f7daa8566c2045f790876f970579246befb08d6b7674e6

Observation f19735e6-1e20-4659-95aa-7e1a785bf69b · outbound

This paper cites Flow-guided one-shot talk- ing face generation with a high-resolution audio-visual dataset,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Flow-guided one-shot talk- ing face generation with a high-resolution audio-visual dataset,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:25.569588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.403521Z digest=sha256:cd801c22e8167e208d4f553ae22aa9643d1aa9917180c21b2c23946cafb33f52

Observation 6b5e534e-7341-447d-a3ce-b64e0398625b · outbound

This paper cites On estimation of a probability density function and mode,.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation On estimation of a probability density function and mode,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.405874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.405874Z digest=sha256:ede59d7b0c74228515301c6cc521602caea17ddd13abe6d67553fca77cb550e5

Observation d65e1152-961a-4841-97ab-07921088c082 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Adam: A Method for Stochastic Optimization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:25.409083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:00:25.409083Z digest=sha256:3cb1f76e2d0e16367bf53f942448eedf44ac17e392424927718daeaf8088e421

Pith citing papers

Observation 2c27026a-919e-41fc-8f9f-5703037ab91b · inbound

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation cites this paper.

Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:00:25.556673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T18:00:25.313393Z digest=sha256:6dfb97fc628f4af717f1a9a3f146ee83147f7e02c35d2d2254bdfd32cf4ded0a