Training a talking-face generator with an auxiliary uncertainty module that predicts its own pixel errors and matches error and uncertainty histograms improves image-quality metrics, though not lip-sync relative to Wav2Lip.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Audio-Driven Talking Face Video Generation with Joint Uncertainty Learning
Training a talking-face generator with an auxiliary uncertainty module that predicts its own pixel errors and matches error and uncertainty histograms improves image-quality metrics, though not lip-sync relative to Wav2Lip.