TurboTalk uses progressive distillation from 4 steps to 1 step with distribution matching and adversarial training to achieve 120x faster single-step audio-driven talking avatar video generation.
Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation
2 Pith papers cite this work. Polarity classification is still indexing.
fields
cs.CV 2years
2026 2representative citing papers
Applying Encoding-Decoding Direction Pairs to an Xception deepfake detector reveals 16 interpretable concepts (e.g., fake-mouth, real-eyes) that drive real/fake predictions, with concept-level interventions achieving 99.8% correction of misclassified samples.
citing papers explorer
-
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
TurboTalk uses progressive distillation from 4 steps to 1 step with distribution matching and adversarial training to achieve 120x faster single-step audio-driven talking avatar video generation.
-
Why Fake ? Unveiling the Semantic Vocabulary of Deepfake Detectors
Applying Encoding-Decoding Direction Pairs to an Xception deepfake detector reveals 16 interpretable concepts (e.g., fake-mouth, real-eyes) that drive real/fake predictions, with concept-level interventions achieving 99.8% correction of misclassified samples.