Pith. sign in

Talkvid: A large-scale diversified dataset for audio-driven talking head synthesis

10 Pith papers cite this work. Polarity classification is still indexing.

10 Pith papers citing it

citation-role summary

dataset 1

citation-polarity summary

years

2026 9 2025 1

roles

dataset 1

polarities

use dataset 1

representative citing papers

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

cs.CV · 2025-11-29 · conditional · novelty 7.0

MVAD is the first comprehensive benchmark dataset for AI-generated multimodal video-audio detection, with three realistic forgery patterns, high-quality outputs from state-of-the-art models, and diversity across visual styles and content categories.

Generate Your Talking Avatar from Video Reference

cs.CV · 2026-04-30 · unverdicted · novelty 6.0

TAVR generates high-fidelity talking avatars from cross-scene video references via token selection and three-stage training (same-scene pretraining, cross-scene fine-tuning, identity RL), outperforming baselines on a new 158-pair benchmark.

citing papers explorer

Showing 10 of 10 citing papers.