REVIEW 3 cited by
Are audio DeepFake detection models polyglots?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Since the majority of audio DeepFake (DF) detection methods are trained on English-centric datasets, their applicability to non-English languages remains largely unexplored. In this work, we present a benchmark for the multilingual audio DF detection challenge by evaluating various adaptation strategies. Our experiments focus on analyzing models trained on English benchmark datasets, as well as intra-linguistic (same-language) and cross-linguistic adaptation approaches. Our results indicate considerable variations in detection efficacy, highlighting the difficulties of multilingual settings. We show that limiting the dataset to English negatively impacts the efficacy, while stressing the importance of the data in the target language.
Forward citations
Cited by 3 Pith papers
-
Tell me Habibi, is it Real or Fake?
ArEnAV, the first large-scale Arabic-English code-switched audio-visual deepfake dataset, makes current state-of-the-art detectors fail much more than on monolingual data.
-
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
The first multilingual source-tracing benchmark for speech deepfakes, showing LFCC-ECAPA-TDNN generalizes best across languages.
-
Multi-level SSL Feature Gating for Audio Deepfake Detection
An XLS-R based audio deepfake detector combining gated multi-kernel convolutions with a CKA dissimilarity loss reports top EERs on 19LA, 21DF, and In-The-Wild benchmarks.
Discussion (0). Sign in to comment.