Pith. sign in

REVIEW 1 cited by

One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.16716 v1 pith:2XN3PHIU submitted 2024-06-24 eess.AS cs.CRcs.SD

One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection

classification eess.AS cs.CRcs.SD
keywords bonafidecentroidsystemsdeepfakelearningmethodone-classadaptive
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

As speech synthesis systems continue to make remarkable advances in recent years, the importance of robust deepfake detection systems that perform well in unseen systems has grown. In this paper, we propose a novel adaptive centroid shift (ACS) method that updates the centroid representation by continually shifting as the weighted average of bonafide representations. Our approach uses only bonafide samples to define their centroid, which can yield a specialized centroid for one-class learning. Integrating our ACS with one-class learning gathers bonafide representations into a single cluster, forming well-separated embeddings robust to unseen spoofing attacks. Our proposed method achieves an equal error rate (EER) of 2.19% on the ASVspoof 2021 deepfake dataset, outperforming all existing systems. Furthermore, the t-SNE visualization illustrates that our method effectively maps the bonafide embeddings into a single cluster and successfully disentangles the bonafide and spoof classes.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Layer-Wise Decision Fusion for Fake Audio Detection Using XLS-R

    cs.SD 2026-07 conditional novelty 5.0

    Per-layer late fusion of one-class softmax classifiers on frozen XLS-R features achieves 6.90% EER on In-the-Wild fake-audio detection, outperforming feature-fusion and single-layer baselines.