Pith. sign in

REVIEW 5 cited by

Explainable Attribute-Based Speaker Verification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.19796 v1 pith:GH4MLVE6 submitted 2024-05-30 cs.SD cs.AIeess.AS

Explainable Attribute-Based Speaker Verification

classification cs.SD cs.AIeess.AS
keywords speakerapproachattributesexplainableattribute-basedbelievemethodsperformance
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

This paper proposes a fully explainable approach to speaker verification (SV), a task that fundamentally relies on individual speaker characteristics. The opaque use of speaker attributes in current SV systems raises concerns of trust. Addressing this, we propose an attribute-based explainable SV system that identifies speakers by comparing personal attributes such as gender, nationality, and age extracted automatically from voice recordings. We believe this approach better aligns with human reasoning, making it more understandable than traditional methods. Evaluated on the Voxceleb1 test set, the best performance of our system is comparable with the ground truth established when using all correct attributes, proving its efficacy. Whilst our approach sacrifices some performance compared to non-explainable methods, we believe that it moves us closer to the goal of transparent, interpretable AI and lays the groundwork for future enhancements through attribute expansion.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Why Do You Say It Like That? A Phoneme-Level Framework for Explainable Speech Deepfake Detection

    eess.AS 2026-07 conditional novelty 6.0

    Phoneme-aligned Grad-CAM on a WavLM-CNN detector reveals significant attack- and speaker-dependent importance of vowels, fricatives and pauses for spoof vs bona-fide decisions on ASVspoof 5.

  2. Towards Dys-XAI: Influence-Based Explanations for Dysarthria Severity Assessment

    cs.AI 2026-06 unverdicted novelty 6.0

    Introduces an instance-level influence-based XAI method for dysarthria severity assessment that explains predictions by computing per-utterance influence scores from training samples and validates them via controlled ...

  3. SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

    cs.SD 2026-05 unverdicted novelty 6.0

    SpeakerLLM unifies speaker profiling, recording-condition understanding, and structured verification reasoning in an audio-LLM via a hierarchical tokenizer and decision traces.

  4. PhiNet: Speaker Verification with Phonetic Interpretability

    eess.AS 2026-04 unverdicted novelty 6.0

    PhiNet adds phonetic interpretability to speaker verification while matching the accuracy of standard black-box models on VoxCeleb, SITW, and LibriSpeech.

  5. LISE : Listenable Interpretable Speaker Embeddings

    cs.SD 2026-06 unverdicted novelty 5.0

    LISE decomposes pretrained speaker embeddings into components that preserve ASV performance with negligible EER degradation and enable listeners to distinguish speakers at 83.9% accuracy.