On the NIST SRE24 audio track, a ResNet152 pre-trained on 8kHz/GSM-augmented VoxBlink2 and fine-tuned on telephone speech with 40s segments achieved the best EER and Cprimary among the tested frontends.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Analysis of ABC Frontend Audio Systems for the NIST-SRE24
On the NIST SRE24 audio track, a ResNet152 pre-trained on 8kHz/GSM-augmented VoxBlink2 and fine-tuned on telephone speech with 40s segments achieved the best EER and Cprimary among the tested frontends.