A compact 1D-CNN trained on FFT coefficients and noise-augmented clips reaches 97.87% validation accuracy on a small, same-session custom speaker dataset.
Voxceleb: alarge-scalespeakeridentificationdataset,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Towards Speaker Identification with Minimal Dataset and Constrained Resources using 1D-Convolution Neural Network
A compact 1D-CNN trained on FFT coefficients and noise-augmented clips reaches 97.87% validation accuracy on a small, same-session custom speaker dataset.