A Fast-Converged Acoustic Modeling for Korean Speech Recognition: A Preliminary Study on Time Delay Neural Network

Donghyun Lee; Hosung Park; Ji-Hwan Kim; Juneseok Oh; Minkyu Lim; Yoseb Kang

arxiv: 1807.05855 · v1 · pith:XGWHWGSTnew · submitted 2018-07-11 · 💻 cs.CL · cs.SD· eess.AS

A Fast-Converged Acoustic Modeling for Korean Speech Recognition: A Preliminary Study on Time Delay Neural Network

Hosung Park , Donghyun Lee , Minkyu Lim , Yoseb Kang , Juneseok Oh , Ji-Hwan Kim This is my paper

classification 💻 cs.CL cs.SDeess.AS

keywords acoustickoreanmodelnetworkneuralspeechtdnndelay

0 comments

read the original abstract

In this paper, a time delay neural network (TDNN) based acoustic model is proposed to implement a fast-converged acoustic modeling for Korean speech recognition. The TDNN has an advantage in fast-convergence where the amount of training data is limited, due to subsampling which excludes duplicated weights. The TDNN showed an absolute improvement of 2.12% in terms of character error rate compared to feed forward neural network (FFNN) based modelling for Korean speech corpora. The proposed model converged 1.67 times faster than a FFNN-based model did.

This paper has not been read by Pith yet.

A Fast-Converged Acoustic Modeling for Korean Speech Recognition: A Preliminary Study on Time Delay Neural Network

discussion (0)