Precision Scaling of Neural Networks for Efficient Audio Processing

Ivan Tashev; Jong Hwan Ko; Josh Fromm; Matthai Philipose; Shuayb Zarar

arxiv: 1712.01340 · v1 · pith:SEF4CS4Bnew · submitted 2017-12-04 · 📡 eess.AS · cs.SD

Precision Scaling of Neural Networks for Efficient Audio Processing

Jong Hwan Ko , Josh Fromm , Matthai Philipose , Ivan Tashev , Shuayb Zarar This is my paper

classification 📡 eess.AS cs.SD

keywords processingnetworksneuralperformanceprecisionaudioimpactdeep

0 comments

read the original abstract

While deep neural networks have shown powerful performance in many audio applications, their large computation and memory demand has been a challenge for real-time processing. In this paper, we study the impact of scaling the precision of neural networks on the performance of two common audio processing tasks, namely, voice-activity detection and single-channel speech enhancement. We determine the optimal pair of weight/neuron bit precision by exploring its impact on both the performance and processing time. Through experiments conducted with real user data, we demonstrate that deep neural networks that use lower bit precision significantly reduce the processing time (up to 30x). However, their performance impact is low (< 3.14%) only in the case of classification tasks such as those present in voice activity detection.

This paper has not been read by Pith yet.

Precision Scaling of Neural Networks for Efficient Audio Processing

discussion (0)