Pith. sign in

REVIEW

Efficient Trainable Front-Ends for Neural Speech Enhancement

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.09286 v1 pith:BENUCBQX submitted 2020-02-20 eess.AS cs.LGcs.NEcs.SDstat.ML

classification eess.AScs.LGcs.NEcs.SDstat.ML
keywords trainableenhancementfourierneuralspeechtransformefficientfront-ends
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Many neural speech enhancement and source separation systems operate in the time-frequency domain. Such models often benefit from making their Short-Time Fourier Transform (STFT) front-ends trainable. In current literature, these are implemented as large Discrete Fourier Transform matrices; which are prohibitively inefficient for low-compute systems. We present an efficient, trainable front-end based on the butterfly mechanism to compute the Fast Fourier Transform, and show its accuracy and efficiency benefits for low-compute neural speech enhancement models. We also explore the effects of making the STFT window trainable.

Discussion (0). Continue with ORCID to comment.

Pith tools