Pith. sign in

REVIEW

Continuous Speech Separation with Conformer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.05773 v2 pith:2GFPLOXO submitted 2020-08-13 eess.AS cs.CL

classification eess.AScs.CL
keywords separationspeechconformercontinuousevaluationmodelreductionachieves
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Continuous speech separation plays a vital role in complicated speech related tasks such as conversation transcription. The separation model extracts a single speaker signal from a mixed speech. In this paper, we use transformer and conformer in lieu of recurrent neural networks in the separation system, as we believe capturing global information with the self-attention based method is crucial for the speech separation. Evaluating on the LibriCSS dataset, the conformer separation model achieves state of the art results, with a relative 23.5% word error rate (WER) reduction from bi-directional LSTM (BLSTM) in the utterance-wise evaluation and a 15.4% WER reduction in the continuous evaluation.

Discussion (0). Sign in to comment.

Pith tools