Pith. sign in

REVIEW

Accelerating Natural Language Understanding in Task-Oriented Dialog

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2006.03701 v1 pith:47UGJRLN submitted 2020-06-05 cs.CL cs.LG

classification cs.CLcs.LG
keywords dialoglanguagemodelmodelsnaturalparameterstask-orientedunderstanding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Task-oriented dialog models typically leverage complex neural architectures and large-scale, pre-trained Transformers to achieve state-of-the-art performance on popular natural language understanding benchmarks. However, these models frequently have in excess of tens of millions of parameters, making them impossible to deploy on-device where resource-efficiency is a major concern. In this work, we show that a simple convolutional model compressed with structured pruning achieves largely comparable results to BERT on ATIS and Snips, with under 100K parameters. Moreover, we perform acceleration experiments on CPUs, where we observe our multi-task model predicts intents and slots nearly 63x faster than even DistilBERT.

Discussion (0). Sign in to comment.

Pith tools