Pith. sign in

REVIEW

Deep Reinforcement Learning for On-line Dialogue State Tracking

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2009.10321 v1 pith:R2ZTLU2V submitted 2020-09-22 cs.CL

classification cs.CL
keywords dialogueon-lineoptimizationpolicydeepframeworkfurtherimprove
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Dialogue state tracking (DST) is a crucial module in dialogue management. It is usually cast as a supervised training problem, which is not convenient for on-line optimization. In this paper, a novel companion teaching based deep reinforcement learning (DRL) framework for on-line DST optimization is proposed. To the best of our knowledge, this is the first effort to optimize the DST module within DRL framework for on-line task-oriented spoken dialogue systems. In addition, dialogue policy can be further jointly updated. Experiments show that on-line DST optimization can effectively improve the dialogue manager performance while keeping the flexibility of using predefined policy. Joint training of both DST and policy can further improve the performance.

Discussion (0). Sign in to comment.

Pith tools