REVIEW 2 cited by
Continuously Learning Neural Dialogue Management
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We describe a two-step approach for dialogue management in task-oriented spoken dialogue systems. A unified neural network framework is proposed to enable the system to first learn by supervision from a set of dialogue data and then continuously improve its behaviour via reinforcement learning, all using gradient-based algorithms on one single model. The experiments demonstrate the supervised model's effectiveness in the corpus-based evaluation, with user simulation, and with paid human subjects. The use of reinforcement learning further improves the model's performance in both interactive settings, especially under higher-noise conditions.
Forward citations
Cited by 2 Pith papers
-
Towards End-to-End Learning for Efficient Dialogue Agent by Modeling Looking-ahead Ability
A supervised end-to-end dialogue model with a bidirectional 'looking-ahead' module predicts future turns to guide response generation, showing modest and inconsistent gains on two datasets.
-
LSTM vs. GRU vs. Bidirectional RNN for script generation
A case study comparing LSTM, GRU and Bidirectional RNN for character-level TV script generation, with internally inconsistent reported results.
Discussion (0). Continue with ORCID to comment.