Pith. sign in

REVIEW 1 cited by

Adaptive Natural Language Generation for Task-oriented Dialogue via Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.07873 v1 pith:BXQLFBBG submitted 2022-09-16 cs.CL cs.AI

classification cs.CLcs.AI
keywords languagenaturalutterancesdialoguegenerationadaptiveantorlearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

When a natural language generation (NLG) component is implemented in a real-world task-oriented dialogue system, it is necessary to generate not only natural utterances as learned on training data but also utterances adapted to the dialogue environment (e.g., noise from environmental sounds) and the user (e.g., users with low levels of understanding ability). Inspired by recent advances in reinforcement learning (RL) for language generation tasks, we propose ANTOR, a method for Adaptive Natural language generation for Task-Oriented dialogue via Reinforcement learning. In ANTOR, a natural language understanding (NLU) module, which corresponds to the user's understanding of system utterances, is incorporated into the objective function of RL. If the NLG's intentions are correctly conveyed to the NLU, which understands a system's utterances, the NLG is given a positive reward. We conducted experiments on the MultiWOZ dataset, and we confirmed that ANTOR could generate adaptive utterances against speech recognition errors and the different vocabulary levels of users.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Measuring How (Not Just Whether) VLMs Build Common Ground

    cs.CL 2025-09 conditional novelty 7.0 of 10

    A four-metric suite shows VLM self-play in referential games diverges from human grounding patterns, with GPT4o-mini closest and task success not implying common ground.

Pith tools