pith. sign in

arxiv: 1806.04327 · v1 · pith:K7Z6Z2PUnew · submitted 2018-06-12 · 💻 cs.CL

ISO-Standard Domain-Independent Dialogue Act Tagging for Conversational Agents

classification 💻 cs.CL
keywords dialogueannotationavailableconversationalcorporacorpusdatadifferent
0
0 comments X
read the original abstract

Dialogue Act (DA) tagging is crucial for spoken language understanding systems, as it provides a general representation of speakers' intents, not bound to a particular dialogue system. Unfortunately, publicly available data sets with DA annotation are all based on different annotation schemes and thus incompatible with each other. Moreover, their schemes often do not cover all aspects necessary for open-domain human-machine interaction. In this paper, we propose a methodology to map several publicly available corpora to a subset of the ISO standard, in order to create a large task-independent training corpus for DA classification. We show the feasibility of using this corpus to train a domain-independent DA tagger testing it on out-of-domain conversational data, and argue the importance of training on multiple corpora to achieve robustness across different DA categories.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Towards Universal Dialogue Act Tagging for Task-Oriented Dialogues

    cs.CL 2019-07 unverdicted novelty 6.0

    The authors define a universal dialogue act schema, align several task-oriented dialogue datasets to it, and report a tagger reaching 54.1% F1 unsupervised and 57.7% semi-supervised on human-human dialogues.