Pith. sign in

DialoGLUE: A Natural Language Understanding Benchmark for Task-Oriented Dialogue

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

A long-standing goal of task-oriented dialogue research is the ability to flexibly adapt dialogue models to new domains. To progress research in this direction, we introduce DialoGLUE (Dialogue Language Understanding Evaluation), a public benchmark consisting of 7 task-oriented dialogue datasets covering 4 distinct natural language understanding tasks, designed to encourage dialogue research in representation-based transfer, domain adaptation, and sample-efficient task learning. We release several strong baseline models, demonstrating performance improvements over a vanilla BERT architecture and state-of-the-art results on 5 out of 7 tasks, by pre-training on a large open-domain dialogue corpus and task-adaptive self-supervised training. Through the DialoGLUE benchmark, the baseline methods, and our evaluation scripts, we hope to facilitate progress towards the goal of developing more general task-oriented dialogue models.

citation-role summary

background 1

citation-polarity summary

fields

cs.CL 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Few-Shot Query Intent Detection via Relation-Aware Prompt Learning

cs.CL · 2025-09-06 · conditional · novelty 6.0

SAID pretrains language models with query-query and query-answer relation-aware soft prompts, then transfers them via intent-specific prompts for few-shot intent detection, reporting up to 27% relative accuracy gains.

citing papers explorer

Showing 1 of 1 citing paper.

  • Few-Shot Query Intent Detection via Relation-Aware Prompt Learning cs.CL · 2025-09-06 · conditional · none · ref 18 · internal anchor

    SAID pretrains language models with query-query and query-answer relation-aware soft prompts, then transfers them via intent-specific prompts for few-shot intent detection, reporting up to 27% relative accuracy gains.