Pith. sign in

REVIEW 1 cited by

DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2109.02492 v2 pith:KIS6Q6D5 submitted 2021-09-06 cs.CL

classification cs.CL
keywords dialoguelongmodeldialoglmpre-trainedsummarizationattentiondatasets
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Dialogue is an essential part of human communication and cooperation. Existing research mainly focuses on short dialogue scenarios in a one-on-one fashion. However, multi-person interactions in the real world, such as meetings or interviews, are frequently over a few thousand words. There is still a lack of corresponding research and powerful tools to understand and process such long dialogues. Therefore, in this work, we present a pre-training framework for long dialogue understanding and summarization. Considering the nature of long conversations, we propose a window-based denoising approach for generative pre-training. For a dialogue, it corrupts a window of text with dialogue-inspired noise, and guides the model to reconstruct this window based on the content of the remaining conversation. Furthermore, to process longer input, we augment the model with sparse attention which is combined with conventional attention in a hybrid manner. We conduct extensive experiments on five datasets of long dialogues, covering tasks of dialogue summarization, abstractive question answering and topic segmentation. Experimentally, we show that our pre-trained model DialogLM significantly surpasses the state-of-the-art models across datasets and tasks. Source code and all the pre-trained models are available on our GitHub repository (https://github.com/microsoft/DialogLM).

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Refining Text Generation for Realistic Conversational Recommendation via Direct Preference Optimization

    cs.IR 2025-08 conditional novelty 5.0 of 10

    DPO fine-tuning of the summary and recommendation writers improves conversational recommendation ranking on two Japanese datasets, but the evaluation shares the scorer that generated the training signal.

Pith tools