Pith. sign in

REVIEW 1 cited by

EM Pre-training for Multi-party Dialogue Response Generation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.12412 v1 pith:KKSV45F7 submitted 2023-05-21 cs.CL

classification cs.CL
keywords responsedialoguegenerationmulti-partydialoguestwo-partyaddresseebeen
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Dialogue response generation requires an agent to generate a response according to the current dialogue history, in terms of which two-party dialogues have been well studied, but leaving a great gap for multi-party dialogues at the same time. Different from two-party dialogues where each response is a direct reply to its previous utterance, the addressee of a response utterance should be specified before it is generated in the multi-party scenario. Thanks to the huge amount of two-party conversational data, various pre-trained language models for two-party dialogue response generation have been proposed. However, due to the lack of annotated addressee labels in multi-party dialogue datasets, it is hard to use them to pre-train a response generation model for multi-party dialogues. To tackle this obstacle, we propose an Expectation-Maximization (EM) approach that iteratively performs the expectation steps to generate addressee labels, and the maximization steps to optimize a response generation model. Theoretical analyses and extensive experiments have justified the feasibility and effectiveness of our proposed method.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. An LLM Benchmark for Addressee Recognition in Multi-modal Multi-party Dialogue

    cs.CL 2025-01 conditional novelty 6.0 of 10

    On a new 387-turn Japanese triadic dialogue subset, GPT-4o's addressee recognition accuracy (80.9%) matches the always-O baseline (80.6%), and its next-speaker prediction (46.0%) is below the 50% baseline.

Pith tools