Pith. sign in

REVIEW 7 cited by

Topical-Chat: Towards Knowledge-Grounded Open-Domain Conversations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.11995 v1 pith:A4A4SKZB submitted 2023-08-23 cs.CL cs.AI

classification cs.CLcs.AI
keywords conversationconversationsknowledgeknowledge-groundedopen-domaintopical-chatconversationaldatasets
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Building socialbots that can have deep, engaging open-domain conversations with humans is one of the grand challenges of artificial intelligence (AI). To this end, bots need to be able to leverage world knowledge spanning several domains effectively when conversing with humans who have their own world knowledge. Existing knowledge-grounded conversation datasets are primarily stylized with explicit roles for conversation partners. These datasets also do not explore depth or breadth of topical coverage with transitions in conversations. We introduce Topical-Chat, a knowledge-grounded human-human conversation dataset where the underlying knowledge spans 8 broad topics and conversation partners don't have explicitly defined roles, to help further research in open-domain conversational AI. We also train several state-of-the-art encoder-decoder conversational models on Topical-Chat and perform automated and human evaluation for benchmarking.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Beyond One-Size-Fits-All: Inversion Learning for Highly Effective NLG Evaluation Prompts

    cs.CL 2025-04 conditional novelty 6.0 of 10

    Inversion learning generates model-specific NLG evaluation prompts from a single human-annotated sample, and these prompts outperform hand-crafted and search-based prompts in correlation with human scores.

  2. Stop Overvaluing Multi-Agent Debate -- We Must Rethink Evaluation and Embrace Model Heterogeneity

    cs.CL 2025-02 conditional novelty 6.0 of 10

    Multi-agent debate mostly underperforms simple chain-of-thought baselines when tested broadly, while randomly mixing different models into the debate reliably improves performance.

  3. DRE: An Effective Dual-Refined Method for Integrating Small and Large Language Models in Open-Domain Dialogue Evaluation

    cs.CL 2025-06 conditional novelty 5.0 of 10

    A dual-refinement method (DRE) that uses a contrastively trained small language model to guide and rescale an LLM's dialogue quality scores achieves higher correlation with human ratings than LLM-only baselines on thr...

  4. Improving Linguistic Diversity of Large Language Models with Possibility Exploration Fine-Tuning

    cs.CL 2024-12 conditional novelty 5.0 of 10

    Possibility Exploration Fine-Tuning (PEFT) conditions LLMs on a random possibility number and trains with unlikelihood to generate diverse, controllable responses without added latency, as shown on dialogue and story tasks.

  5. Commonsense Generation and Evaluation for Dialogue Systems using Large Language Models

    cs.CL 2025-06 reject novelty 4.0 of 10

    GPT-3.5 and GPT-4 can generate and evaluate turn-level commonsense dialogue expansions guided by ATOMIC relation definitions, with GPT-4 achieving the highest reranking accuracy.

  6. Real-Time Textless Dialogue Generation

    cs.CL 2025-01 conditional novelty 4.0 of 10

    A streaming, textless dialogue model predicts turn-taking actions every 160 ms and generates speech units, improving naturalness over cascaded systems at the cost of lower semantic coherence.

  7. MAPS: Modeling Co-Existing Subjective Perspectives and Shared Meaning in Multi-Agent Cognitive Dialogue

    cs.CL 2026-05 reject novelty 3.0 of 10

    MAPS combines hand-coded domain weights, a GRU memory, and attention to let dialogue agents keep distinct subjective profiles while their hidden states are trained to move closer together.

Pith tools