Pith. sign in

REVIEW 1 cited by

Frugal Prompting for Dialog Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.14919 v2 pith:EX2OYYZ7 submitted 2023-05-24 cs.CL

classification cs.CL
keywords dialogllmsmodelsunderstandingvariousabilitiesbetterbuilding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The use of large language models (LLMs) in natural language processing (NLP) tasks is rapidly increasing, leading to changes in how researchers approach problems in the field. To fully utilize these models' abilities, a better understanding of their behavior for different input protocols is required. With LLMs, users can directly interact with the models through a text-based interface to define and solve various tasks. Hence, understanding the conversational abilities of these LLMs, which may not have been specifically trained for dialog modeling, is also important. This study examines different approaches for building dialog systems using LLMs by considering various aspects of the prompt. As part of prompt tuning, we experiment with various ways of providing instructions, exemplars, current query and additional context. The research also analyzes the representations of dialog history that have the optimal usable-information density. Based on the findings, the paper suggests more compact ways of providing dialog history information while ensuring good performance and reducing model's inference-API costs. The research contributes to a better understanding of how LLMs can be effectively used for building interactive systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MCST-Mamba: Multivariate Mamba-Based Model for Traffic Prediction

    cs.LG 2025-07 reject novelty 4.0 of 10

    MCST-Mamba combines STAEformer-style adaptive embeddings with two Mamba blocks to jointly predict speed, flow, and occupancy, but its claimed state-of-the-art results rest on comparing aggregated multi-channel errors ...

Pith tools