Pith. sign in

REVIEW

Strings from the Library of Babel: Random Sampling as a Strong Baseline for Prompt Optimisation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.09569 v2 pith:TTNJYAXT submitted 2023-11-16 cs.CL cs.AI

classification cs.CLcs.AI
keywords languagepromptseparatorsmodelsoptimisationrandomstrongaverage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent prompt optimisation approaches use the generative nature of language models to produce prompts -- even rivaling the performance of human-curated prompts. In this paper, we demonstrate that randomly sampling tokens from the model vocabulary as ``separators'' can be as effective as language models for prompt-style text classification. Our experiments show that random separators are competitive baselines, having less than a 1% difference compared to previous self-optimisation methods and showing a 12% average relative improvement over strong human baselines across nine text classification tasks and eight language models. We further analyse this phenomenon in detail using three different random generation strategies, establishing that the language space is rich with potentially good separators, with a greater than 40% average chance that a randomly drawn separator performs better than human-curated separators. These observations challenge the common assumption that an effective prompt should be human readable or task relevant and establish a strong baseline for prompt optimisation research.

Discussion (0). Continue with ORCID to comment.

Pith tools