Pith. sign in

REVIEW 1 cited by

Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.08141 v2 pith:F2QIFAIL submitted 2022-09-16 cs.CL cs.AIcs.LG

Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models

classification cs.CL cs.AIcs.LG
keywords modelslanguagepromptsunderstandingchain-of-thoughtmetaphorprobabilisticthey
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Probabilistic models of language understanding are valuable tools for investigating human language use. However, they need to be hand-designed for a particular domain. In contrast, large language models (LLMs) are trained on text that spans a wide array of domains, but they lack the structure and interpretability of probabilistic models. In this paper, we use chain-of-thought prompts to introduce structures from probabilistic models into LLMs. We explore this approach in the case of metaphor understanding. Our chain-of-thought prompts lead language models to infer latent variables and reason about their relationships in order to choose appropriate paraphrases for metaphors. The latent variables and relationships chosen are informed by theories of metaphor understanding from cognitive psychology. We apply these prompts to the two largest versions of GPT-3 and show that they can improve performance in a paraphrase selection task.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?

    cs.CL 2026-07 conditional novelty 6.5

    Δacc between low-frequency and novel xiehouyu is ~23.6% for Chinese frontier LLMs vs ~5.1% for English-centric models and ~2.9% for humans, while LLM-created xiehouyu rate below human creations.