Pith. sign in

REVIEW 1 cited by

Violation of Expectation via Metacognitive Prompting Reduces Theory of Mind Prediction Error in Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.06983 v1 pith:BQUKMK32 submitted 2023-10-10 cs.CL cs.LG

classification cs.CLcs.LG
keywords expectationhumanlanguagelargellmsmetacognitivemindmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent research shows that Large Language Models (LLMs) exhibit a compelling level of proficiency in Theory of Mind (ToM) tasks. This ability to impute unobservable mental states to others is vital to human social cognition and may prove equally important in principal-agent relations between individual humans and Artificial Intelligences (AIs). In this paper, we explore how a mechanism studied in developmental psychology known as Violation of Expectation (VoE) can be implemented to reduce errors in LLM prediction about users by leveraging emergent ToM affordances. And we introduce a \textit{metacognitive prompting} framework to apply VoE in the context of an AI tutor. By storing and retrieving facts derived in cases where LLM expectation about the user was violated, we find that LLMs are able to learn about users in ways that echo theories of human learning. Finally, we discuss latent hazards and augmentative opportunities associated with modeling user psychology and propose ways to mitigate risk along with possible directions for future inquiry.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Amico: An Event-Driven Modular Framework for Persistent and Embedded Autonomy

    cs.AI 2025-07 reject novelty 4.0 of 10

    A new event-driven Rust agent framework for embedded and WebAssembly deployment reports better WebShop reward (0.61 vs 0.47) than an LLM baseline, but without error bars or released code.

Pith tools