In a new persuasion game, humans outperformed the LLM o1-preview when an opponent's preferences had to be inferred, while o1-preview outperformed humans when those preferences were disclosed.
How much do you like attribute A?
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Do Large Language Models Have a Planning Theory of Mind? Evidence from MindGames: a Multi-Step Persuasion Task
In a new persuasion game, humans outperformed the LLM o1-preview when an opponent's preferences had to be inferred, while o1-preview outperformed humans when those preferences were disclosed.