Pith. sign in

Solving the Baby Intuitions Benchmark with a Hierarchically Bayesian Theory of Mind

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

To facilitate the development of new models to bridge the gap between machine and human social intelligence, the recently proposed Baby Intuitions Benchmark (arXiv:2102.11938) provides a suite of tasks designed to evaluate commonsense reasoning about agents' goals and actions that even young infants exhibit. Here we present a principled Bayesian solution to this benchmark, based on a hierarchically Bayesian Theory of Mind (HBToM). By including hierarchical priors on agent goals and dispositions, inference over our HBToM model enables few-shot learning of the efficiency and preferences of an agent, which can then be used in commonsense plausibility judgements about subsequent agent behavior. This approach achieves near-perfect accuracy on most benchmark tasks, outperforming deep learning and imitation learning baselines while producing interpretable human-like inferences, demonstrating the advantages of structured Bayesian models of human social cognition.

citation-role summary

background 1

citation-polarity summary

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Machine Theory of Mind and the Structure of Human Values

cs.AI · 2025-05-24 · conditional · novelty 5.0

Human values are claimed to have a rational instrumental structure that lets AI infer unseen values from known ones, framing this as the 'value generalization problem' in AI safety.

citing papers explorer

Showing 1 of 1 citing paper.

  • Machine Theory of Mind and the Structure of Human Values cs.AI · 2025-05-24 · conditional · none · ref 44 · internal anchor

    Human values are claimed to have a rational instrumental structure that lets AI infer unseen values from known ones, framing this as the 'value generalization problem' in AI safety.