REVIEW 4 major objections 5 minor 35 references
Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality
T0 review · 4 major / 5 minor · reviewed 2026-08-07 · deepseek-v4-flash
Pith's one-line read People attribute beliefs that causally explain actions, not merely true ones.
desk verdict A clean, small-scale experiment showing causal relevance predicts belief attribution best, but the headline correlation is an in-sample fit that needs out-of-sample confirmation. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing machinery is a language-augmented Bayesian theory-of-mind model: a probabilistic generative model in which an agent with known goals updates beliefs about key locations and chooses actions by Boltzmann-rational planning. Candidate belief statements are represented as epistemic-logic formulas that evaluate to Boolean truth at each time step given the agent's belief state and environment. From this, the model computes three explanatory factors: accuracy as the posterior probability that the statement is true; informativity as the Kullback-Leibler divergence a statement would provide to a listener; and causal relevance as a normality-weighted product of causal necessity (the probability that the observed actions would not have occurred if the belief were intervened to be false at the last moment $t_c$ when beliefs changed) and causal sufficiency (the probability the actions would occur if the belief were intervened true), with atypical causes weighted more heavily.
What would settle it
A direct experiment could pit a highly accurate and informative but causally inert belief against a less accurate but causally necessary belief in scenarios where the two factor rankings diverge. If participants consistently rank the accurate-and-informative statement above the causally relevant one, the central claim would fail; conversely, the paper predicts they should prefer the belief whose falsity would have changed the agent's observed path.
Extended reading notes
Core claim
The central discovery is that causal relevance, not truth or informativeness, best explains which belief statements people choose to attribute to another agent. The model computes causal relevance from a generative model of belief-driven action by evaluating what would have happened under hypothetical interventions on the belief: a statement scores highly when setting the belief false would have prevented the observed actions and setting it true would have produced them, with unusual causes weighted more heavily. This causal score predicted average human rankings of three candidate belief statements per scenario at $r=0.81$, beating every other single factor and matching the fit of the full three-factor model ($r=0.82$). The paper also shows that accuracy and informativity together do reasonably well ($r=0.68$) and correlate with causal relevance, but qualitative cases reveal divergences where humans side with causal relevance over mere accuracy plus informativity.
Load-bearing premise
The central claim rests on the assumption that people evaluate a belief statement by simulating a hypothetical intervention at the precise moment the agent's beliefs last changed and judging only the subsequent actions, with the agent's action choices treated as Boltzmann-rational noise.
Editorial extensions
If this is right
- If causal relevance is the dominant factor, then models of belief attribution should prioritize interventions on mental states over posterior beliefs, which suggests new prediction targets for theory-of-mind models.
- Accuracy and informativity alone are insufficient predictors of human belief attribution, so communicative accounts of explanation need to be supplemented with causal structure rather than replaced by it.
- The fact that a full three-factor model only marginally improves on causal relevance alone indicates that causal relevance already captures much of what accuracy and informativity contribute in these scenarios.
- The model's formalism extends naturally to richer natural-language belief statements, including knowledge claims and compositional beliefs, because it computes factors from a grounded epistemic semantics.
- The findings open a path toward unifying causal and communicative accounts of explanation: a causally relevant belief is often the most useful thing to tell a listener, even without an explicit communicative context.
Reading between the lines
- An unstated implication is that people may attribute beliefs the way scientists select causes: as the minimal intervention that would change the observed outcome, which suggests belief attribution could track counterfactual dependence over whole trajectories, not just the last belief change.
- If tested across different mental states, the same causal-relevance logic might predict which desires, intentions, or emotions people cite as explanations, since those states also drive action in a generative model.
- A testable extension is to vary the time at which beliefs change: the model's intervention point is fixed at $t_c$, so scenarios with multiple belief changes or delayed belief updates could distinguish this measure from a full counterfactual simulation account.
- The residual error cases point toward bounded rationality and ambiguous statement readings as the next limiting factors, so incorporating suboptimal sub-goal selection and multiple readings of disjunctive statements may raise the ceiling beyond $r=0.81$.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This paper asks which belief statements people prefer to attribute to other agents, and proposes that people select beliefs that function as good mental explanations. Building on the LaBToM framework, the authors define three explanatory factors: accuracy (posterior probability of the belief statement given actions and observations), informativity (KL information gain of the statement relative to a listener), and causal relevance (a combination of causal necessity, causal sufficiency, and normality). The model is evaluated with an online experiment in which 41 participants watch gridworld agents collect keys and rank three belief statements for each of 18 scenarios. The α weights in the scoring function are fitted by grid search to maximize Pearson correlation with averaged human rankings. The paper reports that causal relevance is the single best predictor (r = 0.81), outperforming accuracy plus informativity (r = 0.68), accuracy alone (r = 0.36), and informativity alone (r = 0.43). The central claim is that people gravitate toward causally relevant beliefs when attributing mental states.
Significance. The paper addresses a genuine gap by connecting belief attribution with computational accounts of explanation. A notable strength is that the explanatory factor scores are computed from a generative theory-of-mind model that is independent of the human ranking data; only the α combination weights are fitted. If the causal-relevance advantage survives proper model comparison, the result would be an important contribution to both theory-of-mind research and computational pragmatics. However, the current evidence for the headline model comparison is not yet conclusive: the α weights are fitted to the same data used to compute the reported correlations, no complexity-penalized or out-of-sample comparison is reported, and the scenarios are acknowledged by the authors to only weakly disambiguate the competing accounts. The framework and stimuli are nevertheless a promising foundation for stronger tests.
major comments (4)
- [Model Fitting and Table 1] The α coefficients are fit by maximizing Pearson's correlation on the same averaged human rankings used to compute the reported r values. This makes the correlations in Table 1 in-sample fits, and the reported 95% confidence intervals do not account for the parameter fitting. The central claim that Causal (r = 0.81) outperforms Acc+Info (r = 0.68) rests on this comparison, so the authors should report leave-one-scenario-out or leave-one-statement-out cross-validated correlations, or a penalized model-comparison statistic such as AIC, BIC, or WAIC, for each factor combination.
- [Table 1] Several fitted coefficients sit at the upper boundary of the grid search range [0,10], most notably αCSucc = 10.00 for the Causal model and αAcc = 10.00 for Acc+Info and other combinations. Boundary solutions suggest the grid may truncate the optimum and that the data do not strongly constrain the parameters. A sensitivity analysis over a wider range, or a continuous optimizer with reported standard errors, is needed before the parameter values can be interpreted as meaningful and before the Causal advantage can be attributed to the model rather than to the chosen grid.
- [Results and Discussion] The authors themselves state in the Discussion that the stimuli 'do not strongly disambiguate' causal relevance from the combined accuracy and informativity account, and the Results report a high correlation (r = 0.81) between the Acc+Info and Causal models. With only 18 scenarios and 54 statement-level data points, the 0.13 correlation advantage of Causal over Acc+Info is fragile. The paper should report per-scenario correlations or the proportion of scenarios in which Causal ranks the human-preferred statement more accurately than Acc+Info, along with a leave-one-scenario-out analysis, to show that the aggregate advantage is not driven by a few scenarios.
- [Causal Relevance equations] The causal relevance measure intervenes at the most recent time step tc at which the agent's beliefs changed and evaluates only subsequent actions (CNecc and CSuff equations). This is a strong structural assumption about how people perform causal reasoning about beliefs. Because the paper's central claim depends on this specific formalization, the authors should compare at least one alternative (e.g., intervening at the current time, or counterfactual reasoning over the entire trajectory) or provide quantitative evidence that the model rankings are robust to this choice. The Discussion mentions that hypothetical and counterfactual interventions 'do not come apart significantly' in this experiment, but no such comparison is shown.
minor comments (5)
- [Experiment Design] The manuscript states that there are 18 scenarios (6 maps × 3 key allocations), but the Experiment Design paragraph says 'Each participant completed all 21 scenarios in a randomized order.' This inconsistency should be corrected.
- [Figure 3] The caption contains a typo: 'enivonment' should be 'environment'; in addition, the text at the end of the Qualitative Analysis section refers to 'Acc+sfInfo', which should be 'Acc+Info'.
- [Results section] There are minor typos in the text: 'lister' should be 'listener' and 'casual intuitions' in the Discussion should be 'causal intuitions'.
- [Probabilistic Attribution of Belief Statements] The Score equation is written as a linear combination of logarithms 'except for Info, which is in log-space.' This wording is confusing because Info is already a KL divergence; please clarify whether the model uses log(Info) in the sum and why this choice was made.
- [Model Fitting] The computation of the 95% confidence intervals in Table 1 is not described. Please state whether these are bootstrap intervals, and if so, whether the bootstrap resamples participants, scenarios, or statements, and whether the α parameters were refit during resampling.
Circularity Check
The reported correlations are optimized in-sample fits, so the Causal advantage is a goodness-of-fit claim rather than an independent prediction.
-
fitted input called prediction
[Model Fitting / Results (Correlation Analysis); Table 1]
"We then fit the coefficients α f for each explanatory factor by maximizing the Pearson’s correlation between the average predicted ranks and the average human-provided ranks for each statement. ... Our results indicate that the causal relevance factor (Causal) was the single factor that best explained human rankings over belief attributions, with a correlation of r = 0.81."
The model parameters α_f are fitted by maximizing exactly the Pearson correlation that is then reported as evidence for the central claim. With no held-out data, cross-validation, or model-comparison statistic, the r values in Table 1 are optimized in-sample goodness-of-fit values, not predictions. The comparison of Causal (r=0.81) against Acc+Info (r=0.68) is therefore a comparison of fitted correlations on the same 54 averaged statement ranks used as the fitting target. This is reinforced by α_CSucc=10.00 sitting at the upper boundary of the grid-search range.
full rationale
The explanatory factors — accuracy, informativity, and causal relevance — are computed from the generative model via probabilistic, information-theoretic, and interventionist quantities that do not use the human ranking data. That part of the derivation is self-contained. The central circularity concern is that the α weights in the probabilistic attribution model are fitted to the very same averaged human rankings that are later correlated with model rankings, and no cross-validation or held-out evaluation is reported. Thus the reported r values, including the 0.81 vs 0.68 advantage of Causal over Acc+Info, are optimized in-sample fits rather than predictive validations. This does not reduce the central claim to a definitional tautology, because the factor scores themselves are not derived from the human data and the fitted correlations are not guaranteed to order the models in the observed way; but the statistical evidence for the headline claim is weaker than presented. The paper's self-citations to LaBToM and ELoT are used as modeling infrastructure, not as a substitute for evidence, so they are not load-bearing circularity.
Assumptions & free parameters
free parameters (6)
- alpha_Acc =
10.00 (in All Factors model)
- alpha_Info =
0.086 (in All Factors model)
- alpha_CNecc =
1.216 (in All Factors model)
- alpha_CSucc =
2.432 (in All Factors model)
- 50-50 prior for accuracy =
0.5 (chosen, not fitted here)
- K = 3 initial belief weights =
3 (chosen)
assumptions (5)
- domain assumption Actions are Boltzmann-rational with respect to the agent's beliefs and goal.
- domain assumption Belief update is deterministic and based on consistency with observations.
- domain assumption Uniform prior over initial environment states and over K=3 weight distributions.
- domain assumption Each belief statement has a single valid reading.
- domain assumption Intervention at the last belief-change time step captures causality.
Cite this review
Pith. "Pith review of Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality." pith.science (2026). https://pith.science/paper/RTQA2Z3J
@misc{pith2026250519376,
author = {Pith},
title = {Pith review of: Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality},
year = {2026},
howpublished = {\url{https://pith.science/paper/RTQA2Z3J}},
note = {Machine review of arXiv:2505.19376}
}
read the original abstract
A key feature of human theory-of-mind is the ability to attribute beliefs to other agents as mentalistic explanations for their behavior. But given the wide variety of beliefs that agents may hold about the world and the rich language we can use to express them, which specific beliefs are people inclined to attribute to others? In this paper, we investigate the hypothesis that people prefer to attribute beliefs that are good explanations for the behavior they observe. We develop a computational model that quantifies the explanatory strength of a (natural language) statement about an agent's beliefs via three factors: accuracy, informativity, and causal relevance to actions, each of which can be computed from a probabilistic generative model of belief-driven behavior. Using this model, we study the role of each factor in how people selectively attribute beliefs to other agents. We investigate this via an experiment where participants watch an agent collect keys hidden in boxes in order to reach a goal, then rank a set of statements describing the agent's beliefs about the boxes' contents. We find that accuracy and informativity perform reasonably well at predicting these rankings when combined, but that causal relevance is the single factor that best explains participants' responses.
Figures
Reference graph
Works this paper leans on
-
[1]
Modeling the Mistakes of Boundedly Rational Agents Within a Bayesian Theory of Mind
alanqary2021modeling APACrefauthors Alanqary, A. , Lin, G Z. , Le, J. , Zhi-Xuan, T. , Mansinghka, V K. \ Tenenbaum, J B. APACrefauthors \ 2021 . Modeling the mistakes of boundedly rational agents within a Bayesian theory of mind Modeling the mistakes of boundedly rational agents within a bayesian theory of mind . arXiv preprint arXiv:2106.13249
work page Pith review arXiv 2021
-
[2]
alicke2015causal APACrefauthors Alicke, M D. , Mandel, D R. , Hilton, D J. , Gerstenberg, T. \ Lagnado, D A. APACrefauthors \ 2015 . Causal conceptions in social explanation and moral evaluation: A historical tour Causal conceptions in social explanation and moral evaluation: A historical tour . Perspectives on Psychological Science 10 6 790--812
work page 2015
-
[3]
baker2017rational APACrefauthors Baker, C L. , Jara-Ettinger, J. , Saxe, R. \ Tenenbaum, J B. APACrefauthors \ 2017 . Rational quantitative attribution of beliefs, desires and percepts in human mentalizing Rational quantitative attribution of beliefs, desires and percepts in human mentalizing . Nature Human Behaviour 1 4 1--10
work page 2017
-
[4]
cedegao2021does APACrefauthors Cedegao, Z. , Ham, H. \ Holliday, W H. APACrefauthors \ 2021 . Does Amy know Ben knows you know your cards? A computational model of higher-order epistemic reasoning Does amy know ben knows you know your cards? a computational model of higher-order epistemic reasoning . Proceedings of the Annual Meeting of the Cognitive Scie...
work page 2021
-
[5]
chandra2024cooperative APACrefauthors Chandra, K. , Chen, T. , Li, T M. , Ragan-Kelley, J. \ Tenenbaum, J. APACrefauthors \ 2024 . Cooperative Explanation as Rational Communication Cooperative explanation as rational communication . Proceedings of the Annual Meeting of the Cognitive Science Society Proceedings of the annual meeting of the cognitive scienc...
work page 2024
-
[6]
chen2024intervening APACrefauthors Chen, T. , Houlihan, S D. , Chandra, K. , Tenenbaum, J. \ Saxe, R. APACrefauthors \ 2024 . Intervening on Emotions by Planning Over a Theory of Mind Intervening on emotions by planning over a theory of mind . Proceedings of the Annual Meeting of the Cognitive Science Society Proceedings of the annual meeting of the cogni...
work page 2024
-
[7]
gerstenberg2015whether APACrefauthors Gerstenberg, T. , Goodman, N D. , Lagnado, D A. \ Tenenbaum, J B. APACrefauthors \ 2015 . How, whether, why: Causal judgments as counterfactual contrasts. How, whether, why: Causal judgments as counterfactual contrasts. CogSci. Cogsci
work page 2015
-
[8]
gerstenberg2021counterfactual APACrefauthors Gerstenberg, T. , Goodman, N D. , Lagnado, D A. \ Tenenbaum, J B. APACrefauthors \ 2021 . A counterfactual simulation model of causal judgments for physical events. A counterfactual simulation model of causal judgments for physical events. Psychological review 128 5 936
work page 2021
Show all 35 references
-
[9]
\ Icard, T
gerstenberg2020expectations APACrefauthors Gerstenberg, T. \ Icard, T. APACrefauthors \ 2020 . Expectations affect physical causation judgments. Expectations affect physical causation judgments. Journal of Experimental Psychology: General 149 3 599
2020
-
[10]
, Saxe, R
ho2022planning APACrefauthors Ho, M K. , Saxe, R. \ Cushman, F. APACrefauthors \ 2022 . Planning with theory of mind Planning with theory of mind . Trends in Cognitive Sciences 26 11 959--971
2022
-
[11]
\ Knobe, J
icard2016causality APACrefauthors Icard, T F. \ Knobe, J. APACrefauthors \ 2016 . Causality, normality, and sampling propensity Causality, normality, and sampling propensity . Proceedings of the Annual Meeting of the Cognitive Science Society Proceedings of the annual meeting ...
2016
-
[12]
, Kominsky, J F
icard2017normality APACrefauthors Icard, T F. , Kominsky, J F. \ Knobe, J. APACrefauthors \ 2017 . Normality and actual causal strength Normality and actual causal strength . Cognition 161 80--93
2017
-
[13]
, Harding, J
kirfel2024explain APACrefauthors Kirfel, L. , Harding, J. , Shin, J Y. , Xin, C. , Icard, T. \ Gerstenberg, T. APACrefauthors \ 2024 . Do as I explain: Explanations communicate optimal interventions Do as i explain: Explanations communicate optimal interventions . Proceedings ...
2024
-
[14]
, Icard, T
kirfel2022inference APACrefauthors Kirfel, L. , Icard, T. \ Gerstenberg, T. APACrefauthors \ 2022 . Inference from explanation. Inference from explanation. Journal of Experimental Psychology: General 151 7 1481
2022
-
[15]
, Gerstenberg, T
lagnado2013causal APACrefauthors Lagnado, D A. , Gerstenberg, T. \ Zultan, R. APACrefauthors \ 2013 . Causal responsibility and counterfactuals Causal responsibility and counterfactuals . Cognitive science 37 6 1036--1073
2013
-
[16]
APACrefauthors \ 2006
lombrozo2006structure APACrefauthors Lombrozo, T. APACrefauthors \ 2006 . The structure and function of explanations The structure and function of explanations . Trends in cognitive sciences 10 10 464--470
2006
-
[17]
APACrefauthors \ 2007
lombrozo2007simplicity APACrefauthors Lombrozo, T. APACrefauthors \ 2007 . Simplicity and probability in causal explanation Simplicity and probability in causal explanation . Cognitive psychology 55 3 232--257
2007
-
[18]
, Shprints, R
machino2024listener APACrefauthors Machino, Y. , Shprints, R. , Siegel, M. , Wong, L. \ Tenenbaum, J. APACrefauthors \ 2024 . Listener Knowledge Structures Commonsense Explanation Listener knowledge structures commonsense explanation . Proceedings of the Annual Meeting of the ...
2024
-
[19]
\ Baillargeon, R
onishi200515 APACrefauthors Onishi, K H. \ Baillargeon, R. APACrefauthors \ 2005 . Do 15-month-old infants understand false beliefs? Do 15-month-old infants understand false beliefs? science 308 5719 255--258
2005
-
[20]
APACrefauthors \ 2009
pearl2009causality APACrefauthors Pearl, J. APACrefauthors \ 2009 . Causality Causality . Cambridge university press
2009
-
[21]
\ Lucas, C
quillien2022logic APACrefauthors Quillien, T. \ Lucas, C. APACrefauthors \ 2022 . The logic of guesses: how people communicate probabilistic information The logic of guesses: how people communicate probabilistic information . Proceedings of the Annual Meeting of the Cognitive ...
2022
-
[22]
\ Lucas, C G
quillien2023counterfactuals APACrefauthors Quillien, T. \ Lucas, C G. APACrefauthors \ 2023 . Counterfactuals and the logic of causal selection. Counterfactuals and the logic of causal selection. Psychological Review
2023
-
[23]
\ Lapata, M
saparina2024ambrosia APACrefauthors Saparina, I. \ Lapata, M. APACrefauthors \ 2024 . Ambrosia: A benchmark for parsing ambiguous questions into database queries Ambrosia: A benchmark for parsing ambiguous questions into database queries . Advances in Neural Information Proces...
2024
-
[24]
, Tessler, M H
scontras2021practical APACrefauthors Scontras, G. , Tessler, M H. \ Franke, M. APACrefauthors \ 2021 . A practical introduction to the Rational Speech Act modeling framework A practical introduction to the rational speech act modeling framework . arXiv preprint arXiv:2105.09867
2021 arXiv
-
[25]
, Klassen, T Q
shvo2020epistemic APACrefauthors Shvo, M. , Klassen, T Q. , Sohrabi, S. \ McIlraith, S A. APACrefauthors \ 2020 . Epistemic plan recognition Epistemic plan recognition . Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems Proceedings of...
2020
-
[26]
\ Pearl, J
tian2000probabilities APACrefauthors Tian, J. \ Pearl, J. APACrefauthors \ 2000 . Probabilities of causation: Bounds and identification Probabilities of causation: Bounds and identification . Annals of Mathematics and Artificial Intelligence 28 1 287--313
2000
-
[27]
\ Perner, J
wimmer1983beliefs APACrefauthors Wimmer, H. \ Perner, J. APACrefauthors \ 1983 . Beliefs about beliefs: Representation and constraining function of wrong beliefs in young children's understanding of deception Beliefs about beliefs: Representation and constraining function of w...
1983
-
[28]
\ DeDeo, S
wojtowicz2020probability APACrefauthors Wojtowicz, Z. \ DeDeo, S. APACrefauthors \ 2020 . From probability to consilience: How explanatory values implement B ayesian reasoning From probability to consilience: How explanatory values implement B ayesian reasoning . Trends in Cog...
2020
-
[29]
, Schulz, L
wu2024change APACrefauthors Wu, S. , Schulz, L. \ Saxe, R. APACrefauthors \ 2024 . How to Change a Mind: Adults and Children Use the Causal Structure of Theory of Mind to Intervene on Others’ Behaviors How to change a mind: Adults and children use the causal structure of theor...
2024
-
[30]
, Sridhar, S
wu2023computational APACrefauthors Wu, S A. , Sridhar, S. \ Gerstenberg, T. APACrefauthors \ 2023 . A computational model of responsibility judgments from counterfactual simulations and intention inferences A computational model of responsibility judgments from counterfactual ...
2023
-
[31]
, Collins, K M
ying2023neuro APACrefauthors Ying, L. , Collins, K M. , Wei, M. , Zhang, C E. , Zhi-Xuan, T. , Weller, A. Wong, L. APACrefauthors \ 2023 . The N euro- S ymbolic I nverse P lanning E ngine ( NIPE ): Modeling Probabilistic Social Inferences from Linguistic Inputs The N euro- S y...
2023 arXiv
-
[32]
, Collins, K M
ying2025benchmarking APACrefauthors Ying, L. , Collins, K M. , Wong, L. , Sucholutsky, I. , Liu, R. , Weller, A. Tenenbaum, J B. APACrefauthors \ 2025 . On benchmarking human-like intelligence in machines On benchmarking human-like intelligence in machines . arXiv preprint arX...
2025 arXiv
-
[33]
, Zhi-Xuan, T
ying2024grounding APACrefauthors Ying, L. , Zhi-Xuan, T. , Wong, L. , Mansinghka, V. \ Tenenbaum, J. APACrefauthors \ 2024 . Grounding Language about B elief in a B ayesian Theory-of-Mind Grounding language about B elief in a B ayesian theory-of-mind . Proceedings of the Annua...
2024
-
[34]
, Zhi-Xuan, T
ying2025understanding APACrefauthors Ying, L. , Zhi-Xuan, T. , Wong, L. , Mansinghka, V. \ Tenenbaum, J B. APACrefauthors \ 2025 . Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind Understanding epistemic language with a language-augmented baye...
2025
-
[35]
, Mann, J
zhi2020online APACrefauthors Zhi-Xuan, T. , Mann, J. , Silver, T. , Tenenbaum, J. \ Mansinghka, V. APACrefauthors \ 2020 . Online B ayesian Goal Inference for Boundedly Rational Planning Agents Online B ayesian goal inference for boundedly rational planning agents . Advances i...
2020
Reviewed August 7, 2026 · model on record in the stance chip above.
Discussion (0). Sign in to comment.