Pith. sign in

Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

We propose a general approach to quantitatively assessing the risk and vulnerability of artificial intelligence (AI) systems to biased decisions. The guiding principle of the proposed approach is that any AI algorithm must outperform a random guesser. This may appear trivial, but empirical results from a simplistic sequential decision-making scenario involving roulette games show that sophisticated AI-based approaches often underperform the random guesser by a significant margin. We highlight that modern recommender systems may exhibit a similar tendency to favor overly low-risk options. We argue that this "random guesser test" can serve as a useful tool for evaluating the utility of AI actions, and also points towards increasing exploration as a potential improvement to such systems.

citation-role summary

background 1

citation-polarity summary

fields

q-bio.NC 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

support 1

representative citing papers

AI Agent Behavioral Science

q-bio.NC · 2025-06-04 · conditional · novelty 4.0

AI agents should be studied as behavioral entities shaped by context and interaction, not only as trained models.

citing papers explorer

Showing 1 of 1 citing paper.

  • AI Agent Behavioral Science q-bio.NC · 2025-06-04 · conditional · none · ref 72 · internal anchor

    AI agents should be studied as behavioral entities shaped by context and interaction, not only as trained models.