Pith. sign in

REVIEW 1 cited by

Brain-Model Evaluations Need the NeuroAI Turing Test

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.16238 v1 pith:6LJLEDHP submitted 2025-02-22 q-bio.NC cs.AIcs.LGcs.NE

classification q-bio.NCcs.AIcs.LGcs.NE
keywords brainbehaviorneuroaitestturingintelligenceinternalmodel
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

What makes an artificial system a good model of intelligence? The classical test proposed by Alan Turing focuses on behavior, requiring that an artificial agent's behavior be indistinguishable from that of a human. While behavioral similarity provides a strong starting point, two systems with very different internal representations can produce the same outputs. Thus, in modeling biological intelligence, the field of NeuroAI often aims to go beyond behavioral similarity and achieve representational convergence between a model's activations and the measured activity of a biological system. This position paper argues that the standard definition of the Turing Test is incomplete for NeuroAI, and proposes a stronger framework called the ``NeuroAI Turing Test'', a benchmark that extends beyond behavior alone and \emph{additionally} requires models to produce internal neural representations that are empirically indistinguishable from those of a brain up to measured individual variability, i.e. the differences between a computational model and the brain is no more than the difference between one brain and another brain. While the brain is not necessarily the ceiling of intelligence, it remains the only universally agreed-upon example, making it a natural reference point for evaluating computational models. By proposing this framework, we aim to shift the discourse from loosely defined notions of brain inspiration to a systematic and testable standard centered on both behavior and internal representations, providing a clear benchmark for neuroscientific modeling and AI development.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Representation biases: will we achieve complete understanding by analyzing representations?

    q-bio.NC 2025-07 accept novelty 4.0 of 10

    Feature representation biases in trained models can distort PCA, regression, RSA, and model-brain comparisons, so representational analyses may not reveal all of a system's computations.

Pith tools