**arXiv ID:** 2502.16238 **Authors:** Jenelle Feather, Meenakshi Khosla, N. Apurva Ratan Murty, Aran Nayebi **Published:** 2025-02-22T14:16:28Z **Abstract:** What makes an artificial system a good model of intelligence? The classical test proposed by Alan Turing focuses on behavior, requiring that an artificial agent's behavior be indistinguishable from that of a human. While behavioral similarity provides a strong starting point, two systems with very different internal representations can p...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill brainmodel-evaluations-need-the-neuroai-turing-test --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Brainmodel Evaluations Need The Neuroai Turing Test?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-brainmodel-evaluations-need-the-neuroai-turing-tes)More formats (shields.io, HTML) on the badges page.
# Brain-Model Evaluations Need the NeuroAI Turing Test
**arXiv ID:** 2502.16238
**Authors:** Jenelle Feather, Meenakshi Khosla, N. Apurva Ratan Murty, Aran Nayebi
**Published:** 2025-02-22T14:16:28Z
**Abstract:**
What makes an artificial system a good model of intelligence? The classical test proposed by Alan Turing focuses on behavior, requiring that an artificial agent's behavior be indistinguishable from that of a human. While behavioral similarity provides a strong starting point, two systems with very different internal representations can produce the same outputs. Thus, in modeling biological intelligence, the field of NeuroAI often aims to go beyond behavioral similarity and achieve representational convergence between a model's activations and the measured activity of a biological system. This position paper argues that the standard definition of the Turing Test is incomplete for NeuroAI, and proposes a stronger framework called the ``NeuroAI Turing Test'', a benchmark that extends beyond behavior alone and \emph{additionally} requires models to produce internal neural representations that are empirically indistinguishable from those of a brain up to measured individual variability, i.e. the differences between a computational model and the brain is no more than the difference between one brain and another brain. While the brain is not necessarily the ceiling of intelligence, it remains the only universally agreed-upon example, making it a natural reference point for evaluating computational models. By proposing this framework, we aim to shift the discourse from loosely defined notions of brain inspiration to a systematic and testable standard centered on both behavior and internal representations, providing a clear benchmark for neuroscientific modeling and AI development.
## Skill Description
This skill is generated from the arXiv paper: Brain-Model Evaluations Need the NeuroAI Turing Test (2502.16238).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2502.16238](http://arxiv.org/abs/2502.16238v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!