paper-with-me

홈 › Papers

Beyond Behavior: Why AI Evaluation Needs a Cognitive Revolution

2026-04-07 · Amir Konigsberg arxiv

In 1950, Alan Turing proposed replacing the question "Can machines think?" with a behavioral test: if a machine's outputs are indistinguishable from those of a thinking being, the question of whether it truly thinks can be set aside. This paper argues that Turing's move was not only a pragmatic simplification but also an epistemological commitment, a decision about what kind of evidence counts as relevant to intelligence attribution, and that this commitment has quietly constrained AI research for seven decades. We trace how Turing's behavioral epistemology became embedded in the field's evaluative infrastructure, rendering unaskable a class of questions about process, mechanism, and internal organization that cognitive psychology, neuroscience, and related disciplines learned to ask. We draw a structural parallel to the behaviorist-to-cognitivist transition in psychology: just as psychology's commitment to studying only observable behavior prevented it from asking productive questions about internal mental processes until that commitment was abandoned, AI's commitment to behavioral evaluation prevents it from distinguishing between systems that achieve identical outputs through fundamentally different computational processes, a distinction on which intelligence attribution depends. We argue that the field requires an epistemological transition comparable to the cognitive revolution: not an abandonment of behavioral evidence, but a recognition that behavioral evidence alone is insufficient for the construct claims the field wishes to make. We articulate what a post-behaviorist epistemology for AI would involve and identify the specific questions it would make askable that the field currently has no way to ask.

📄 PDF Abstract BibTeX arXiv:2604.05631

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bridging Values and Behavior: A Hierarchical Framework for Proactive Embodied Agents

2026-04-30 · Chunhui Zhang, Yuxuan Wang, Aoyang Qin, Yi-Long Lu 외 arxiv

Current embodied agents are often limited to passive instruction-following or reactive need-satisfaction, lacking a stable, high-order value framework essential for long-term, self-directed behavior and resolving motivat…

Bias Beware: The Impact of Cognitive Biases on LLM-Driven Product Recommendations

2025-02-03 · Giorgos Filandrianos, Angeliki Dimitriou, Maria Lymperaiou, Konstantinos Thomas 외

The advent of Large Language Models (LLMs) has revolutionized product recommenders, yet their susceptibility to adversarial manipulation poses critical challenges, particularly in real-world commercial applications. Our …

Product Recommendation

Developing the PsyCogMetrics AI Lab to Evaluate Large Language Models and Advance Cognitive Science -- A Three-Cycle Action Design Science Study

2026-03-13 · Zhiye Jin, Yibai Li, K. D. Joshi, Xuefei 외 arxiv

This study presents the development of the PsyCogMetrics AI Lab (psycogmetrics.ai), an integrated, cloud-based platform that operationalizes psychometric and cognitive-science methodologies for Large Language Model (LLM)…

Closer to Language than Steam: AI as the Cognitive Engine of a New Productivity Revolution

2025-06-12 · Xinmin Fang, Lingfeng Tao, Zhengxiong Li

Artificial Intelligence (AI) is reframed as a cognitive engine driving a novel productivity revolution distinct from the Industrial Revolution's physical thrust. This paper develops a theoretical framing of AI as a cogni…

Can Large Language Models Simulate Human Cognition Beyond Behavioral Imitation?

2026-03-29 · Yuxuan Gu, Lunjun Liu, Xiaocheng Feng, Kun Zhu 외 arxiv

An essential problem in artificial intelligence is whether LLMs can simulate human cognition or merely imitate surface-level behaviors, while existing datasets suffer from either synthetic reasoning traces or population-…