paper-with-me

홈 › Papers

Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models

2026-02-25 · Christian Nickel, Laura Schrewe, Florian Mai, Lucie Flek arxiv

Theory of Mind (ToM) refers to an agent's ability to model the internal states of others. Contributing to the debate whether large language models (LLMs) exhibit genuine ToM capabilities, our study investigates their ToM robustness using perturbations on false-belief tasks and examines the potential of Chain-of-Thought prompting (CoT) to enhance performance and explain the LLM's decision. We introduce a handcrafted, richly annotated ToM dataset, including classic and perturbed false belief tasks, the corresponding spaces of valid reasoning chains for correct task completion, subsequent reasoning faithfulness, task solutions, and propose metrics to evaluate reasoning chain correctness and to what extent final answers are faithful to reasoning traces of the generated CoT. We show a steep drop in ToM capabilities under task perturbation for all evaluated LLMs, questioning the notion of any robust form of ToM being present. While CoT prompting improves the ToM performance overall in a faithful manner, it surprisingly degrades accuracy for some perturbation classes, indicating that selective application is necessary.

📄 PDF Abstract BibTeX arXiv:2602.22072

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Revisiting the Evaluation of Theory of Mind through Question Answering

2019-11-01 · IJCNLP 2019 11 · Matthew Le, Y-Lan Boureau, Maximilian Nickel

Theory of mind, i.e., the ability to reason about intents and beliefs of agents is an important task in artificial intelligence and central to resolving ambiguous references in natural language dialogue. In this work, we…

Question Answering

Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

2023-02-16 · Tomer Ullman

Intuitive psychology is a pillar of common-sense reasoning. The replication of this reasoning in machine intelligence is an important stepping-stone on the way to human-like artificial intelligence. Several recent tasks …

Common Sense Reasoning

Proceedings of the 2nd Workshop on Advancing Artificial Intelligence through Theory of Mind

2026-03-19 · Nitay Alon, Joseph M. Barnby, Reuth Mirsky, Stefan Sarkadi arxiv

This volume includes a selection of papers presented at the 2nd Workshop on Advancing Artificial Intelligence through Theory of Mind held at AAAI 2026 in Singapore on 26th January 2026. The purpose of this volume is to p…

Multi-Agent Language Models: Advancing Cooperation, Coordination, and Adaptation

2025-06-11 · Arjun Vaithilingam Sudhakar

Modern Large Language Models (LLMs) exhibit impressive zero-shot and few-shot generalization capabilities across complex natural language tasks, enabling their widespread use as virtual assistants for diverse application…

Multi-agent Reinforcement Learning

What should I say? -- Interacting with AI and Natural Language Interfaces

2024-01-12 · Mark Adkins

As Artificial Intelligence (AI) technology becomes more and more prevalent, it becomes increasingly important to explore how we as humans interact with AI. The Human-AI Interaction (HAI) sub-field has emerged from the Hu…