paper-with-me

홈 › Papers

A dataset of questions on decision-theoretic reasoning in Newcomb-like problems

2024-11-15 · Caspar Oesterheld, Emery Cooper, Miles Kodama, Linh Chi Nguyen, Ethan Perez

We introduce a dataset of natural-language questions in the decision theory of so-called Newcomb-like problems. Newcomb-like problems include, for instance, decision problems in which an agent interacts with a similar other agent, and thus has to reason about the fact that the other agent will likely reason in similar ways. Evaluating LLM reasoning about Newcomb-like problems is important because interactions between foundation-model-based agents will often be Newcomb-like. Some ways of reasoning about Newcomb-like problems may allow for greater cooperation between models. Our dataset contains both capabilities questions (i.e., questions with a unique, uncontroversially correct answer) and attitude questions (i.e., questions about which decision theorists would disagree). We use our dataset for an investigation of decision-theoretical capabilities and expressed attitudes and their interplay in existing models (different models by OpenAI, Anthropic, Meta, GDM, Reka, etc.), as well as models under simple prompt-based interventions. We find, among other things, that attitudes vary significantly between existing models; that high capabilities are associated with attitudes more favorable toward so-called evidential decision theory; and that attitudes are consistent across different types of questions.

📄 PDF Abstract BibTeX arXiv:2411.10588

Code (1)

casparoe/newcomblike_questions_dataset 공식 구현

Similar Papers 제목 키워드 기반

Reinforcement Learning in Newcomblike Environments

2021-12-01 · NeurIPS 2021 12 · James Bell, Linda Linsefors, Caspar Oesterheld, Joar Skalse

Newcomblike decision problems have been studied extensively in the decision theory literature, but they have so far been largely absent in the reinforcement learning literature. In this paper we study value-based reinfor…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

A Comparison of Decision Algorithms on Newcomblike Problems

2023-05-31 · Alex Altair

When formulated using Bayesian networks, two standard decision algorithms (Evidential Decision Theory and Causal Decision Theory) can be shown to fail systematically when faced with aspects of the prisoner's dilemma and …

Can CDT rationalise the ex ante optimal policy via modified anthropics?

2024-11-07 · Emery Cooper, Caspar Oesterheld, Vincent Conitzer

In Newcomb's problem, causal decision theory (CDT) recommends two-boxing and thus comes apart from evidential decision theory (EDT) and ex ante policy optimisation (which prescribe one-boxing). However, in Newcomb's prob…

Purely Bayesian counterfactuals versus Newcomb's paradox

2020-08-10 · Lê Nguyên Hoang

This paper proposes a careful separation between an entity's epistemic system and their decision system. Crucially, Bayesian counterfactuals are estimated by the epistemic system; not by the decision system. Based on thi…

counterfactual

Reinforcement Learning in Conflicting Environments for Autonomous Vehicles

2016-10-22 · Dominik Meyer, Johannes Feldmaier, Hao Shen

In this work, we investigate the application of Reinforcement Learning to two well known decision dilemmas, namely Newcomb's Problem and Prisoner's Dilemma. These problems are exemplary for dilemmas that autonomous agent…

Autonomous Vehiclesreinforcement-learningReinforcement LearningReinforcement Learning (RL)