An AGI with Time-Inconsistent Preferences
This paper reveals a trap for artificial general intelligence (AGI) theorists who use economists' standard method of discounting. This trap is implicitly and falsely assuming that a rational AGI would have time-consistent preferences. An agent with time-inconsistent preferences knows that its future self will disagree with its current self concerning intertemporal decision making. Such an agent cannot automatically trust its future self to carry out plans that its current self considers optimal.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingSimilar Papers 제목 키워드 기반
Optimal exit decision of venture capital under time-inconsistent preferences
This paper proposes two kinds of time-inconsistent preferences (i.e. time flow inconsistency and critical time point inconsistency) to further advance the research on the exit decision of venture capital. Time-inconsiste…
An Integral Equation in Portfolio Selection with Time-Inconsistent Preferences
This paper discusses a nonlinear integral equation arising from portfolio selection with a class of time-inconsistent preferences. We propose a unified framework requiring minimal assumptions, such as right-continuity of…
A Rule-Based Approach to Specifying Preferences over Conflicting Facts and Querying Inconsistent Knowledge Bases
Repair-based semantics have been extensively studied as a means of obtaining meaningful answers to queries posed over inconsistent knowledge bases (KBs). While several works have considered how to exploit a priority rela…
Learning the Preferences of Ignorant, Inconsistent Agents
An important use of machine learning is to learn what people value. What posts or photos should a user be shown? Which jobs or activities would a person find rewarding? In each case, observations of people's past choices…
Decision MakingTeaching Precommitted Agents: Model-Free Policy Evaluation and Control in Quasi-Hyperbolic Discounted MDPs
Time-inconsistent preferences, where agents favor smaller-sooner over larger-later rewards, are a key feature of human and animal decision-making. Quasi-Hyperbolic (QH) discounting provides a simple yet powerful model fo…
Reinforcement Learning