paper-with-me

홈 › Papers

An AGI with Time-Inconsistent Preferences

2019-06-23 · James D. Miller, Roman Yampolskiy

This paper reveals a trap for artificial general intelligence (AGI) theorists who use economists' standard method of discounting. This trap is implicitly and falsely assuming that a rational AGI would have time-consistent preferences. An agent with time-inconsistent preferences knows that its future self will disagree with its current self concerning intertemporal decision making. Such an agent cannot automatically trust its future self to carry out plans that its current self considers optimal.

📄 PDF Abstract BibTeX arXiv:1906.10536

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Optimal exit decision of venture capital under time-inconsistent preferences

2021-03-22 · Yanzhao Li, Ju'e Guo, Yongwu Li, Xu Zhang

This paper proposes two kinds of time-inconsistent preferences (i.e. time flow inconsistency and critical time point inconsistency) to further advance the research on the exit decision of venture capital. Time-inconsiste…

An Integral Equation in Portfolio Selection with Time-Inconsistent Preferences

2024-12-03 · Zongxia Liang, Sheng Wang, Jianming Xia

This paper discusses a nonlinear integral equation arising from portfolio selection with a class of time-inconsistent preferences. We propose a unified framework requiring minimal assumptions, such as right-continuity of…

A Rule-Based Approach to Specifying Preferences over Conflicting Facts and Querying Inconsistent Knowledge Bases

2025-08-11 · Meghyn Bienvenu, Camille Bourgaux, Katsumi Inoue, Robin Jean arxiv

Repair-based semantics have been extensively studied as a means of obtaining meaningful answers to queries posed over inconsistent knowledge bases (KBs). While several works have considered how to exploit a priority rela…

Learning the Preferences of Ignorant, Inconsistent Agents

2015-12-18 · Owain Evans, Andreas Stuhlmueller, Noah D. Goodman

An important use of machine learning is to learn what people value. What posts or photos should a user be shown? Which jobs or activities would a person find rewarding? In each case, observations of people's past choices…

Decision Making

Teaching Precommitted Agents: Model-Free Policy Evaluation and Control in Quasi-Hyperbolic Discounted MDPs

2025-09-07 · S. R. Eshwar arxiv

Time-inconsistent preferences, where agents favor smaller-sooner over larger-later rewards, are a key feature of human and animal decision-making. Quasi-Hyperbolic (QH) discounting provides a simple yet powerful model fo…

Reinforcement Learning