paper-with-me

Papers

Toward Virtuous Reinforcement Learning: A Critique and Roadmap

2025-12-03 · Majid Ghasemi, Mark Crowley arxiv

This paper critiques common patterns in machine ethics for Reinforcement Learning (RL) and argues for a virtue focused alternative. We highlight two recurring limitations in much of the current literature: (i) rule based (deontological) methods that encode duties as constraints or shields often struggle under ambiguity and nonstationarity and do not cultivate lasting habits, and (ii) many reward based approaches, especially single objective RL, implicitly compress diverse moral considerations into a single scalar signal, which can obscure trade offs and invite proxy gaming in practice. We instead treat ethics as policy level dispositions, that is, relatively stable habits that hold up when incentives, partners, or contexts change. This shifts evaluation beyond rule checks or scalar returns toward trait summaries, durability under interventions, and explicit reporting of moral trade offs. Our roadmap combines four components: (1) social learning in multi agent RL to acquire virtue like patterns from imperfect but normatively informed exemplars; (2) multi objective and constrained formulations that preserve value conflicts and incorporate risk aware criteria to guard against harm; (3) affinity based regularization toward updateable virtue priors that support trait like stability under distribution shift while allowing norms to evolve; and (4) operationalizing diverse ethical traditions as practical control signals, making explicit the value and cultural assumptions that shape ethical RL benchmarks.

📄 PDF Abstract BibTeX arXiv:2512.04246

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

2026-04-30 · Jiacheng Liu, Zichen Tang, Zhongjun Yang, Xinyi Hu 외 arxiv

People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers through hierarchical subtasks to solve complex research problems step …

Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learning in a Game Environment

2026-06-03 · Ajay Vishwanath, Christian Omlin arxiv

Instilling virtuous behavior in artificial intelligence has seen increasing interest. One of the techniques proposed is known as affinity-based reinforcement learning, which uses policy regularization on the objective fu…

Reinforcement Learning

Bootstrapping Exploration with Group-Level Natural Language Feedback in Reinforcement Learning

2026-03-04 · Lei Huang, Xiang Cheng, Chenxiao Zhao, Guobin Shen 외 arxiv

Large language models (LLMs) typically receive diverse natural language (NL) feedback through interaction with the environment. However, current reinforcement learning (RL) algorithms rely solely on scalar rewards, leavi…

Reinforcement Learning

Neuroprospecting with DeepRL agents

2021-09-24 · NeurIPS Workshop AI4Scien 2021 12 · Satpreet Harcharan Singh

A virtuous cycle between neuroscience and deep reinforcement learning is emerging, and the AI community can do much to enable and accelerate it.

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Virtuously Safe Reinforcement Learning

2018-05-29 · Henrik Aslund, El Mahdi El Mhamdi, Rachid Guerraoui, Alexandre Maurer

We show that when a third party, the adversary, steps into the two-party setting (agent and operator) of safely interruptible reinforcement learning, a trade-off has to be made between the probability of following the op…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Exploration+1