paper-with-me

홈 › Papers

Avoiding Negative Side Effects due to Incomplete Knowledge of AI Systems

2020-08-24 · Sandhya Saisubramanian, Shlomo Zilberstein, Ece Kamar

Autonomous agents acting in the real-world often operate based on models that ignore certain aspects of the environment. The incompleteness of any given model -- handcrafted or machine acquired -- is inevitable due to practical limitations of any modeling technique for complex real-world settings. Due to the limited fidelity of its model, an agent's actions may have unexpected, undesirable consequences during execution. Learning to recognize and avoid such negative side effects of an agent's actions is critical to improve the safety and reliability of autonomous systems. Mitigating negative side effects is an emerging research topic that is attracting increased attention due to the rapid growth in the deployment of AI systems and their broad societal impacts. This article provides a comprehensive overview of different forms of negative side effects and the recent research efforts to address them. We identify key characteristics of negative side effects, highlight the challenges in avoiding negative side effects, and discuss recently developed approaches, contrasting their benefits and limitations. The article concludes with a discussion of open questions and suggestions for future research directions.

📄 PDF Abstract BibTeX arXiv:2008.12146

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Avoiding Side Effects in Complex Environments

2020-06-11 · NeurIPS 2020 12 · Alexander Matt Turner, Neale Ratzlaff, Prasad Tadepalli

Reward function specification can be difficult. Rewarding the agent for making a widget may be easy, but penalizing the multitude of possible negative side effects is hard. In toy environments, Attainable Utility Preserv…

DICNet: Deep Instance-Level Contrastive Network for Double Incomplete Multi-View Multi-Label Classification

2023-03-15 · Chengliang Liu, Jie Wen, Xiaoling Luo, Chao Huang 외

In recent years, multi-view multi-label learning has aroused extensive research enthusiasm. However, multi-view multi-label data in the real world is commonly incomplete due to the uncertain factors of data collection an…

Contrastive LearningMissing LabelsMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION+1

Semantic Vector Spaces for Broadening Consideration of Consequences

2018-02-23 · Douglas Summers Stay

Reasoning systems with too simple a model of the world and human intent are unable to consider potential negative side effects of their actions and modify their plans to avoid them (e.g., avoiding potential errors). Howe…

Common Sense Reasoning

Towards Socially and Morally Aware RL agent: Reward Design With LLM

2024-01-23 · Zhaoyue Wang

When we design and deploy an Reinforcement Learning (RL) agent, reward functions motivates agents to achieve an objective. An incorrect or incomplete specification of the objective can result in behavior that does not al…

Reinforcement Learning (RL)Safe Exploration

Temporal Planning with Incomplete Knowledge and Perceptual Information

2022-07-20 · Yaniel Carreno, Yvan Petillot, Ronald P. A. Petrick

In real-world applications, the ability to reason about incomplete knowledge, sensing, temporal notions, and numeric constraints is vital. While several AI planners are capable of dealing with some of these requirements,…