paper-with-me

홈 › Papers

Understanding and Avoiding AI Failures: A Practical Guide

2021-04-22 · Heather M. Williams, Roman V. Yampolskiy

As AI technologies increase in capability and ubiquity, AI accidents are becoming more common. Based on normal accident theory, high reliability theory, and open systems theory, we create a framework for understanding the risks associated with AI applications. In addition, we also use AI safety principles to quantify the unique risks of increased intelligence and human-like qualities in AI. Together, these two fields give a more complete picture of the risks of contemporary AI. By focusing on system properties near accidents instead of seeking a root cause of accidents, we identify where attention should be paid to safety for current generation AI systems.

📄 PDF Abstract BibTeX arXiv:2104.12582

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reinforcement Learning with Probabilistic Boolean Network Models of Smart Grid Devices

2021-02-02 · Pedro J. Rivera Torres, Carlos Gershenson García, Samir Kanaan Izquierdo

The area of Smart Power Grids needs to constantly improve its efficiency and resilience, to pro-vide high quality electrical power, in a resistant grid, managing faults and avoiding failures. Achieving this requires high…

reinforcement-learningReinforcement Learning (RL)

ERR@HRI 2.0 Challenge: Multimodal Detection of Errors and Failures in Human-Robot Conversations

2025-07-17 · Shiye Cao, Maia Stiber, Amama Mahmood, Maria Teresa Parreira 외 arxiv

The integration of large language models (LLMs) into conversational robots has made human-robot conversations more dynamic. Yet, LLM-powered conversational robots remain prone to errors, e.g., misunderstanding user inten…

SoK: Access Control Policy Generation from High-level Natural Language Requirements

2023-10-05 · Sakuna Harinda Jayasundara, Nalin Asanka Gamagedara Arachchilage, Giovanni Russello

Administrator-centered access control failures can cause data breaches, putting organizations at risk of financial loss and reputation damage. Existing graphical policy configuration tools and automated policy generation…

Systematic Literature Review

RefactorAssist: Agentic Refinement for Reliable Code Refactoring

2026-08-02 · Jonathan Cordeiro, Shayan Noei, Ying Zou arxiv

Code refactoring aims to enhance the internal structure of source code without affecting its functional behavior. The recent advancements of Large Language Models (LLMs) have demonstrated potential for automating softwar…

SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding

2026-01-14 · Shuyang Hou, Yi Hu, Muhan Zhang arxiv

Recent advancements in large language models (LLMs) have significantly enhanced their reasoning capabilities. However, they continue to struggle with basic character-level tasks, such as counting letters in words, a prob…