paper-with-me

홈 › Papers

Trustworthy, Responsible, and Safe AI: A Comprehensive Architectural Framework for AI Safety with Challenges and Mitigations

2024-08-23 · Chen Chen, Xueluan Gong, Ziyao Liu, Weifeng Jiang, Si Qi Goh, Kwok-Yan Lam

AI Safety is an emerging area of critical importance to the safe adoption and deployment of AI systems. With the rapid proliferation of AI and especially with the recent advancement of Generative AI (or GAI), the technology ecosystem behind the design, development, adoption, and deployment of AI systems has drastically changed, broadening the scope of AI Safety to address impacts on public safety and national security. In this paper, we propose a novel architectural framework for understanding and analyzing AI Safety; defining its characteristics from three perspectives: Trustworthy AI, Responsible AI, and Safe AI. We provide an extensive review of current research and advancements in AI safety from these perspectives, highlighting their key challenges and mitigation approaches. Through examples from state-of-the-art technologies, particularly Large Language Models (LLMs), we present innovative mechanism, methodologies, and techniques for designing and testing AI safety. Our goal is to promote advancement in AI safety research, and ultimately enhance people's trust in digital transformation.

📄 PDF Abstract BibTeX arXiv:2408.12935

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Trustworthy and Responsible AI for Human-Centric Autonomous Decision-Making Systems

2024-08-28 · Farzaneh Dehghani, Mahsa Dibaji, Fahim Anzum, Lily Dey 외

Artificial Intelligence (AI) has paved the way for revolutionary decision-making processes, which if harnessed appropriately, can contribute to advancements in various sectors, from healthcare to economics. However, its …

Decision MakingFairness

Towards Trustworthy GUI Agents: A Survey

2025-03-30 · Yucheng Shi, Wenhao Yu, Wenlin Yao, Wenhu Chen 외

GUI agents, powered by large foundation models, can interact with digital interfaces, enabling various applications in web automation, mobile navigation, and software testing. However, their increasing autonomy has raise…

Decision MakingSequential Decision Makingsoftware testingSurvey

A Checklist for Trustworthy, Safe, and User-Friendly Mental Health Chatbots

2026-01-21 · Shreya Haran, Samiha Thatikonda, Dong Whi Yoo, Koustuv Saha arxiv

Mental health concerns are rising globally, prompting increased reliance on technology to address the demand-supply gap in mental health services. In particular, mental health chatbots are emerging as a promising solutio…

ResponsibleRobotBench: Benchmarking Responsible Robot Manipulation using Multi-modal Large Language Models

2025-12-03 · Lei Zhang, Ju Dong, Kaixin Bai, Minheng Ni 외 arxiv

Recent advances in large multimodal models have enabled new opportunities in embodied AI, particularly in robotic manipulation. These models have shown strong potential in generalization and reasoning, but achieving reli…

Robot Manipulation

Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI

2024-12-08 · Chao Yang, Chaochao Lu, Yingchun Wang, BoWen Zhou

Ensuring Artificial General Intelligence (AGI) reliably avoids harmful behaviors is a critical challenge, especially for systems with high autonomy or in safety-critical domains. Despite various safety assurance proposal…

Decision Making