paper-with-me

홈 › Papers

Building Trust: Foundations of Security, Safety and Transparency in AI

2024-11-19 · Huzaifa Sidhpurwala, Garth Mollett, Emily Fox, Mark Bestavros, Huamin Chen

This paper explores the rapidly evolving ecosystem of publicly available AI models, and their potential implications on the security and safety landscape. As AI models become increasingly prevalent, understanding their potential risks and vulnerabilities is crucial. We review the current security and safety scenarios while highlighting challenges such as tracking issues, remediation, and the apparent absence of AI model lifecycle and ownership processes. Comprehensive strategies to enhance security and safety for both model developers and end-users are proposed. This paper aims to provide some of the foundational pieces for more standardized security, safety, and transparency in the development and operation of AI models and the larger open ecosystems and communities forming around them.

📄 PDF Abstract BibTeX arXiv:2411.12275

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

2026-05-17 · Jinhu Qi, Muzhi Li, Jiahong Liu, Yuqin Shu 외 arxiv

Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks autonomously, but their multi-step trajectories introduce new failu…

Blueprints of Trust: AI System Cards for End to End Transparency and Governance

2025-09-23 · Huzaifa Sidhpurwala, Emily Fox, Garth Mollett, Florencio Cano Gabarda 외 arxiv

This paper introduces the Hazard-Aware System Card (HASC), a novel framework designed to enhance transparency and accountability in the development and deployment of AI systems. The HASC builds upon existing model card a…

Building Trust in AI-Driven Decision Making for Cyber-Physical Systems (CPS): A Comprehensive Review

2024-05-10 · Rahul Umesh Mhapsekar, Muhammad Iftikhar Umrani, Malik Faizan, Omer Ali 외

Recent advancements in technology have led to the emergence of Cyber-Physical Systems (CPS), which seamlessly integrate the cyber and physical domains in various sectors such as agriculture, autonomous systems, and healt…

Decision Making

Inference-Time Safety For Code LLMs Via Retrieval-Augmented Revision

2026-03-02 · Manisha Mukherjee, Vincent J. Hellendoorn arxiv

Large Language Models (LLMs) are increasingly deployed for code generation in high-stakes software development, yet their limited transparency in security reasoning and brittleness to evolving vulnerability patterns rais…

Code Generation

Explainable AI is Responsible AI: How Explainability Creates Trustworthy and Socially Responsible Artificial Intelligence

2023-12-04 · Stephanie Baker, Wei Xiang

Artificial intelligence (AI) has been clearly established as a technology with the potential to revolutionize fields from healthcare to finance - if developed and deployed responsibly. This is the topic of responsible AI…

Fairness