Bridging the AI Trustworthiness Gap between Functions and Norms
Trustworthy Artificial Intelligence (TAI) is gaining traction due to regulations and functional benefits. While Functional TAI (FTAI) focuses on how to implement trustworthy systems, Normative TAI (NTAI) focuses on regulations that need to be enforced. However, gaps between FTAI and NTAI remain, making it difficult to assess trustworthiness of AI systems. We argue that a bridge is needed, specifically by introducing a conceptual language which can match FTAI and NTAI. Such a semantic language can assist developers as a framework to assess AI systems in terms of trustworthiness. It can also help stakeholders translate norms and regulations into concrete implementation steps for their systems. In this position paper, we describe the current state-of-the-art and identify the gap between FTAI and NTAI. We will discuss starting points for developing a semantic language and the envisioned effects of it. Finally, we provide key considerations and discuss future actions towards assessment of TAI.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
MORAL: Aligning AI with Human Norms through Multi-Objective Reinforced Active Learning
Inferring reward functions from demonstrations and pairwise preferences are auspicious approaches for aligning Reinforcement Learning (RL) agents with human intentions. However, state-of-the art methods typically focus o…
Active LearningEthicsReinforcement Learning (RL)Generalization Analysis for Deep Contrastive Representation Learning
In this paper, we present generalization bounds for the unsupervised risk in the Deep Contrastive Representation Learning framework, which employs deep neural networks as representation functions. We approach this proble…
Contrastive LearningGeneralization BoundsRepresentation LearningAI Ethics and Social Norms: Exploring ChatGPT's Capabilities From What to How
Using LLMs in healthcare, Computer-Supported Cooperative Work, and Social Computing requires the examination of ethical and social norms to ensure safe incorporation into human life. We conducted a mixed-method study, in…
EthicsOn the relation between Loss Functions and T-Norms
Deep learning has been shown to achieve impressive results in several domains like computer vision and natural language processing. A key element of this success has been the development of new loss functions, like the p…
RelationMapping Trustworthiness in Large Language Models: A Bibliometric Analysis Bridging Theory to Practice
The rapid proliferation of Large Language Models (LLMs) has raised significant trustworthiness and ethical concerns. Despite the widespread adoption of LLMs across domains, there is still no clear consensus on how to def…
EthicsFairnessRAGRetrieval-augmented Generation