paper-with-me

홈 › Papers

The five Is: Key principles for interpretable and safe conversational AI

2021-08-31 · Mattias Wahde, Marco Virgolin

In this position paper, we present five key principles, namely interpretability, inherent capability to explain, independent data, interactive learning, and inquisitiveness, for the development of conversational AI that, unlike the currently popular black box approaches, is transparent and accountable. At present, there is a growing concern with the use of black box statistical language models: While displaying impressive average performance, such systems are also prone to occasional spectacular failures, for which there is no clear remedy. In an effort to initiate a discussion on possible alternatives, we outline and exemplify how our five principles enable the development of conversational AI systems that are transparent and thus safer for use. We also present some of the challenges inherent in the implementation of those principles.

📄 PDF Abstract BibTeX arXiv:2108.13766

Code (0)

등록된 구현이 없습니다.

Tasks

Position

Similar Papers 제목 키워드 기반

Safety Bench: Identifying Safety-Sensitive Situations for Open-domain Conversational Systems

2021-10-16 · ACL ARR October 2021 10 · Anonymous

The social impact of natural language processing and its applications has received increasing attention. Here, we focus on the problem of safety for end-to-end conversational AI. We survey the problem landscape therein,…

Designing LMS and Instructional Strategies for Integrating Generative-Conversational AI

2025-08-31 · Elias Ra, Seung Je Kim, Eui-Yeong Seo, Geunju So arxiv

Higher education faces growing challenges in delivering personalized, scalable, and pedagogically coherent learning experiences. This study introduces a structured framework for designing an AI-powered Learning Managemen…

The Chai Platform's AI Safety Framework

2023-06-05 · Xiaoding Lu, Aleksey Korshuk, Zongyi Liu, William Beauchamp

Chai empowers users to create and interact with customized chatbots, offering unique and engaging experiences. Despite the exciting prospects, the work recognizes the inherent challenges of a commitment to modern safety …

Chatbot

Evaluating LLM Agent Adherence to Hierarchical Safety Principles: A Lightweight Benchmark for Probing Foundational Controllability Components

2025-06-03 · Ram Potham

Credible safety plans for advanced AI development require methods to verify agent behavior and detect potential control deficiencies early. A fundamental aspect is ensuring agents adhere to safety-critical principles, es…

SynBullying: A Multi LLM Synthetic Conversational Dataset for Cyberbullying Detection

2025-10-30 · Arefeh Kazemi, Hamza Qadeer, Joachim Wagner, Hossein Hosseini 외 arxiv

We introduce SynBullying, a synthetic multi-LLM conversational dataset for studying and detecting cyberbullying (CB). SynBullying provides a scalable and ethically safe alternative to human data collection by leveraging …