paper-with-me

Papers

AI Alignment Dialogues: An Interactive Approach to AI Alignment in Support Agents

2023-01-16 · Pei-Yu Chen, Myrthe L. Tielman, Dirk K. J. Heylen, Catholijn M. Jonker, M. Birna van Riemsdijk

AI alignment is about ensuring AI systems only pursue goals and activities that are beneficial to humans. Most of the current approach to AI alignment is to learn what humans value from their behavioural data. This paper proposes a different way of looking at the notion of alignment, namely by introducing AI Alignment Dialogues: dialogues with which users and agents try to achieve and maintain alignment via interaction. We argue that alignment dialogues have a number of advantages in comparison to data-driven approaches, especially for behaviour support agents, which aim to support users in achieving their desired future behaviours rather than their current behaviours. The advantages of alignment dialogues include allowing the users to directly convey higher-level concepts to the agent, and making the agent more transparent and trustworthy. In this paper we outline the concept and high-level structure of alignment dialogues. Moreover, we conducted a qualitative focus group user study from which we developed a model that describes how alignment dialogues affect users, and created design suggestions for AI alignment dialogues. Through this we establish foundations for AI alignment dialogues and shed light on what requires further development and research.

📄 PDF Abstract BibTeX arXiv:2301.06421

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models

2024-02-05 · Ivar Frisch, Mario Giulianelli

While both agent interaction and personalisation are vibrant topics in research on large language models (LLMs), there has been limited focus on the effect of language interaction on the behaviour of persona-conditioned …

Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes

2025-09-07 · Abhijnan Nath, Carine Graff, Nikhil Krishnaswamy arxiv

As Large Language Models (LLMs) get integrated into diverse workflows, they are increasingly being regarded as "collaborators" with humans, and required to work in coordination with other AI systems. If such AI collabora…

ERABAL: Enhancing Role-Playing Agents through Boundary-Aware Learning

2024-09-23 · Yihong Tang, Jiao Ou, Che Liu, Fuzheng Zhang 외

Role-playing is an emerging application in the field of Human-Computer Interaction (HCI), primarily implemented through the alignment training of a large language model (LLM) with assigned characters. Despite significant…

Language ModelingLanguage ModellingLarge Language Model

Hybrid LLM-Embedded Dialogue Agents for Learner Reflection: Designing Responsive and Theory-Driven Interactions

2026-02-24 · Paras Sharma, YuePing Sha, Janet Shufor Bih Epse Fofang, Brayden Yan 외 arxiv

Dialogue systems have long supported learner reflections, with theoretically grounded, rule-based designs offering structured scaffolding but often struggling to respond to shifts in engagement. Large Language Models (LL…

Enhancing User Engagement in Socially-Driven Dialogue through Interactive LLM Alignments

2025-06-26 · Jiashuo Wang, Kaitao Song, Chunpu Xu, Changhe Song 외

Enhancing user engagement through interactions plays an essential role in socially-driven dialogues. While prior works have optimized models to reason over relevant knowledge or plan a dialogue act flow, the relationship…