paper-with-me

Papers

Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach

2025-05-14 · Shannon Lodoen, Alexi Orchard

Since 2022, versions of generative AI chatbots such as ChatGPT and Claude have been trained using a specialized technique called Reinforcement Learning from Human Feedback (RLHF) to fine-tune language model output using feedback from human annotators. As a result, the integration of RLHF has greatly enhanced the outputs of these large language models (LLMs) and made the interactions and responses appear more "human-like" than those of previous versions using only supervised learning. The increasing convergence of human and machine-written text has potentially severe ethical, sociotechnical, and pedagogical implications relating to transparency, trust, bias, and interpersonal relations. To highlight these implications, this paper presents a rhetorical analysis of some of the central procedures and processes currently being reshaped by RLHF-enhanced generative AI chatbots: upholding language conventions, information seeking practices, and expectations for social relationships. Rhetorical investigations of generative AI and LLMs have, to this point, focused largely on the persuasiveness of the content generated. Using Ian Bogost's concept of procedural rhetoric, this paper shifts the site of rhetorical investigation from content analysis to the underlying mechanisms of persuasion built into RLHF-enhanced LLMs. In doing so, this theoretical investigation opens a new direction for further inquiry in AI ethics that considers how procedures rerouted through AI-driven technologies might reinforce hegemonic language use, perpetuate biases, decontextualize learning, and encroach upon human relationships. It will therefore be of interest to educators, researchers, scholars, and the growing number of users of generative AI chatbots.

📄 PDF Abstract BibTeX arXiv:2505.09576

Code (0)

등록된 구현이 없습니다.

Tasks

EthicsPersuasiveness

Similar Papers 제목 키워드 기반

Refine and Imitate: Reducing Repetition and Inconsistency in Persuasion Dialogues via Reinforcement Learning and Human Demonstration

2020-12-31 · Findings (EMNLP) 2021 11 · Weiyan Shi, Yu Li, Saurav Sahay, Zhou Yu

Persuasion dialogue systems reflect the machine's ability to make strategic moves beyond verbal communication, and therefore differentiate themselves from task-oriented or open-domain dialogue systems and have their own …

Language ModellingReinforcement Learning (RL)Response GenerationSentence

The Ethics of Generative AI

2025-12-04 · Michael Klenk arxiv

This chapter discusses the ethics of generative AI. It provides a technical primer to show how generative AI affords experiencing technology as if it were human, and this affordance provides a fruitful focus for the phil…

Towards Strategic Persuasion with Language Models

2025-09-26 · Zirui Cheng, Jiaxuan You arxiv

Large language models (LLMs) have demonstrated strong persuasive capabilities comparable to those of humans, offering promising benefits while raising societal concerns. However, systematically evaluating the persuasive …

Reinforcement Learning

Efficient Model-agnostic Alignment via Bayesian Persuasion

2024-05-29 · Fengshuo Bai, Mingzhi Wang, Zhaowei Zhang, Boyuan Chen 외

With recent advancements in large language models (LLMs), alignment has emerged as an effective technique for keeping LLMs consensus with human intent. Current methods primarily involve direct training through Supervised…

Code GenerationMathematical Reasoningmodel

Refine and Imitate: Reducing Repetition and Inconsistency in Dialogue Generation via Reinforcement Learning and Human Demonstration

2021-01-01 · Weiyan Shi, Yu Li, Saurav Sahay, Zhou Yu

Despite the recent success of large-scale language models on various downstream NLP tasks, the repetition and inconsistency problems still persist in dialogue response generation. Previous approaches have attempted to av…

Dialogue GenerationLanguage ModelingLanguage ModellingResponse Generation+1