Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?
We argue for the epistemic and ethical advantages of pluralism in Reinforcement Learning from Human Feedback (RLHF) in the context of Large Language Models (LLM). Drawing on social epistemology and pluralist philosophy of science, we suggest ways in which RHLF can be made more responsive to human needs and how we can address challenges along the way. The paper concludes with an agenda for change, i.e. concrete, actionable steps to improve LLM development.
Code (0)
등록된 구현이 없습니다.
Tasks
Philosophyreinforcement-learningReinforcement LearningSimilar Papers 제목 키워드 기반
Whose Norms? Disentangling Cultural and Personal Alignment in Large Language Models
Large language models are increasingly used for social decision-making situations that require balancing cultural norms with personal preferences. For example, a user preferring honesty might ask whether to correct a cow…
NaRLE: Natural Language Models using Reinforcement Learning with Emotion Feedback
Current research in dialogue systems is focused on conversational assistants working on short conversations in either task-oriented or open domain settings. In this paper, we focus on improving task-based conversational …
Deep Reinforcement Learningintent-classificationIntent ClassificationNatural Language Understanding+3Reinforcement Learning from User Feedback
As large language models (LLMs) are increasingly deployed in diverse user facing applications, aligning them with real user preferences becomes essential. Existing methods like Reinforcement Learning from Human Feedback …
reinforcement-learningReinforcement LearningLatency-aware Human-in-the-Loop Reinforcement Learning for Semantic Communications
Semantic communication promises task-aligned transmission but must reconcile semantic fidelity with stringent latency guarantees in immersive and safety-critical services. This paper introduces a time-constrained human-i…
Reinforcement LearningSemantic CommunicationFinding Culture-Sensitive Neurons in Vision-Language Models
Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs process culturally grounded information, we study the presence of culture-sensitive…
Visual Question Answering