paper-with-me

홈 › Papers

Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?

2024-07-02 · Kristian González Barman, Simon Lohse, Henk de Regt

We argue for the epistemic and ethical advantages of pluralism in Reinforcement Learning from Human Feedback (RLHF) in the context of Large Language Models (LLM). Drawing on social epistemology and pluralist philosophy of science, we suggest ways in which RHLF can be made more responsive to human needs and how we can address challenges along the way. The paper concludes with an agenda for change, i.e. concrete, actionable steps to improve LLM development.

📄 PDF Abstract BibTeX arXiv:2407.17482

Code (0)

등록된 구현이 없습니다.

Tasks

Philosophyreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Whose Norms? Disentangling Cultural and Personal Alignment in Large Language Models

2026-06-05 · Angana Borah, Isabelle Augenstein, Rada Mihalcea arxiv

Large language models are increasingly used for social decision-making situations that require balancing cultural norms with personal preferences. For example, a user preferring honesty might ask whether to correct a cow…

NaRLE: Natural Language Models using Reinforcement Learning with Emotion Feedback

2021-10-05 · Ruijie Zhou, Soham Deshmukh, Jeremiah Greer, Charles Lee

Current research in dialogue systems is focused on conversational assistants working on short conversations in either task-oriented or open domain settings. In this paper, we focus on improving task-based conversational …

Deep Reinforcement Learningintent-classificationIntent ClassificationNatural Language Understanding+3

Reinforcement Learning from User Feedback

2025-05-20 · Eric Han, Jun Chen, Karthik Abinav Sankararaman, Xiaoliang Peng 외

As large language models (LLMs) are increasingly deployed in diverse user facing applications, aligning them with real user preferences becomes essential. Existing methods like Reinforcement Learning from Human Feedback …

reinforcement-learningReinforcement Learning

Latency-aware Human-in-the-Loop Reinforcement Learning for Semantic Communications

2026-02-17 · Peizheng Li, Xinyi Lin, Adnan Aijaz arxiv

Semantic communication promises task-aligned transmission but must reconcile semantic fidelity with stringent latency guarantees in immersive and safety-critical services. This paper introduces a time-constrained human-i…

Reinforcement LearningSemantic Communication

Finding Culture-Sensitive Neurons in Vision-Language Models

2025-10-28 · Xiutian Zhao, Rochelle Choenni, Rohit Saxena, Ivan Titov arxiv

Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs process culturally grounded information, we study the presence of culture-sensitive…

Visual Question Answering