paper-with-me

Papers

Pareto-Optimal Learning from Preferences with Hidden Context

2024-06-21 · Ryan Boldi, Li Ding, Lee Spector, Scott Niekum

Ensuring AI models align with human values is essential for their safety and functionality. Reinforcement learning from human feedback (RLHF) uses human preferences to achieve this alignment. However, preferences sourced from diverse populations can result in point estimates of human values that may be sub-optimal or unfair to specific groups. We propose Pareto Optimal Preference Learning (POPL), which frames discrepant group preferences as objectives with potential trade-offs, aiming for policies that are Pareto-optimal on the preference dataset. POPL utilizes Lexicase selection, an iterative process to select diverse and Pareto-optimal solutions. Our empirical evaluations demonstrate that POPL surpasses baseline methods in learning sets of reward functions, effectively catering to distinct groups without access to group numbers or membership labels. Furthermore, we illustrate that POPL can serve as a foundation for techniques optimizing specific notions of group fairness, ensuring inclusive and equitable AI model alignment.

📄 PDF Abstract BibTeX arXiv:2406.15599

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Provably Efficient Multi-Objective Bandit Algorithms under Preference-Centric Customization

2025-02-19 · Linfeng Cao, Ming Shi, Ness B. Shroff

Multi-objective multi-armed bandit (MO-MAB) problems traditionally aim to achieve Pareto optimality. However, real-world scenarios often involve users with varying preferences across objectives, resulting in a Pareto-opt…

Efficiency and Sequenceability in Fair Division of Indivisible Goods with Additive Preferences

2016-04-06 · Sylvain Bouveret, Michel Lemaître

In fair division of indivisible goods, using sequences of sincere choices (or picking sequences) is a natural way to allocate the objects. The idea is the following: at each stage, a designated agent picks one object amo…

Fairness

Computing and Testing Pareto Optimal Committees

2018-03-18 · Haris Aziz, Jerome Lang, Jerome Monnot

Selecting a set of alternatives based on the preferences of agents is an important problem in committee selection and beyond. Among the various criteria put forth for the desirability of a committee, Pareto optimality is…

An Optimal Procedure to Check Pareto-Optimality in House Markets with Single-Peaked Preferences

2020-02-14 · Aurélie Beynier, Nicolas Maudet, Simon Rey, Parham Shams

Recently, the problem of allocating one resource per agent with initial endowments (house markets) has seen a renewed interest: indeed, while in the domain of strict preferences the Top Trading Cycle algorithm is known t…

User-Preference Meets Pareto-Optimality: Multi-Objective Bayesian Optimization with Local Gradient Search

2025-02-10 · Joshua Hang Sai Ip, Ankush Chakrabarty, Ali Mesbah, Diego Romeres

Incorporating user preferences into multi-objective Bayesian optimization (MOBO) allows for personalization of the optimization procedure. Preferences are often abstracted in the form of an unknown utility function, esti…

Bayesian Optimization