paper-with-me

Papers

Personalized Reward Learning with Interaction-Grounded Learning (IGL)

2022-11-28 · Jessica Maghakian, Paul Mineiro, Kishan Panaganti, Mark Rucker, Akanksha Saran, Cheng Tan

In an era of countless content offerings, recommender systems alleviate information overload by providing users with personalized content suggestions. Due to the scarcity of explicit user feedback, modern recommender systems typically optimize for the same fixed combination of implicit feedback signals across all users. However, this approach disregards a growing body of work highlighting that (i) implicit signals can be used by users in diverse ways, signaling anything from satisfaction to active dislike, and (ii) different users communicate preferences in different ways. We propose applying the recent Interaction Grounded Learning (IGL) paradigm to address the challenge of learning representations of diverse user communication modalities. Rather than requiring a fixed, human-designed reward function, IGL is able to learn personalized reward functions for different users and then optimize directly for the latent user satisfaction. We demonstrate the success of IGL with experiments using simulations as well as with real-world production traces.

📄 PDF Abstract BibTeX arXiv:2211.15823

Code (1)

asaran/IGL-P 공식 구현

Tasks

Recommendation Systems

Similar Papers 제목 키워드 기반

Interaction-Grounded Learning for Contextual Markov Decision Processes with Personalized Feedback

2026-02-09 · Mengxiao Zhang, Yuheng Zhang, Haipeng Luo, Paul Mineiro arxiv

In this paper, we study Interaction-Grounded Learning (IGL) [Xie et al., 2021], a paradigm designed for realistic scenarios where the learner receives indirect feedback generated by an unknown mechanism, rather than expl…

Provably Efficient Interactive-Grounded Learning with Personalized Reward

2024-05-31 · Mengxiao Zhang, Yuheng Zhang, Haipeng Luo, Paul Mineiro

Interactive-Grounded Learning (IGL) [Xie et al., 2021] is a powerful framework in which a learner aims at maximizing unobservable rewards through interacting with an environment and observing reward-dependent feedback on…

Recommendation Systems

Interaction-Grounded Learning

2021-06-09 · Tengyang Xie, John Langford, Paul Mineiro, Ida Momennejad

Consider a prosthetic arm, learning to adapt to its user's control signals. We propose Interaction-Grounded Learning for this novel setting, in which a learner's goal is to interact with the environment with no grounding…

P-Check: Advancing Personalized Reward Model via Learning to Generate Dynamic Checklist

2026-01-06 · Kwangwook Seo, Dongha Lee arxiv

Recent approaches in personalized reward modeling have primarily focused on leveraging user interaction history to align model judgments with individual preferences. However, existing approaches largely treat user contex…

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users

2026-05-30 · Wuqiang Zheng, Chengbing Wang, Yilin Yang, Junyi Cheng 외 arxiv

As Large Language Models (LLMs) are increasingly deployed in long-term interactions with users, empathy has become an increasingly important capability. However, existing research overlooks the influence of users' person…