paper-with-me

홈 › Papers

Generative Adversarial Reward Learning for Generalized Behavior Tendency Inference

2021-05-03 · Xiaocong Chen, Lina Yao, Xianzhi Wang, Aixin Sun, Wenjie Zhang, Quan Z. Sheng

Recent advances in reinforcement learning have inspired increasing interest in learning user modeling adaptively through dynamic interactions, e.g., in reinforcement learning based recommender systems. Reward function is crucial for most of reinforcement learning applications as it can provide the guideline about the optimization. However, current reinforcement-learning-based methods rely on manually-defined reward functions, which cannot adapt to dynamic and noisy environments. Besides, they generally use task-specific reward functions that sacrifice generalization ability. We propose a generative inverse reinforcement learning for user behavioral preference modelling, to address the above issues. Instead of using predefined reward functions, our model can automatically learn the rewards from user's actions based on discriminative actor-critic network and Wasserstein GAN. Our model provides a general way of characterizing and explaining underlying behavioral tendencies, and our experiments show our method outperforms state-of-the-art methods in a variety of scenarios, namely traffic signal control, online recommender systems, and scanpath prediction.

📄 PDF Abstract BibTeX arXiv:2105.00822

Code (0)

등록된 구현이 없습니다.

Tasks

Recommendation Systemsreinforcement-learningReinforcement LearningReinforcement Learning (RL)Scanpath predictionTraffic Signal Control

Similar Papers 제목 키워드 기반

Adversarial Imitation via Variational Inverse Reinforcement Learning

2018-09-17 · ICLR 2019 5 · Ahmed H. Qureshi, Byron Boots, Michael C. Yip

We consider a problem of learning the reward and policy from expert examples under unknown dynamics. Our proposed method builds on the framework of generative adversarial networks and introduces the empowerment-regulariz…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Transfer Learning

Task-Relevant Adversarial Imitation Learning

2019-10-02 · Konrad Zolna, Scott Reed, Alexander Novikov, Sergio Gomez Colmenarejo 외

We show that a critical vulnerability in adversarial imitation is the tendency of discriminator networks to learn spurious associations between visual features and expert labels. When the discriminator focuses on task-ir…

Imitation Learning

Multi-Agent Generative Adversarial Imitation Learning

2018-07-26 · NeurIPS 2018 12 · Jiaming Song, Hongyu Ren, Dorsa Sadigh, Stefano Ermon

Imitation learning algorithms can be used to learn a policy from expert demonstrations without access to a reward signal. However, most existing approaches are not applicable in multi-agent settings due to the existence …

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Generative Adversarial User Model for Reinforcement Learning Based Recommendation System

2018-12-27 · Xinshi Chen, Shuang Li, Hui Li, Shaohua Jiang 외

There are great interests as well as many challenges in applying reinforcement learning (RL) to recommendation systems. In this setting, an online user is the environment; neither the reward function nor the environment …

Generative Adversarial NetworkModel-based Reinforcement LearningRecommendation Systemsreinforcement-learning+2

Generalized Dual Discriminator GANs

2025-07-23 · Penukonda Naga Chandana, Tejas Srivastava, Gowtham R. Kurri, V. Lalitha arxiv

Dual discriminator generative adversarial networks (D2 GANs) were introduced to mitigate the problem of mode collapse in generative adversarial networks. In D2 GANs, two discriminators are employed alongside a generator:…