paper-with-me

홈 › Papers

Connecting Generative Adversarial Networks and Actor-Critic Methods

2016-10-06 · David Pfau, Oriol Vinyals

Both generative adversarial networks (GAN) in unsupervised learning and actor-critic methods in reinforcement learning (RL) have gained a reputation for being difficult to optimize. Practitioners in both fields have amassed a large number of strategies to mitigate these instabilities and improve training. Here we show that GANs can be viewed as actor-critic methods in an environment where the actor cannot affect the reward. We review the strategies for stabilizing training for each class of models, both those that generalize between the two and those that are particular to that model. We also review a number of extensions to GANs and RL algorithms with even more complicated information flow. We hope that by highlighting this formal connection we will encourage both GAN and RL communities to develop general, scalable, and stable algorithms for multilevel optimization with deep networks, and to draw inspiration across communities.

📄 PDF Abstract BibTeX arXiv:1610.01945

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

AC-SUM-GAN: Connecting Actor-Critic and Generative Adversarial Networks for Unsupervised Video Summarization

2020-11-16 · IEEE Transactions on Circuits and Systems for Video Technology 2020 11 · Evlampios Apostolidis, Eleni Adamantidou, Alexandros I. Metsai, Vasileios Mezaris 외

This paper presents a new method for unsupervised video summarization. The proposed architecture embeds an Actor-Critic model into a Generative Adversarial Network and formulates the selection of important video fragment…

Generative Adversarial NetworkUnsupervised Video SummarizationVideo Summarization

Adversarial Advantage Actor-Critic Model for Task-Completion Dialogue Policy Learning

2017-10-31 · Baolin Peng, Xiujun Li, Jianfeng Gao, Jingjing Liu 외

This paper presents a new method --- adversarial advantage actor-critic (Adversarial A2C), which significantly improves the efficiency of dialogue policy learning in task-completion dialogue systems. Inspired by generati…

Task-Completion Dialogue Policy Learning

The Limit Points of (Optimistic) Gradient Descent in Min-Max Optimization

2018-07-11 · NeurIPS 2018 12 · Constantinos Daskalakis, Ioannis Panageas

Motivated by applications in Optimization, Game Theory, and the training of Generative Adversarial Networks, the convergence properties of first order methods in min-max problems have received extensive study. It has bee…

Critical heat flux diagnosis using conditional generative adversarial networks

2023-05-04 · UngJin Na, Moonhee Choi, HangJin Jo

The critical heat flux (CHF) is an essential safety boundary in boiling heat transfer processes employed in high heat flux thermal-hydraulic systems. Identifying CHF is vital for preventing equipment damage and ensuring …

Image-to-Image Translation

ACtuAL: Actor-Critic Under Adversarial Learning

2017-11-13 · Anirudh Goyal, Nan Rosemary Ke, Alex Lamb, R. Devon Hjelm 외

Generative Adversarial Networks (GANs) are a powerful framework for deep generative modeling. Posed as a two-player minimax problem, GANs are typically trained end-to-end on real-valued data and can be used to train a ge…

Language ModelingLanguage ModellingReinforcement Learning