paper-with-me

Papers

Reinforcement Learning for Generative AI: A Survey

2023-08-28 · Yuanjiang Cao, Quan Z. Sheng, Julian McAuley, Lina Yao

Deep Generative AI has been a long-standing essential topic in the machine learning community, which can impact a number of application areas like text generation and computer vision. The major paradigm to train a generative model is maximum likelihood estimation, which pushes the learner to capture and approximate the target data distribution by decreasing the divergence between the model distribution and the target distribution. This formulation successfully establishes the objective of generative tasks, while it is incapable of satisfying all the requirements that a user might expect from a generative model. Reinforcement learning, serving as a competitive option to inject new training signals by creating new objectives that exploit novel signals, has demonstrated its power and flexibility to incorporate human inductive bias from multiple angles, such as adversarial learning, hand-designed rules and learned reward model to build a performant model. Thereby, reinforcement learning has become a trending research field and has stretched the limits of generative AI in both model design and application. It is reasonable to summarize and conclude advances in recent years with a comprehensive review. Although there are surveys in different application areas recently, this survey aims to shed light on a high-level review that spans a range of application areas. We provide a rigorous taxonomy in this area and make sufficient coverage on various models and applications. Notably, we also surveyed the fast-developing large language model area. We conclude this survey by showing the potential directions that might tackle the limit of current models and expand the frontiers for generative AI.

📄 PDF Abstract BibTeX arXiv:2308.14328

Code (0)

등록된 구현이 없습니다.

Tasks

Inductive BiasLanguage ModellingLarge Language Modelreinforcement-learningReinforcement LearningSurveyText Generation

Similar Papers 제목 키워드 기반

Reinforcement Learning for Generative AI: State of the Art, Opportunities and Open Research Challenges

2023-07-31 · Giorgio Franceschelli, Mirco Musolesi

Generative Artificial Intelligence (AI) is one of the most exciting developments in Computer Science of the last decade. At the same time, Reinforcement Learning (RL) has emerged as a very successful paradigm for a varie…

Reinforcement Learning (RL)Survey

A Survey on Data-Centric AI: Tabular Learning from Reinforcement Learning and Generative AI Perspective

2025-02-12 · Wangyang Ying, Cong Wei, Nanxu Gong, Xinyuan Wang 외

Tabular data is one of the most widely used data formats across various domains such as bioinformatics, healthcare, and marketing. As artificial intelligence moves towards a data-centric perspective, improving data quali…

Feature Engineeringfeature selectionMarketingReinforcement Learning (RL)+1

Diffusion Models for Reinforcement Learning: A Survey

2023-11-02 · Zhengbang Zhu, Hanye Zhao, Haoran He, Yichao Zhong 외

Diffusion models surpass previous generative models in sample quality and training stability. Recent works have shown the advantages of diffusion models in improving reinforcement learning (RL) solutions. This survey aim…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Survey

Advances in GRPO for Generation Models: A Survey

2026-02-21 · Zexiang Liu, Xianglong He, Yangguang Li arxiv

Large-scale flow matching models have achieved strong performance across generative tasks such as text-to-image, video, 3D, and speech synthesis. However, aligning their outputs with human preferences and task-specific o…

Reinforcement LearningSpeech SynthesisVideo GenerationImage Editing

Deep Generative Models in Robotics: A Survey on Learning from Multimodal Demonstrations

2024-08-08 · Julen Urain, Ajay Mandlekar, Yilun Du, Mahi Shafiullah 외

Learning from Demonstrations, the field that proposes to learn robot behavior models from data, is gaining popularity with the emergence of deep generative models. Although the problem has been studied for years under na…

Grasp GenerationImitation LearningSurvey