paper-with-me

홈 › Papers

A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning

2018-06-20 · Amy Zhang, Nicolas Ballas, Joelle Pineau

The risks and perils of overfitting in machine learning are well known. However most of the treatment of this, including diagnostic tools and remedies, was developed for the supervised learning case. In this work, we aim to offer new perspectives on the characterization and prevention of overfitting in deep Reinforcement Learning (RL) methods, with a particular focus on continuous domains. We examine several aspects, such as how to define and diagnose overfitting in MDPs, and how to reduce risks by injecting sufficient training diversity. This work complements recent findings on the brittleness of deep RL methods and offers practical observations for RL researchers and practitioners.

📄 PDF Abstract BibTeX arXiv:1806.07937

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningDeep Reinforcement LearningDiagnosticDiversityreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Grokking of Diffusion Models: Case Study on Modular Addition

2026-04-20 · Joon Hyeok Kim, Yong-Hyun Park, Mattis Dalsætra Østby, Jiatao Gu arxiv

Despite their empirical success, how diffusion models generalize remains poorly understood from a mechanistic perspective. We demonstrate that diffusion models trained with flow-matching objectives exhibit grokking--dela…

Generalization in Transfer Learning

2019-09-03 · Suzan Ece Ada, Emre Ugur, H. Levent Akin

Agents trained with deep reinforcement learning algorithms are capable of performing highly complex tasks including locomotion in continuous environments. We investigate transferring the learning acquired in one task to …

continuous-controlContinuous ControlDeep Reinforcement LearningFriction+5

A Survey Analyzing Generalization in Deep Reinforcement Learning

2024-01-04 · Ezgi Korkmaz

Reinforcement learning research obtained significant success and attention with the utilization of deep neural networks to solve problems in high dimensional state or action spaces. While deep reinforcement learning poli…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningSurvey

Learning dissection trajectories from expert surgical videos via imitation learning with equivariant diffusion

2025-06-05 · Hongyu Wang, Yonghao Long, Yueyao Chen, Hon-Chi Yip 외

Endoscopic Submucosal Dissection (ESD) is a well-established technique for removing epithelial lesions. Predicting dissection trajectories in ESD videos offers significant potential for enhancing surgical skill training …

Imitation LearningRepresentation LearningTrajectory Prediction

A Study on Overfitting in Deep Reinforcement Learning

2018-04-18 · Chiyuan Zhang, Oriol Vinyals, Remi Munos, Samy Bengio

Recent years have witnessed significant progresses in deep Reinforcement Learning (RL). Empowered with large scale neural networks, carefully designed architectures, novel training algorithms and massively parallel compu…

Deep Reinforcement LearningInductive Biasreinforcement-learningReinforcement Learning+1