paper-with-me

홈 › Papers

Faster and Safer Training by Embedding High-Level Knowledge into Deep Reinforcement Learning

2019-10-22 · Haodi Zhang, Zihang Gao, Yi Zhou, Hao Zhang, Kaishun Wu, Fangzhen Lin

Deep reinforcement learning has been successfully used in many dynamic decision making domains, especially those with very large state spaces. However, it is also well-known that deep reinforcement learning can be very slow and resource intensive. The resulting system is often brittle and difficult to explain. In this paper, we attempt to address some of these problems by proposing a framework of Rule-interposing Learning (RIL) that embeds high level rules into the deep reinforcement learning. With some good rules, this framework not only can accelerate the learning process, but also keep it away from catastrophic explorations, thus making the system relatively stable even during the very early stage of training. Moreover, given the rules are high level and easy to interpret, they can be easily maintained, updated and shared with other similar tasks.

📄 PDF Abstract BibTeX arXiv:1910.09986

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

2026-01-13 · Renyang Liu, Kangjie Chen, Han Qiu, Jie Zhang 외 arxiv

Image generation models (IGMs), while capable of producing impressive and creative content, often memorize a wide range of undesirable concepts from their training data, leading to the reproduction of unsafe content such…

Image Generation

SafeRoPE: Risk-specific Head-wise Embedding Rotation for Safe Generation in Rectified Flow Transformers

2026-04-02 · Xiang Yang, Feifei Li, Mi Zhang, Geng Hong 외 arxiv

Recent Text-to-Image (T2I) models based on rectified-flow transformers (e.g., SD3, FLUX) achieve high generative fidelity but remain vulnerable to unsafe semantics, especially when triggered by multi-token interactions. …

Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion Models

2025-10-28 · Byeonghu Na, Mina Kang, Jiseok Kwak, Minsang Park 외 arxiv

Text-to-image models have recently made significant advances in generating realistic and semantically coherent images, driven by advanced diffusion models and large-scale web-crawled datasets. However, these datasets oft…

CBF-RL: Safety Filtering Reinforcement Learning in Training with Control Barrier Functions

2025-10-16 · Lizhi Yang, Blake Werner, Massimiliano de Sa, Aaron D. Ames arxiv

Reinforcement learning (RL), while powerful and expressive, can often prioritize performance at the expense of safety. Yet safety violations can lead to catastrophic outcomes in real-world deployments. Control Barrier Fu…

Reinforcement Learning

A Hybrid Computational Intelligence Framework with Metaheuristic Optimization for Drug-Drug Interaction Prediction

2025-10-08 · Maryam Abdollahi Shamami, Babak Teimourpour, Farshad Sharifi arxiv

Drug-drug interactions (DDIs) are a leading cause of preventable adverse events, often complicating treatment and increasing healthcare costs. At the same time, knowing which drugs do not interact is equally important, a…