paper-with-me

Papers

Towards Governing Agent's Efficacy: Action-Conditional $β$-VAE for Deep Transparent Reinforcement Learning

2018-11-11 · John Yang, Gyujeong Lee, Minsung Hyun, Simyung Chang, Nojun Kwak

We tackle the blackbox issue of deep neural networks in the settings of reinforcement learning (RL) where neural agents learn towards maximizing reward gains in an uncontrollable way. Such learning approach is risky when the interacting environment includes an expanse of state space because it is then almost impossible to foresee all unwanted outcomes and penalize them with negative rewards beforehand. Unlike reverse analysis of learned neural features from previous works, our proposed method \nj{tackles the blackbox issue by encouraging} an RL policy network to learn interpretable latent features through an implementation of a disentangled representation learning method. Toward this end, our method allows an RL agent to understand self-efficacy by distinguishing its influences from uncontrollable environmental factors, which closely resembles the way humans understand their scenes. Our experimental results show that the learned latent factors not only are interpretable, but also enable modeling the distribution of entire visited state space with a specific action condition. We have experimented that this characteristic of the proposed structure can lead to ex post facto governance for desired behaviors of RL agents.

📄 PDF Abstract BibTeX arXiv:1811.04350

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Representation Learning

Similar Papers 제목 키워드 기반

After the Party: Governing What a Viral Agent-Skill Ecosystem Left Behind

2026-09-15 · Yunpeng Xiong, Ting Zhang arxiv

AI agents increasingly act through agent skills, i.e., natural-language instructions, that direct a host agent toward shell, network, credential, file, and process actions, and public registries distribute them at scale.…

Governing Actions, Not Agents: Institutional Attestation as a Governance Model for Autonomous AI Systems

2026-06-24 · Jakob Salfeld-Nebgen arxiv

Autonomous AI agents may begin to perform consequential, irreversible actions such as clinical prescribing and production software deployment. This paper observes that human institutions have governed powerful autonomous…

Joint Action Language Modelling for Transparent Policy Execution

2025-04-14 · Theodor Wulff, Rahul Singh Maharjan, Xinyun Chi, Angelo Cangelosi

An agent's intention often remains hidden behind the black-box nature of embodied policies. Communication using natural language statements that describe the next action can provide transparency towards the agent's behav…

Language ModellingText Generation

Bridging Symbolic Control and Neural Reasoning in LLM Agents -- The Structured Cognitive Loop

2025-11-21 · Myung Ho Kim arxiv

Large language model agents suffer from architectural fragilities such as entangled reasoning and execution, memory volatility, and uncontrolled action sequences. We introduce Structured Cognitive Loop (SCL), a modular a…

Risk-sensitive Actor-free Policy via Convex Optimization

2023-06-30 · Ruoqi Zhang, Jens Sjölund

Traditional reinforcement learning methods optimize agents without considering safety, potentially resulting in unintended consequences. In this paper, we propose an optimal actor-free policy that optimizes a risk-sensit…

reinforcement-learningReinforcement Learning