Structured Policy Representation: Imposing Stability in arbitrarily conditioned dynamic systems
We present a new family of deep neural network-based dynamic systems. The presented dynamics are globally stable and can be conditioned with an arbitrary context state. We show how these dynamics can be used as structured robot policies. Global stability is one of the most important and straightforward inductive biases as it allows us to impose reasonable behaviors outside the region of the demonstrations.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Latent Convergence Modulation in Large Language Models: A Novel Approach to Iterative Contextual Realignment
Token prediction stability remains a challenge in autoregressive generative models, where minor variations in early inference steps often lead to significant semantic drift over extended sequences. A structured modulatio…
Computational EfficiencySentenceText GenerationImposing Robust Structured Control Constraint on Reinforcement Learning of Linear Quadratic Regulator
This paper discusses learning a structured feedback control to obtain sufficient robustness to exogenous inputs for linear dynamic systems with unknown state matrix. The structural constraint on the controller is necessa…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Proactive Constrained Policy Optimization with Preemptive Penalty
Safe Reinforcement Learning (RL) often faces significant issues such as constraint violations and instability, necessitating the use of constrained policy optimization, which seeks optimal policies while ensuring adheren…
Reinforcement LearningDemystifying Action Space Design for Robotic Manipulation Policies
The specification of the action space plays a pivotal role in imitation-based robotic manipulation policy learning, fundamentally shaping the optimization landscape of policy learning. While recent advances have focused …
Stability-Constrained Markov Decision Processes Using MPC
In this paper, we consider solving discounted Markov Decision Processes (MDPs) under the constraint that the resulting policy is stabilizing. In practice MDPs are solved based on some form of policy approximation. We wil…
Model Predictive Control